{"doi":"10.1109/icassp.2004.1325927","title":"Text-independent speaker recognition by combining speaker-specific GMM with speaker adapted syllable-based HMM","abstract":null,"journal":"2004 IEEE International Conference on Acoustics, Speech, and Signal Processing","year":null,"id":646791,"datarank":0.5333022092234121,"base_score":3.5553480614894135,"endowment":3.5553480614894135,"self_citation_contribution":0.5333022092234121,"citation_network_contribution":0.0,"self_endowment_contribution":0.5333022092234121,"citer_contribution":0.0,"corpus_percentile":null,"corpus_rank":null,"citation_count":34,"citer_count":0,"citers_with_citation_signal":0,"citers_with_endowment":0,"datacite_reuse_total":0,"is_dataset":false,"is_dataset_confidence":null,"is_data_producer":false,"deposit_databanks":null,"is_oa":false,"file_count":0,"downloads":0,"has_version_chain":false,"published_date":null,"fair_score":null,"fair_percentile":null,"algorithm_id":"datarank_citation_only_1hop_v6","ranking_scope":"data_only","authors":[{"id":706308,"name":"W. Zhang","orcid":"0009-0003-7137-7294","position":1,"is_corresponding":false},{"id":47043,"name":"M. Takahashi","orcid":"0000-0003-1171-5960","position":2,"is_corresponding":false},{"id":549494,"name":"S. Nakagawa","orcid":null,"position":0,"is_corresponding":false}],"reference_count":0,"raw_metadata":{"has_enrichment":true,"resolved":true,"title":"Text-independent speaker recognition by combining speaker-specific GMM with speaker adapted syllable-based HMM","abstract":"We presented a new text-independent speaker recognition method by combining a speaker-specific Gaussian mixture model (GMM) with a syllable-based HMM adapted by MLLR or MAP (S. Nakagawa et al., Proc. Eurospeech, p.3017-3020, 2003). The robustness of this speaker recognition method for speaking style changes was evaluated in this paper. A speaker identification experiment, using an NTT database, which consists of sentences of data uttered at three speed modes (normal, fast and slow) by 35 Japanese speakers (22 males and 13 females) on five sessions over ten months, was conducted. Each speaker uttered only 5 training utterances (about 20 seconds in total). We obtained an accuracy of 98.8% for text-independent speaker identification for three speaking style modes (normal, fast, slow) by using a short test utterance (about 4 seconds). This result was superior to conventional methods for the same database. We show that the attractive result was brought from the compensational effect between speaker specific GMM and speaker adapted syllable based HMM.","is_dataset_classified":null,"base_score":3.5553480614894135,"endowment":3.5553480614894135,"datacite_reuse_total":0,"file_count":0,"downloads":0,"views":0,"has_version_chain":false,"is_dataset":false,"is_oa":false,"pmid":"19910364","pmcid":null,"openalex_id":"https://openalex.org/W2120097768","authors":[],"funders":[],"total_grants":0,"fwci":1.7641,"citation_percentile":0.85563943,"influential_citations":0,"citation_trend":[{"year":2012,"count":2},{"year":2013,"count":3},{"year":2014,"count":2},{"year":2015,"count":1},{"year":2016,"count":1},{"year":2018,"count":3},{"year":2021,"count":1},{"year":2023,"count":1}],"oa_status":"closed","license":null,"oa_locations":[{"url":"https://doi.org/10.1109/icassp.2004.1325927","host_type":""}],"fields_of_study":["Speech Recognition and Synthesis","Speech and Audio Processing"],"mesh_terms":[],"keywords":["Speech recognition","Speaker recognition","Speaker identification","Computer science","Hidden Markov model","Speaker diarisation","Robustness (evolution)","Utterance","Syllable","Mixture model","Artificial intelligence","Speaker verification","Pattern recognition (psychology)"],"sdg_mappings":[{"sdg_number":0,"sdg_label":"Quality Education"}],"linked_datasets":[],"clinical_trials":[],"software_tools":[],"database_accessions":[],"source":"live","citation_network_status":"fetched"},"created_at":"2026-08-09T15:02:50.126004Z","pmid":null,"pmcid":null,"fwci":null,"citation_percentile":null,"influential_citations":0,"oa_status":null,"license":null,"views":0,"total_file_size_bytes":0,"version_count":0,"fair_f":null,"fair_a":null,"fair_i":null,"fair_r":null,"fair_zscore":null,"fair_rationale":null,"fair_model":null,"fair_agent_version":null,"fair_fulltext_source":null,"fair_has_llm":null,"fair_computed_at":null,"clinical_trials":[],"software_tools":[],"db_accessions":[],"linked_datasets":[],"topics":[]}