{"doi":"10.1093/bioinformatics/bth380","title":"The UniMarker (UM) method for synteny mapping of large genomes","abstract":"<jats:title>Abstract</jats:title>\n               <jats:p>Motivation: Synteny mapping, or detecting regions that are orthologous between two genomes, is a key step in studies of comparative genomics. For completely sequenced genomes, this is increasingly accomplished by whole-genome sequence alignment. However, such methods are computationally expensive, especially for large genomes, and require rather complicated post-processing procedures to filter out non-orthologous sequence matches.</jats:p>\n               <jats:p>Results: We have developed a novel method that does not require sequence alignment for synteny mapping of two large genomes, such as the human and mouse. In this method, the occurrence spectra of genome-wide unique 16mer sequences present in both the human and mouse genome are used to directly detect orthologous genomic segments. Being sequence alignment-free, the method is very fast and able to map the two mammalian genomes in one day of computing time on a single Pentium IV personal computer. The resulting human–mouse synteny map was shown to be in excellent agreement with those produced by the Mouse Genome Sequencing Consortium (MGSC) and by the Ensembl team; furthermore, the syntenic relationship of segments found only by our method was supported by BLASTZ sequence alignment.</jats:p>\n               <jats:p>Availability: The source code of our method and the resulting human–mouse synteny maps have been placed at http://synteny.ibms.sinica.edu.tw/ for free access.</jats:p>\n               <jats:p>Supplementary information: Seven supplementary figures can be found at the same website.</jats:p>","journal":"Bioinformatics","year":2004,"id":46868,"datarank":0.6734595534040597,"base_score":2.3978952727983707,"endowment":2.3978952727983707,"self_citation_contribution":0.3596842909197557,"citation_network_contribution":0.313775262484304,"self_endowment_contribution":0.3596842909197557,"citer_contribution":0.313775262484304,"corpus_percentile":null,"corpus_rank":null,"citation_count":10,"citer_count":9,"citers_with_citation_signal":7,"citers_with_endowment":7,"datacite_reuse_total":0,"is_dataset":false,"is_dataset_confidence":null,"is_data_producer":false,"deposit_databanks":null,"is_oa":false,"file_count":0,"downloads":0,"has_version_chain":false,"published_date":null,"fair_score":null,"fair_percentile":null,"algorithm_id":"datarank_citation_only_1hop_v6","ranking_scope":"data_only","authors":[{"id":216699,"name":"Yu-Jung Chang","orcid":null,"position":1,"is_corresponding":false},{"id":193300,"name":"Jan-Ming Ho","orcid":null,"position":2,"is_corresponding":false},{"id":216700,"name":"Ming-Jing Hwang","orcid":null,"position":3,"is_corresponding":false},{"id":112739,"name":"Ben-Yang Liao","orcid":null,"position":0,"is_corresponding":false}],"reference_count":0,"raw_metadata":{"has_enrichment":true,"resolved":true,"title":"The UniMarker (UM) method for synteny mapping of large genomes","abstract":"<jats:title>Abstract</jats:title>\n               <jats:p>Motivation: Synteny mapping, or detecting regions that are orthologous between two genomes, is a key step in studies of comparative genomics. For completely sequenced genomes, this is increasingly accomplished by whole-genome sequence alignment. However, such methods are computationally expensive, especially for large genomes, and require rather complicated post-processing procedures to filter out non-orthologous sequence matches.</jats:p>\n               <jats:p>Results: We have developed a novel method that does not require sequence alignment for synteny mapping of two large genomes, such as the human and mouse. In this method, the occurrence spectra of genome-wide unique 16mer sequences present in both the human and mouse genome are used to directly detect orthologous genomic segments. Being sequence alignment-free, the method is very fast and able to map the two mammalian genomes in one day of computing time on a single Pentium IV personal computer. The resulting human–mouse synteny map was shown to be in excellent agreement with those produced by the Mouse Genome Sequencing Consortium (MGSC) and by the Ensembl team; furthermore, the syntenic relationship of segments found only by our method was supported by BLASTZ sequence alignment.</jats:p>\n               <jats:p>Availability: The source code of our method and the resulting human–mouse synteny maps have been placed at http://synteny.ibms.sinica.edu.tw/ for free access.</jats:p>\n               <jats:p>Supplementary information: Seven supplementary figures can be found at the same website.</jats:p>","is_dataset_classified":null,"base_score":2.3978952727983707,"endowment":2.3978952727983707,"datacite_reuse_total":0,"file_count":0,"downloads":0,"views":0,"has_version_chain":false,"is_dataset":false,"is_oa":false,"pmid":"15217808","pmcid":null,"openalex_id":"https://openalex.org/W2117959411","authors":[],"funders":[],"total_grants":0,"fwci":0.479,"citation_percentile":0.62516087,"influential_citations":0,"citation_trend":[{"year":2012,"count":2},{"year":2013,"count":1},{"year":2014,"count":1},{"year":2016,"count":1}],"oa_status":"bronze","license":null,"oa_locations":[{"url":"https://academic.oup.com/bioinformatics/article-pdf/20/17/3156/48906377/bioinformatics_20_17_3156.pdf","host_type":"journal"},{"url":"https://academic.oup.com/bioinformatics/article-pdf/20/17/3156/48906377/bioinformatics_20_17_3156.pdf","host_type":"BRONZE"},{"url":"https://academic.oup.com/bioinformatics/article-pdf/20/17/3156/48906377/bioinformatics_20_17_3156.pdf","host_type":"publisher"},{"url":"https://doi.org/10.1093/bioinformatics/bth380","host_type":"journal"},{"url":"https://pubmed.ncbi.nlm.nih.gov/15217808","host_type":"repository"}],"fields_of_study":["Genomics and Phylogenetic Studies","RNA and protein synthesis mechanisms","Machine Learning in Bioinformatics","Computer Science","Medicine","Biology"],"mesh_terms":["Algorithms","Animals","Chromosome Mapping","Chromosomes, Human, Pair 16","Humans","Software","Species Specificity","User-Computer Interface","Genome, Human","Sequence Alignment","Conserved Sequence","Sequence Analysis, DNA","Evolution, Molecular","Genomic Islands","Mice"],"keywords":["Synteny","Ensembl","Genome","Comparative genomics","Computational biology","Sequence (biology)","Genomics","Biology","Human genome","Orthologous Gene","Reference genome","Genetics","Gene"],"sdg_mappings":[],"linked_datasets":[],"clinical_trials":[],"software_tools":[],"database_accessions":[],"source":"live","citation_network_status":"fetched"},"created_at":"2026-07-14T21:08:34.551663Z","pmid":null,"pmcid":null,"fwci":null,"citation_percentile":null,"influential_citations":0,"oa_status":null,"license":null,"views":0,"total_file_size_bytes":0,"version_count":0,"fair_f":null,"fair_a":null,"fair_i":null,"fair_r":null,"fair_zscore":null,"fair_rationale":null,"fair_model":null,"fair_agent_version":null,"fair_fulltext_source":null,"fair_has_llm":null,"fair_computed_at":null,"clinical_trials":[],"software_tools":[],"db_accessions":[],"linked_datasets":[],"topics":[]}