{"doi":"10.1002/9780470015902.a0020868","title":"High‐Throughput Automated Subcellular Localisation","abstract":"<jats:title>Abstract</jats:title>\n          <jats:sec>\n            <jats:label/>\n            <jats:p>Defining the subcellular localisation of the proteome for an organism of interest is a critical next step following genome sequencing. Knowledge of protein subcellular localisation provides insight into the functionality of the normal cell, as well during disease states. However, the presence of gene isoforms, alternative splicing and posttranslational modifications significantly increase the number of protein variants encoded by a single gene, making this a complex task. In the last 20 years, parallel approaches using fractionation and mass spectrometry, synthesis of large libraries of open reading frames fused to genes encoding fluorescent proteins, as well as production of thousands of antibodies have all contributed to the systematic analysis of protein localisation. Alongside these methods, improved bioinformatic predictors, machine learning and deep learning algorithms have also evolved as essential tools. A combinatorial approach of these methods now brings us close to systematically defining the subcellular proteome for many organisms.</jats:p>\n          </jats:sec>\n          <jats:sec>\n            <jats:title>Key Concepts</jats:title>\n            <jats:p>\n              <jats:list list-type=\"bullet\">\n                <jats:list-item>\n                  <jats:p>Subcellular localisation is a critical determinant in understanding protein function.</jats:p>\n                </jats:list-item>\n                <jats:list-item>\n                  <jats:p>Data from genome sequencing projects provide the fundamental information from which approaches to understand protein localisation can be initiated.</jats:p>\n                </jats:list-item>\n                <jats:list-item>\n                  <jats:p>Parallel approaches using fluorescence microscopy are being applied in a high‐throughput manner to systematically reveal the subcellular localisation of large numbers of proteins in different cells.</jats:p>\n                </jats:list-item>\n                <jats:list-item>\n                  <jats:p>The primary techniques to determine protein localisation are mass spectrometry‐based proteomics, production of antibodies and expression of fluorescently tagged proteins.</jats:p>\n                </jats:list-item>\n                <jats:list-item>\n                  <jats:p>There is increasing use of computational biology tools to aid the automated classification of subcellular localisation.</jats:p>\n                </jats:list-item>\n                <jats:list-item>\n                  <jats:p>Large image datasets can be interrogated by machine learning software algorithms to automatically classify proteins to specific localisations.</jats:p>\n                </jats:list-item>\n                <jats:list-item>\n                  <jats:p>Deep learning methods, which can work independently of training datasets, have become the newest tool to automatically assign protein localisation from image sets.</jats:p>\n                </jats:list-item>\n                <jats:list-item>\n                  <jats:p>Automated approaches combining both experimental and computational methods are likely to become the primary means by which subcellular localisation is determined from new cell systems.</jats:p>\n                </jats:list-item>\n              </jats:list>\n            </jats:p>\n          </jats:sec>","journal":"Encyclopedia of Life Sciences","year":2020,"id":44407,"datarank":0.0,"base_score":0.0,"endowment":0.0,"self_citation_contribution":0.0,"citation_network_contribution":0.0,"self_endowment_contribution":0.0,"citer_contribution":0.0,"corpus_percentile":null,"corpus_rank":null,"citation_count":0,"citer_count":0,"citers_with_citation_signal":0,"citers_with_endowment":0,"datacite_reuse_total":0,"is_dataset":false,"is_dataset_confidence":null,"is_data_producer":false,"deposit_databanks":null,"is_oa":false,"file_count":0,"downloads":0,"has_version_chain":false,"published_date":null,"fair_score":null,"fair_percentile":null,"algorithm_id":"datarank_citation_only_1hop_v6","ranking_scope":"data_only","authors":[{"id":209095,"name":"Suainibhe Kelly","orcid":null,"position":1,"is_corresponding":false},{"id":209096,"name":"Margaritha M Mysior","orcid":null,"position":2,"is_corresponding":false},{"id":209097,"name":"Jeremy C Simpson","orcid":null,"position":3,"is_corresponding":false},{"id":209094,"name":"Alannah S Chalkley","orcid":null,"position":0,"is_corresponding":false}],"reference_count":0,"raw_metadata":{"has_enrichment":true,"base_score":0.0,"endowment":0.0,"datacite_reuse_total":0,"file_count":0,"downloads":0,"views":0,"has_version_chain":false,"is_dataset":false,"is_oa":false,"pmid":"21062823","pmcid":null,"openalex_id":"https://openalex.org/W3081696891","authors":[],"funders":[],"total_grants":0,"fwci":null,"citation_percentile":null,"influential_citations":0,"citation_trend":[],"oa_status":"closed","license":"http://doi.wiley.com/10.1002/tdm_license_1.1","oa_locations":[{"url":"https://onlinelibrary.wiley.com/doi/pdf/10.1002/9780470015902.a0020868","host_type":"publisher"},{"url":"https://onlinelibrary.wiley.com/doi/full-xml/10.1002/9780470015902.a0020868","host_type":"publisher"},{"url":"https://doi.org/10.1002/9780470015902.a0020868","host_type":"journal"}],"fields_of_study":["Cell Image Analysis Techniques","Image Processing Techniques and Applications","Gene expression and cancer classification","Computer Science","Biology","Materials Science"],"mesh_terms":[],"keywords":["Subcellular localization","Proteome","Computational biology","Proteomics","Protein subcellular localization prediction","Biology","Genome","Human proteome project","Alternative splicing","Protein isoform","Identification (biology)","Gene isoform","Gene","Computer science","Bioinformatics","Biochemistry"],"sdg_mappings":[],"linked_datasets":[],"clinical_trials":[],"software_tools":[],"database_accessions":[],"source":"live","citation_network_status":"fetched"},"created_at":"2026-06-29T00:08:14.794831Z","pmid":null,"pmcid":null,"fwci":null,"citation_percentile":null,"influential_citations":0,"oa_status":null,"license":null,"views":0,"total_file_size_bytes":0,"version_count":0,"fair_f":null,"fair_a":null,"fair_i":null,"fair_r":null,"fair_zscore":null,"fair_rationale":null,"fair_model":null,"fair_agent_version":null,"fair_fulltext_source":null,"fair_has_llm":null,"fair_computed_at":null,"clinical_trials":[],"software_tools":[],"db_accessions":[],"linked_datasets":[],"topics":[]}