{"doi":"10.1002/gepi.21992","title":"Toward the integration of <i>Omics</i> data in epidemiological studies: still a “long and winding road”","abstract":"<jats:title>ABSTRACT</jats:title><jats:p>Primary and secondary prevention can highly benefit a personalized medicine approach through the accurate discrimination of individuals at high risk of developing a specific disease from those at moderate and low risk. To this end precise risk prediction models need to be built. This endeavor requires a precise characterization of the individual exposome, genome, and phenome. Massive molecular <jats:italic>omics</jats:italic> data representing the different layers of the biological processes of the host and the nonhost will enable to build more accurate risk prediction models. Epidemiologists aim to integrate <jats:italic>omics</jats:italic> data along with important information coming from other sources (questionnaires, candidate markers) that has been proved to be relevant in the discrimination risk assessment of complex diseases. However, the integrative models in large‐scale epidemiologic research are still in their infancy and they face numerous challenges, some of them at the analytical stage. So far, there are a small number of studies that have integrated more than two <jats:italic>omics</jats:italic> data sets, and the inclusion of non‐<jats:italic>omics</jats:italic> data in the same models is still missing in most of studies. In this contribution, we aim at approaching the <jats:italic>omics</jats:italic> and non‐<jats:italic>omics</jats:italic> data integration from the epidemiology scope by considering the “massive” inclusion of variables in the risk assessment and predictive models. We also provide already available examples of integrative contributions in the field, propose analytical strategies that allow considering both <jats:italic>omics</jats:italic> and non‐<jats:italic>omics</jats:italic> data in the models, and finally review the challenges imbedding this type of research.</jats:p>","journal":"Genetic Epidemiology","year":2016,"id":645550,"datarank":0.47670807455219194,"base_score":3.1780538303479458,"endowment":3.1780538303479458,"self_citation_contribution":0.47670807455219194,"citation_network_contribution":0.0,"self_endowment_contribution":0.47670807455219194,"citer_contribution":0.0,"corpus_percentile":null,"corpus_rank":null,"citation_count":23,"citer_count":0,"citers_with_citation_signal":0,"citers_with_endowment":0,"datacite_reuse_total":0,"is_dataset":false,"is_dataset_confidence":null,"is_data_producer":false,"deposit_databanks":null,"is_oa":false,"file_count":0,"downloads":0,"has_version_chain":false,"published_date":null,"fair_score":null,"fair_percentile":null,"algorithm_id":"datarank_citation_only_1hop_v6","ranking_scope":"data_only","authors":[{"id":338892,"name":"Sílvia Pineda","orcid":"0000-0003-4017-1480","position":1,"is_corresponding":false},{"id":1680996,"name":"Angela Brand","orcid":null,"position":2,"is_corresponding":false},{"id":112165,"name":"Kristel Van Steen","orcid":null,"position":3,"is_corresponding":false},{"id":416132,"name":"Núria Malats","orcid":"0000-0003-2538-3784","position":4,"is_corresponding":false},{"id":564266,"name":"Evangelina López de Maturana","orcid":"0000-0001-9425-3911","position":0,"is_corresponding":false}],"reference_count":0,"raw_metadata":{"has_enrichment":true,"resolved":true,"title":"Toward the integration of <i>Omics</i> data in epidemiological studies: still a “long and winding road”","abstract":"<jats:title>ABSTRACT</jats:title><jats:p>Primary and secondary prevention can highly benefit a personalized medicine approach through the accurate discrimination of individuals at high risk of developing a specific disease from those at moderate and low risk. To this end precise risk prediction models need to be built. This endeavor requires a precise characterization of the individual exposome, genome, and phenome. Massive molecular <jats:italic>omics</jats:italic> data representing the different layers of the biological processes of the host and the nonhost will enable to build more accurate risk prediction models. Epidemiologists aim to integrate <jats:italic>omics</jats:italic> data along with important information coming from other sources (questionnaires, candidate markers) that has been proved to be relevant in the discrimination risk assessment of complex diseases. However, the integrative models in large‐scale epidemiologic research are still in their infancy and they face numerous challenges, some of them at the analytical stage. So far, there are a small number of studies that have integrated more than two <jats:italic>omics</jats:italic> data sets, and the inclusion of non‐<jats:italic>omics</jats:italic> data in the same models is still missing in most of studies. In this contribution, we aim at approaching the <jats:italic>omics</jats:italic> and non‐<jats:italic>omics</jats:italic> data integration from the epidemiology scope by considering the “massive” inclusion of variables in the risk assessment and predictive models. We also provide already available examples of integrative contributions in the field, propose analytical strategies that allow considering both <jats:italic>omics</jats:italic> and non‐<jats:italic>omics</jats:italic> data in the models, and finally review the challenges imbedding this type of research.</jats:p>","is_dataset_classified":null,"base_score":3.1780538303479458,"endowment":3.1780538303479458,"datacite_reuse_total":0,"file_count":0,"downloads":0,"views":0,"has_version_chain":false,"is_dataset":false,"is_oa":false,"pmid":"27432111","pmcid":null,"openalex_id":"https://openalex.org/W2505609335","authors":[],"funders":[{"funder_name":"Instituto de Salud Carlos III","grant_id":"#PI12‐00815","title":null},{"funder_name":"European Cooperation in Science and Technology","grant_id":"#BM1204: EU_Pancreas","title":null}],"total_grants":2,"fwci":2.3573,"citation_percentile":0.86715546,"influential_citations":0,"citation_trend":[{"year":2017,"count":4},{"year":2018,"count":4},{"year":2019,"count":4},{"year":2020,"count":3},{"year":2021,"count":1},{"year":2022,"count":3},{"year":2023,"count":3},{"year":2024,"count":1}],"oa_status":"closed","license":"http://onlinelibrary.wiley.com/termsAndConditions#vor","oa_locations":[{"url":"https://api.wiley.com/onlinelibrary/tdm/v1/articles/10.1002%2Fgepi.21992","host_type":"publisher"},{"url":"https://onlinelibrary.wiley.com/doi/pdf/10.1002/gepi.21992","host_type":"publisher"},{"url":"https://doi.org/10.1002/gepi.21992","host_type":"journal"},{"url":"https://pubmed.ncbi.nlm.nih.gov/27432111","host_type":"repository"},{"url":"https://cris.maastrichtuniversity.nl/en/publications/c65405d4-a286-4252-841e-4c94d263a10d","host_type":"repository"},{"url":"https://lirias.kuleuven.be/bitstream/123456789/555092/3/gepi21992.pdf","host_type":"repository"},{"url":"https://orbi.uliege.be/handle/2268/233652","host_type":"repository"}],"fields_of_study":["Health, Environment, Cognitive Aging","Nutrition, Genetics, and Disease","Nutritional Studies and Diet","Epidemiologic Studies","Genetic Predisposition to Disease","Genome-Wide Association Study","Genomics","Humans","Models, Genetic","Polymorphism, Single Nucleotide"],"mesh_terms":["Humans","Models, Genetic","Epidemiologic Studies","Genetic Predisposition to Disease","Polymorphism, Single Nucleotide","Genomics","Genome-Wide Association Study"],"keywords":["Omics","Phenome","Data science","Computer science","Exposome","Data integration","Computational biology","Bioinformatics","Data mining","Biology","Medicine","Genome","Pathology","Integration","Exposure","Statistical methods","Genetic susceptibility","Challenges","epidemiology","Outcome","Omics Data"],"sdg_mappings":[{"sdg_number":0,"sdg_label":"Reduced inequalities"}],"linked_datasets":[],"clinical_trials":[],"software_tools":[],"database_accessions":[],"source":"live","citation_network_status":"fetched"},"created_at":"2026-08-09T07:56:22.269144Z","pmid":null,"pmcid":null,"fwci":null,"citation_percentile":null,"influential_citations":0,"oa_status":null,"license":null,"views":0,"total_file_size_bytes":0,"version_count":0,"fair_f":null,"fair_a":null,"fair_i":null,"fair_r":null,"fair_zscore":null,"fair_rationale":null,"fair_model":null,"fair_agent_version":null,"fair_fulltext_source":null,"fair_has_llm":null,"fair_computed_at":null,"clinical_trials":[],"software_tools":[],"db_accessions":[],"linked_datasets":[],"topics":[]}