{"doi":"10.5281/zenodo.604494","title":"Data From: Automatic Definition Of Robust Microbiome Sub-States In Longitudinal Data","abstract":"Output files of the application of our R software (available at https://github.com/wilkinsonlab/robust-clustering-metagenomics) to different microbiome datasets already published. Prefixes: David2014_: original microbiome dataset published in [David et al.,2014] (http://genomebiology.com/2014/15/7/R89) Ballou2016_: original microbiome dataset published in [Ballou et al.,2016] (http://journal.frontiersin.org/article/10.3389/fvets.2016.00002/full) Gajer2012_: original microbiome dataset published in [Gajer et al.,2012] (http://stm.sciencemag.org/content/4/132/132ra52.long) LaRosa2014_: original microbiome dataset published in [LaRosa et al.,2014] (http://www.pnas.org/cgi/doi/10.1073/pnas.1409497111) Suffixes: _All: all taxa _Dominant: only 1% most abundant taxa _NonDominant: remaining taxa after removing above dominant taxa _GenusAll: taxa aggregated at genus level _GenusDominant: taxa aggregated at genes level and then to select only 1% most abundant taxa _GenusNonDominant: taxa aggregated at genus level and then to remove 1% most abundant taxa Each folder contains 3 output files related to the same input dataset:<br> - data.normAndDist_definitiveClustering_XXX.RData: R data file with a) a phyloseq object (including OTU table, meta-data and cluster assigned to each sample); and b) a distance matrix object.<br> - definitiveClusteringResults_XXX.txt: text file with assessment measures of the selected clustering.<br> - sampleId-cluster_pairs_XXX.txt: text file. Two columns, comma separated file: sampleID,clusterID Abstract of the associated paper: The analysis of microbiome dynamics would allow us to elucidate patterns within microbial community evolution; however, microbiome state-transition dynamics have been scarcely studied. This is in part because a necessary first-step in such analyses has not been well-defined: how to deterministically describe a microbiome's \"state\". Clustering in states have been widely studied, although no standard has been concluded yet. We propose a generic, domain-independent and automatic procedure to determine a reliable set of microbiome sub-states within a specific dataset, and with respect to the conditions of the study. The robustness of sub-state identification is established by the combination of diverse techniques for stable cluster verification. We reuse four distinct longitudinal microbiome datasets to demonstrate the broad applicability of our method, analysing results with different taxa subset allowing to adjust it depending on the application goal, and showing that the methodology provides a set of robust sub-states to examine in downstream studies about dynamics in microbiome.","journal":"Zenodo (CERN European Organization for Nuclear Research)","year":2018,"id":8852,"datarank":0.0,"base_score":0.0,"endowment":0.0,"self_citation_contribution":0.0,"citation_network_contribution":0.0,"self_endowment_contribution":0.0,"citer_contribution":0.0,"corpus_percentile":0.0,"corpus_rank":10062,"citation_count":0,"citer_count":0,"citers_with_citation_signal":0,"citers_with_endowment":0,"datacite_reuse_total":0,"is_dataset":true,"is_dataset_confidence":0.9509,"is_data_producer":false,"deposit_databanks":null,"is_oa":true,"file_count":2,"downloads":87,"has_version_chain":false,"published_date":"2018-02-27","fair_score":null,"fair_percentile":null,"algorithm_id":"datarank_citation_only_1hop_v6","ranking_scope":"data_only","authors":[{"id":606,"name":"Mark D D. Wilkinson","orcid":"0000-0001-6960-357X","position":1,"is_corresponding":false},{"id":1680,"name":"Beatriz García-Jiménez","orcid":"0000-0002-8129-6506","position":0,"is_corresponding":false}],"reference_count":0,"raw_metadata":{"citation_network_status":"fetched"},"created_at":"2026-03-01T18:20:47.508186Z","pmid":null,"pmcid":null,"fwci":null,"citation_percentile":null,"influential_citations":0,"oa_status":"green","license":"cc-by","views":0,"total_file_size_bytes":0,"version_count":0,"fair_f":null,"fair_a":null,"fair_i":null,"fair_r":null,"fair_zscore":null,"fair_rationale":null,"fair_model":null,"fair_agent_version":null,"fair_fulltext_source":null,"fair_has_llm":null,"fair_computed_at":null,"clinical_trials":[],"software_tools":[],"db_accessions":[],"linked_datasets":[],"topics":[]}