{"doi":"10.1016/j.xops.2025.100953","title":"A Datasheet for Age-Related Eye Disease Study 2 on the Database of Genotypes and Phenotypes","abstract":"Objective: To provide a comprehensive summary of the controlled-access Age-Related Eye Disease Study 2 (AREDS2) data elements, encompassing phenotypic, imaging, dietary, genetic, and ancillary data. Design: Dataset description of a multicenter, phase III, randomized clinical trial evaluating lutein + zeaxanthin, ω-3, or both long-chain polyunsaturated fatty acid supplementation in intermediate age-related macular degeneration (AMD). Secondary randomization was offered to all AREDS2 participants to evaluate varying levels of zinc and the potential for elimination of β-carotene, which increases the risk of lung cancer in smokers. Participants: A total of 4203 participants aged 50-85 years with bilateral intermediate AMD (bilateral large drusen ≥125 μm) or intermediate AMD in one eye and advanced AMD in the other eye were enrolled at 82 clinical centers between 2006 and 2008. Methods: Participants attended annual clinic visits, including eye examinations, visual acuity, slit lamp, intraocular pressure, and imaging that included stereoscopic 30° color fundus (fields 1-3) and fundus reflex images in all participants, while fundus autofluorescence images and spectral-domain OCT images were acquired in selected clinics. Telephone contacts at 3 and 6 months and annually thereafter collected adverse events and reinforced visit compliance. Main Outcome Measures: Progression to advanced AMD (central geographic atrophy or neovascular AMD), incidence of cataract surgery, and loss of ≥15 letters (≥3 lines) of visual acuity from baseline. Results: Controlled-access data are archived under the database of Genotypes and Phenotypes (dbGaP) website with the accession number phs002015.v2.p1. The data elements include main-study phenotype tables plus multiple ancillary-study tables with cardiovascular, cognitive, nutritional biochemistry, and genetic data. Additional data include dietary assessments, image gradings, visual acuity testing, and cataract surgery documentation. Blood or saliva from >2000 participants was collected; exome-chip data from >1800 and whole-genome sequencing from 1363 participants, including 488 who also participated in the original AREDS, are available under the International AMD Genomics Consortium and dbGaP. Conclusions: The AREDS2 dataset's rigorous interventional design, standardized longitudinal ophthalmic imaging gradings, comprehensive dietary and genetic information, and ancillary cardiovascular and cognitive assessments constitute an invaluable resource for elucidating AMD progression, informing nutritional strategies, and artificial intelligence-driven diagnostics. Financial Disclosures: Proprietary or commercial disclosure may be found in the Footnotes and Disclosures at the end of this article.","journal":"Ophthalmology Science","year":2025,"id":576841,"datarank":0.0,"base_score":0.0,"endowment":0.0,"self_citation_contribution":0.0,"citation_network_contribution":0.0,"self_endowment_contribution":0.0,"citer_contribution":0.0,"corpus_percentile":0.0,"corpus_rank":10062,"citation_count":0,"citer_count":0,"citers_with_citation_signal":0,"citers_with_endowment":0,"datacite_reuse_total":0,"is_dataset":true,"is_dataset_confidence":0.9522,"is_data_producer":false,"deposit_databanks":null,"is_oa":true,"file_count":0,"downloads":0,"has_version_chain":false,"published_date":"2025-01-01","fair_score":60.4167,"fair_percentile":80.03668602873739,"algorithm_id":"datarank_citation_only_1hop_v6","ranking_scope":"data_only","authors":[{"id":1125138,"name":"Akanksha Nagarkar","orcid":"0000-0001-9343-2547","position":1,"is_corresponding":false},{"id":1411878,"name":"Minali Prasad","orcid":"0000-0003-0596-1903","position":2,"is_corresponding":false},{"id":254973,"name":"Elvira Agrón","orcid":"0000-0002-2829-4042","position":3,"is_corresponding":false},{"id":867109,"name":"Claire Weber","orcid":"0000-0002-2352-8479","position":4,"is_corresponding":false},{"id":54567,"name":"Emily Y. Chew","orcid":"0000-0003-0999-9802","position":5,"is_corresponding":false},{"id":585826,"name":"Tharindu De Silva","orcid":"0000-0002-2882-4920","position":6,"is_corresponding":false},{"id":1258475,"name":"Souvick Mukherjee","orcid":"0000-0003-0748-8371","position":0,"is_corresponding":true}],"reference_count":23,"raw_metadata":null,"created_at":"2026-07-19T02:58:00.620755Z","pmid":"41450867","pmcid":"PMC12731274","fwci":null,"citation_percentile":null,"influential_citations":0,"oa_status":null,"license":null,"views":0,"total_file_size_bytes":0,"version_count":0,"fair_f":94.4444,"fair_a":62.5,"fair_i":100.0,"fair_r":20.8333,"fair_zscore":1.0278,"fair_rationale":{"fair_score":60.42,"has_llm":true,"taxonomy_version":"fair_taxonomy_v5","dimensions":{"F":{"name":"Findable","score":94.44,"criteria":[{"key":"f_dataset_pid","label":"Persistent identifier for the data","kind":"llm","weight":2.0,"fraction":1.0,"verdict":"yes","evidence":"The data described in this article are available in the Database of Genotypes and Phenotypes (dbGaP) at https://www.ncbi.nlm.nih.gov/projects/gap/cgi-bin/study.cgi?study_id=phs002015.v2.p1 under accession number phs002015.v2.p1.","grounded":true,"rationale":"The accession number phs002015.v2.p1 is a persistent identifier from the dbGaP scheme, matching the list of accepted PID schemes.","anchors":["RDA-F1-01D — FAIR Data Maturity Model: 'Data is identified by a persistent identifier' (priorit","RDA-F1-02D — FAIR Data Maturity Model: 'Data is identified by a globally unique identifier'","FsF-F1-02D — F-UJI/FAIRsFAIR: 'Data is assigned a persistent identifier'"],"scored":true,"signal":null},{"key":"f_repository_named","label":"Named repository","kind":"llm","weight":2.0,"fraction":1.0,"verdict":"yes","evidence":"The data described in this article are available in the Database of Genotypes and Phenotypes (dbGaP)","grounded":true,"rationale":"dbGaP is a named repository listed in re3data, satisfying the 'yes' class.","anchors":["RDA-F4-01M — FAIR Data Maturity Model: metadata is offered so it can be harvested and indexed (","NIH DMS Policy Element 4 (NOT-OD-21-014) — name the repository where data will be archived","NSTC Desirable Characteristics of Data Repositories (2022) — 'Long-Term Sustainability', 'Reten"],"scored":true,"signal":null},{"key":"f_data_availability_statement","label":"Data-availability statement","kind":"llm","weight":2.0,"fraction":1.0,"verdict":"yes","evidence":"The data described in this article are available in the Database of Genotypes and Phenotypes (dbGaP) at https://www.ncbi.nlm.nih.gov/projects/gap/cgi-bin/study.cgi?study_id=phs002015.v2.p1 under accession number phs002015.v2.p1.","grounded":true,"rationale":"The statement points to a repository record with a persistent identifier, placing it in Colavizza category 3.","anchors":["Colavizza, Hrynaszkiewicz, Staden, Whitaker & McGillivray (2020), 'The citation advantage of li","Springer Nature research data policy — Data Availability Statements: standard statement templat","RDA-F3-01M — metadata clearly and explicitly includes the identifier of the data it describes"],"scored":false,"signal":null},{"key":"f_discovery_metadata","label":"Description of the dataset as an object","kind":"llm","weight":2.0,"fraction":1.0,"verdict":"yes","evidence":"Table 2 Summary of Data Available in AREDS2 with dbGaP Tables","grounded":true,"rationale":"The paper provides an itemised inventory (Table 2) that names the data tables, variables, and descriptions, fulfilling the 'yes' class.","anchors":["RDA-F2-01M — 'Rich metadata is provided to allow discovery' (priority Essential)","FsF-F2-01M — F-UJI: 'Metadata includes descriptive core elements to support data findability'","FsF-R1-01MD — F-UJI: 'Metadata specifies the content of the data'"],"scored":false,"signal":null},{"key":"f_dataset_cited","label":"Dataset formally cited","kind":"llm","weight":1.0,"fraction":0.5,"verdict":"partial","evidence":"The dataset is available to investigators through dbGaP ( https://dbgap.ncbi.nlm.nih.gov/ ) under the accession number phs002015.v2.p1.","grounded":true,"rationale":"The dataset identifier appears only in the body text and not as a separate reference-list entry.","anchors":["FORCE11 Joint Declaration of Data Citation Principles (2014) — data should be cited as a first-","RDA-F3-01M — metadata clearly and explicitly includes the identifier of the data it describes","FsF-F3-01M — F-UJI: 'Metadata includes the identifier of the data it describes'"],"scored":true,"signal":null}]},"A":{"name":"Accessible","score":62.5,"criteria":[{"key":"a_data_openly_accessible","label":"Access route free of preconditions","kind":"llm","weight":2.0,"fraction":0.5,"verdict":"partial","evidence":"The AREDS2 dataset can be obtained by qualified researchers through dbGaP, the National Institutes of Health’s controlled access repository governed by the Genomic Data Sharing Policy.","grounded":true,"rationale":"The text describes a defined, followable access process with preconditions (application, IRB approval, Data Access Committee review), making it partial access, not open.","anchors":["RDA-A1.1-01D — 'Data is accessible through a free access protocol'","FsF-A1-01M — F-UJI: 'Metadata contains access level and access conditions of the data'","NSTC Desirable Characteristics of Data Repositories (2022) — 'Free and Easy Access'"],"scored":true,"signal":null},{"key":"a_access_conditions_stated","label":"Access level labelled","kind":"llm","weight":1.0,"fraction":1.0,"verdict":"yes","evidence":"Controlled-access data are archived under the database of Genotypes and Phenotypes (dbGaP) website with the accession number phs002015.v2.p1.","grounded":true,"rationale":"The paper explicitly labels the data as 'controlled-access' in the abstract and text, meeting the standard access-rights vocabulary. [majority verdict 'yes' (4/5 passes agreed)]","anchors":["FsF-A1-01M — F-UJI: 'Metadata contains access level and access conditions of the data'","RDA-A1-01M — metadata contains information to enable the user to get access to the data","COAR Controlled Vocabularies — Access Rights v1.0 (open / embargoed / restricted / metadata-onl"],"scored":false,"signal":null},{"key":"a_controlled_access_for_sensitive","label":"Gatekeeper for sensitive data","kind":"llm","weight":0.5,"fraction":1.0,"verdict":"yes","evidence":"A centralized Data Access Committee evaluates each request and issues approvals before data release.","grounded":true,"rationale":"The paper names an institutional gatekeeper (the Data Access Committee of dbGaP) for the sensitive human-subject data.","anchors":["NIH Genomic Data Sharing Policy (NOT-OD-14-124) — controlled-access via a Data Access Committee","RDA-A1.2-01D — 'Data is accessible through an access protocol that supports authentication and ","NIH DMS Policy Element 5 (NOT-OD-21-014) — Access, Distribution, or Reuse Considerations (conse"],"scored":false,"signal":null},{"key":"a_timeline_retention","label":"Availability timing & retention","kind":"llm","weight":0.5,"fraction":0.0,"verdict":"no","evidence":null,"grounded":false,"rationale":"The paper does not state a persistence commitment or availability timing for the data, so it is class 'no'. [majority verdict 'no' (4/5 passes agreed)]","anchors":["NIH DMS Plan Element 4 (NOT-OD-21-014) — Data Preservation, Access, and Associated Timelines","NSTC Desirable Characteristics (2022), Organizational Infrastructure: 'Retention Policy'","RDA-A2-01M — 'Metadata is guaranteed to remain available after data is no longer available'"],"scored":false,"signal":null}]},"I":{"name":"Interoperable","score":100.0,"criteria":[{"key":"i_open_nonproprietary_format","label":"Open file format","kind":"llm","weight":1.0,"fraction":1.0,"verdict":"yes","evidence":"These Digital Imaging and Communications in Medicine-embedded images cover 3 30-35° fields per eye","grounded":true,"rationale":"DICOM is an open, community-standard format recognised in the list of open formats. [majority verdict 'yes' (3/5 passes agreed)]","anchors":["FsF-R1.3-02D — F-UJI: 'Data is available in a file format recommended by the target research co","RDA-R1.3-02D — data is expressed in a machine-understandable community standard","RDA-I1-01D — data uses a knowledge representation expressed in a standardised format"],"scored":true,"signal":null},{"key":"i_community_standard_vocabulary","label":"Community standard / vocabulary","kind":"llm","weight":1.0,"fraction":1.0,"verdict":"yes","evidence":"MedDRA coding (preferred term, system organ class, and version)","grounded":true,"rationale":"MedDRA is a community-standard controlled vocabulary for adverse event reporting, registered in FAIRsharing.","anchors":["RDA-R1.3-01M — 'Metadata complies with a community standard' (priority Essential)","RDA-R1.3-01D — 'Data complies with a community standard'","RDA-I2-01M — '(Meta)data use vocabularies that follow FAIR principles'"],"scored":false,"signal":null},{"key":"i_qualified_references","label":"Identifiers for the resources the data depend on","kind":"llm","weight":0.5,"fraction":1.0,"verdict":"yes","evidence":"Genotype calls for more than 1800 of these individuals are available through the International AMD Genomics Consortium Exome Chip project (dbGaP accession phs001039)","grounded":true,"rationale":"The paper provides a dbGaP accession (phs001039) for a resource other than its own dataset, qualifying as a yes.","anchors":["RDA-I3-01M — '(meta)data include references to other (meta)data'","RDA-I3-03M — 'metadata includes qualified references to other metadata'","FsF-I3-01M — F-UJI: 'Metadata includes links between the data and its related entities'"],"scored":false,"signal":null}]},"R":{"name":"Reusable","score":20.83,"criteria":[{"key":"r_reuse_license","label":"Reuse licence","kind":"llm","weight":2.0,"fraction":0.0,"verdict":"no","evidence":null,"grounded":false,"rationale":"No licence or terms document is named for the data; the CC BY-NC-ND footer applies to the article, not the data.","anchors":["RDA-R1.1-01M — 'Metadata includes information about the licence under which the data can be reu","RDA-R1.1-02M — 'Metadata refers to a standard reuse licence'","RDA-R1.1-03M — 'Metadata refers to a machine-understandable reuse licence'"],"scored":true,"signal":null},{"key":"r_provenance_methods","label":"Provenance of the data","kind":"llm","weight":1.0,"fraction":0.0,"verdict":"no","evidence":"Genotype calls for more than 1800 of these individuals are available through the International AMD Genomics Consortium Exome Chip project... whole genome sequencing data for over 1300 participants have been deposited under the Consortium’s whole genome sequencing study","grounded":false,"rationale":"The methods are described in generic terms without naming specific instruments, kits, or software versions. [downgraded to 'no' — no verifiable quote from the paper] [majority verdict 'no' (2/5 passes agreed)]","anchors":["RDA-R1.2-01M — 'Metadata includes provenance information according to community- specific standa","FsF-R1.2-01M — F-UJI: 'Metadata includes provenance information about data creation or generati","W3C PROV-O (W3C Recommendation, 2013) — the entity/activity/agent model of provenance"],"scored":false,"signal":null},{"key":"r_documentation_codebook","label":"Documentation / codebook","kind":"llm","weight":1.0,"fraction":0.5,"verdict":"partial","evidence":"Table 2 summarizes the variables available with the dataset to facilitate lookup and retrieval with the data tables available to download via dbGaP.","grounded":true,"rationale":"The variable definitions are inside the article's Table 2, not in a separate documentation object shipped with the data, so it is partial. [majority verdict 'partial' (4/5 passes agreed)]","anchors":["RDA-R1-01M — '(Meta)data are richly described with a plurality of accurate and relevant attribu","FsF-R1-01MD — F-UJI: 'Metadata specifies the content of the data'","NIH DMS Policy Element 3 (NOT-OD-21-014) — Standards (documentation and metadata to accompany t"],"scored":false,"signal":null},{"key":"r_versioning","label":"Snapshot identified","kind":"llm","weight":0.5,"fraction":1.0,"verdict":"yes","evidence":"The dataset is available to investigators through dbGaP ( https://dbgap.ncbi.nlm.nih.gov/ ) under the accession number phs002015.v2.p1.","grounded":true,"rationale":"The accession includes a version token (v2.p1), providing a specific snapshot identifier.","anchors":["DataCite Metadata Schema 4.6 — the 'Version' property","RDA-R1.2-01M — provenance information (which version was used is provenance)","NSTC Desirable Characteristics of Data Repositories (2022) — 'Provenance', 'Retention Policy'"],"scored":true,"signal":null},{"key":"x_code_availability","label":"Analysis code available","kind":"llm","weight":1.0,"fraction":0.0,"verdict":"no","evidence":null,"grounded":false,"rationale":"No code locator of any kind is given for the study's own code; the paper does not address code availability.","anchors":["NIH DMS Policy Element 2 (NOT-OD-21-014) — 'Related Tools, Software and/or Code'","FAIR4RS Principles v1.0 (Chue Hong et al., 2022; RDA/FORCE11/ReSA) — FAIR Principles for Resear","FORCE11 Software Citation Principles (Smith, Katz & Niemeyer, 2016, PeerJ CS 2:e86)"],"scored":true,"signal":null},{"key":"x_funding_attribution","label":"Funder and award number","kind":"llm","weight":0.5,"fraction":0.5,"verdict":"partial","evidence":"This research is supported by the 10.13039/100030692 NIH Intramural Research Program , 10.13039/100000053 National Eye Institute / 10.13039/100000002 National Institutes of Health (NIH).","grounded":true,"rationale":"Funders are named (NIH, NEI) but no award or grant number is provided, so it is partial. [majority verdict 'partial' (4/5 passes agreed)]","anchors":["DataCite Metadata Schema 4.6 — 'FundingReference' property (funderName, funderIdentifier, award","Crossref Funder Registry — canonical funder identifiers for funding metadata","RDA-F2-01M — rich metadata provided to allow discovery (funding is part of the descriptive reco"],"scored":true,"signal":null}]}},"actions":[{"key":"r_reuse_license","dimension":"R","label":"Reuse licence","action":"Attach a standard, machine-readable open licence to the deposit — CC0 or CC BY, which is what Horizon Europe and most funders expect — and print the licence identifier in the paper. 'Free to use' is not a licence: it grants nothing a reuser's institution can rely on.","anchors":["yes","partial","no"],"verdict":"no","current":0.0,"evidence":null,"why":"No licence or terms document is named for the data; the CC BY-NC-ND footer applies to the article, not the data.","gain":16.67,"priority":"essential","scored":true},{"key":"a_data_openly_accessible","dimension":"A","label":"Access route free of preconditions","action":"Remove the precondition or justify it. Release the data at publication with no embargo, no registration wall, and no approval step — NIH's zero-embargo public- access rule (NOT-OD-25-101) has already made 'available at publication' the federal baseline for the article; the data should not lag behind it. For clinical / human-subjects data, deposit in dbGaP or the European Genome-phenome Archive (EGA).","anchors":["yes","partial","no"],"verdict":"partial","current":0.5,"evidence":"The AREDS2 dataset can be obtained by qualified researchers through dbGaP, the National Institutes of Health’s controlled access repository governed by the Genomic Data Sharing Policy.","why":"The text describes a defined, followable access process with preconditions (application, IRB approval, Data Access Committee review), making it partial access, not open.","gain":8.33,"priority":"essential","scored":true},{"key":"x_code_availability","dimension":"R","label":"Analysis code available","action":"Publish the analysis code in a public forge, archive a tagged release with a DOI (Zenodo/Software Heritage), and cite that DOI in the paper. NIH DMS Element 2 asks for the tools and code, not only the data — and 'available on request' is not a locator. Archive the analysis code in a versioned repository (GitHub + a Zenodo release DOI).","anchors":["yes","partial","no"],"verdict":"no","current":0.0,"evidence":null,"why":"No code locator of any kind is given for the study's own code; the paper does not address code availability.","gain":8.33,"priority":"important","scored":true},{"key":"f_dataset_cited","dimension":"F","label":"Dataset formally cited","action":"Cite the dataset in the reference list like a publication — creator, year, title, repository, DOI/accession — and cite it in-text where it is used. Only a reference- list entry is machine-readable to Crossref/DataCite, and only a citation lets the data earn credit. Cite the clinical / human-subjects repository accession (e.g. from dbGaP or the European Genome-phenome Archive (EGA)) in the reference list.","anchors":["yes","partial","no"],"verdict":"partial","current":0.5,"evidence":"The dataset is available to investigators through dbGaP ( https://dbgap.ncbi.nlm.nih.gov/ ) under the accession number phs002015.v2.p1.","why":"The dataset identifier appears only in the body text and not as a separate reference-list entry.","gain":4.17,"priority":"important","scored":true},{"key":"x_funding_attribution","dimension":"R","label":"Funder and award number","action":"State the funder AND the award number in the paper, and put them in the dataset's FundingReference metadata. A funder name alone cannot be linked back to the award, so the funding provenance of the data is lost the moment the paper is indexed.","anchors":["yes","partial","no"],"verdict":"partial","current":0.5,"evidence":"This research is supported by the 10.13039/100030692 NIH Intramural Research Program , 10.13039/100000053 National Eye Institute / 10.13039/100000002 National Institutes of Health (NIH).","why":"Funders are named (NIH, NEI) but no award or grant number is provided, so it is partial. [majority verdict 'partial' (4/5 passes agreed)]","gain":2.08,"priority":"useful","scored":true},{"key":"r_provenance_methods","dimension":"R","label":"Provenance of the data","action":"Name the instruments, kits, and software — with versions — that produced the data, not just the verbs. 'Reads were aligned' is not provenance; 'aligned with STAR v2.7.9a to GRCh38' is, because someone else can rerun it.","anchors":["yes","partial","no"],"verdict":"no","current":0.0,"evidence":"Genotype calls for more than 1800 of these individuals are available through the International AMD Genomics Consortium Exome Chip project... whole genome sequencing data for over 1300 participants have been deposited under the Consortium’s whole genome sequencing study","why":"The methods are described in generic terms without naming specific instruments, kits, or software versions. [downgraded to 'no' — no verifiable quote from the paper] [majority verdict 'no' (2/5 passes agreed)]","gain":0.0,"priority":"important","scored":false},{"key":"r_documentation_codebook","dimension":"R","label":"Documentation / codebook","action":"Ship a README and a data dictionary IN the deposit — every file, every variable, its units, its allowed values, its missing-value codes. It is the cheapest single thing that makes a dataset usable by someone who was not in the lab, and a table buried in the article does not travel with the data.","anchors":["yes","partial","no"],"verdict":"partial","current":0.5,"evidence":"Table 2 summarizes the variables available with the dataset to facilitate lookup and retrieval with the data tables available to download via dbGaP.","why":"The variable definitions are inside the article's Table 2, not in a separate documentation object shipped with the data, so it is partial. [majority verdict 'partial' (4/5 passes agreed)]","gain":0.0,"priority":"important","scored":false},{"key":"a_timeline_retention","dimension":"A","label":"Availability timing & retention","action":"State when the data become available AND how long they will be retained — cite the repository's preservation policy. NIH DMS Element 4 asks for both; most papers give neither.","anchors":["yes","partial","no"],"verdict":"no","current":0.0,"evidence":null,"why":"The paper does not state a persistence commitment or availability timing for the data, so it is class 'no'. [majority verdict 'no' (4/5 passes agreed)]","gain":0.0,"priority":"useful","scored":false}],"suggestions":["Attach a standard, machine-readable open licence to the deposit — CC0 or CC BY, which is what Horizon Europe and most funders expect — and print the licence identifier in the paper. 'Free to use' is not a licence: it grants nothing a reuser's institution can rely on.","Remove the precondition or justify it. Release the data at publication with no embargo, no registration wall, and no approval step — NIH's zero-embargo public- access rule (NOT-OD-25-101) has already made 'available at publication' the federal baseline for the article; the data should not lag behind it. For clinical / human-subjects data, deposit in dbGaP or the European Genome-phenome Archive (EGA).","Publish the analysis code in a public forge, archive a tagged release with a DOI (Zenodo/Software Heritage), and cite that DOI in the paper. NIH DMS Element 2 asks for the tools and code, not only the data — and 'available on request' is not a locator. Archive the analysis code in a versioned repository (GitHub + a Zenodo release DOI).","Cite the dataset in the reference list like a publication — creator, year, title, repository, DOI/accession — and cite it in-text where it is used. Only a reference- list entry is machine-readable to Crossref/DataCite, and only a citation lets the data earn credit. Cite the clinical / human-subjects repository accession (e.g. from dbGaP or the European Genome-phenome Archive (EGA)) in the reference list.","State the funder AND the award number in the paper, and put them in the dataset's FundingReference metadata. A funder name alone cannot be linked back to the award, so the funding provenance of the data is lost the moment the paper is indexed."],"model":"deepseek/deepseek-v4-flash","agent_version":"fair_agent_v8","fulltext_source":"epmc_xml"},"fair_model":"deepseek/deepseek-v4-flash","fair_agent_version":"fair_agent_v8","fair_fulltext_source":"epmc_xml","fair_has_llm":true,"fair_computed_at":"2026-07-20T14:04:25.377276Z","clinical_trials":[],"software_tools":[],"db_accessions":[],"linked_datasets":[],"topics":[]}