{"doi":"10.1021/acs.jcim.5c01265","title":"Can Reasoning Power Significantly Improve the Knowledge of Large Language Models for Chemistry?─Based on Conversations with LLMs","abstract":null,"journal":"Journal of Chemical Information and Modeling","year":2025,"id":638944,"datarank":0.29188652235829704,"base_score":1.9459101490553132,"endowment":1.9459101490553132,"self_citation_contribution":0.29188652235829704,"citation_network_contribution":0.0,"self_endowment_contribution":0.29188652235829704,"citer_contribution":0.0,"corpus_percentile":null,"corpus_rank":null,"citation_count":6,"citer_count":0,"citers_with_citation_signal":0,"citers_with_endowment":0,"datacite_reuse_total":0,"is_dataset":false,"is_dataset_confidence":null,"is_data_producer":false,"deposit_databanks":null,"is_oa":false,"file_count":0,"downloads":0,"has_version_chain":false,"published_date":null,"fair_score":null,"fair_percentile":null,"algorithm_id":"datarank_citation_only_1hop_v6","ranking_scope":"data_only","authors":[{"id":1659819,"name":"Shi-Yu Long","orcid":"0009-0000-5365-3054","position":1,"is_corresponding":false},{"id":1659820,"name":"Yi-Xuan Tang","orcid":"0009-0007-4744-3967","position":2,"is_corresponding":false},{"id":607967,"name":"Yue Zhao","orcid":"0000-0003-0365-5291","position":3,"is_corresponding":false},{"id":803780,"name":"Qiao Li","orcid":"0000-0002-7472-5333","position":4,"is_corresponding":false},{"id":1659818,"name":"Dong-Xu Cui","orcid":"0009-0006-4866-0560","position":0,"is_corresponding":false}],"reference_count":0,"raw_metadata":{"has_enrichment":true,"resolved":true,"title":"Can Reasoning Power Significantly Improve the Knowledge of Large Language Models for Chemistry?─Based on Conversations with LLMs","abstract":"This study presents a systematic evaluation of five reasoning-enhanced Large Language Models (LLMs)─Deepseek-R1-0528, OpenAI-o4 mini, Gemini-2.5-pro, doubao-seed-1.6-thinking, and qwen-max-latest─across nine key chemistry tasks. By comparing these models with traditional LLMs and established computational tools, we systematically investigate the influence of reasoning capabilities and prompt engineering on chemical cognition. The results demonstrate that reasoning-enabled LLMs achieve significant performance improvements in fundamental tasks and that, in most cases, overly complex prompts are not beneficial for these models. However, domain-specific limitations persist; for instance, all five models exhibited structural inaccuracies in CIF file generation (such as incorrect bond topologies). Notably, while reasoning frameworks enhance logical coherence, they do not fundamentally resolve challenges in stereochemical identification or the recognition of rare symmetry groups. In essence, the spatial recognition capabilities of current Large Language Models remain insufficient. These findings underscore the necessity of developing domain-optimized training paradigms to bridge the gap between general reasoning capabilities and specialized chemical applications.","is_dataset_classified":null,"base_score":1.9459101490553132,"endowment":1.9459101490553132,"datacite_reuse_total":0,"file_count":0,"downloads":0,"views":0,"has_version_chain":false,"is_dataset":false,"is_oa":false,"pmid":"40854079","pmcid":null,"openalex_id":"https://openalex.org/W4413511670","authors":[],"funders":[{"funder_name":"Lanzhou University of Arts and Science","grant_id":"2020BSZX06","title":null}],"total_grants":1,"fwci":1.6986,"citation_percentile":0.84379024,"influential_citations":0,"citation_trend":[{"year":2025,"count":2},{"year":2026,"count":4}],"oa_status":"closed","license":"https://doi.org/10.15223/policy-029","oa_locations":[{"url":"https://pubs.acs.org/doi/pdf/10.1021/acs.jcim.5c01265","host_type":"publisher"},{"url":"https://doi.org/10.1021/acs.jcim.5c01265","host_type":"journal"},{"url":"https://pubmed.ncbi.nlm.nih.gov/40854079","host_type":"repository"}],"fields_of_study":["Machine Learning in Materials Science","Biomedical Text Mining and Ontologies","Topic Modeling","Cheminformatics","Language","Large Language Models"],"mesh_terms":["Cheminformatics","Large Language Models","Language"],"keywords":["Power (physics)","Computer science","Natural language processing","Chemistry","Physics"],"sdg_mappings":[{"sdg_number":0,"sdg_label":"Quality Education"}],"linked_datasets":[],"clinical_trials":[],"software_tools":[],"database_accessions":[],"source":"live","citation_network_status":"fetched"},"created_at":"2026-08-06T21:59:46.041036Z","pmid":null,"pmcid":null,"fwci":null,"citation_percentile":null,"influential_citations":0,"oa_status":null,"license":null,"views":0,"total_file_size_bytes":0,"version_count":0,"fair_f":null,"fair_a":null,"fair_i":null,"fair_r":null,"fair_zscore":null,"fair_rationale":null,"fair_model":null,"fair_agent_version":null,"fair_fulltext_source":null,"fair_has_llm":null,"fair_computed_at":null,"clinical_trials":[],"software_tools":[],"db_accessions":[],"linked_datasets":[],"topics":[]}