{"record_type":"pith_number_record","schema_url":"https://pith.science/schemas/pith-number/v1.json","pith_number":"pith:2023:FDYCWLGEK72SHF7VOMNOJRLDXT","short_pith_number":"pith:FDYCWLGE","schema_version":"1.0","canonical_sha256":"28f02b2cc457f52397f5731ae4c563bcc1775ffd3a95abc7cf9dfa359caf9516","source":{"kind":"arxiv","id":"2305.18365","version":3},"attestation_state":"computed","paper":{"title":"What can Large Language Models do in chemistry? A comprehensive benchmark on eight tasks","license":"http://arxiv.org/licenses/nonexclusive-distrib/1.0/","headline":"","cross_cats":["cs.AI"],"primary_cat":"cs.CL","authors_text":"Bozhao Nan, Kehan Guo, Nitesh V. Chawla, Olaf Wiest, Taicheng Guo, Xiangliang Zhang, Zhenwen Liang, Zhichun Guo","submitted_at":"2023-05-27T14:17:33Z","abstract_excerpt":"Large Language Models (LLMs) with strong abilities in natural language processing tasks have emerged and have been applied in various kinds of areas such as science, finance and software engineering. However, the capability of LLMs to advance the field of chemistry remains unclear. In this paper, rather than pursuing state-of-the-art performance, we aim to evaluate capabilities of LLMs in a wide range of tasks across the chemistry domain. We identify three key chemistry-related capabilities including understanding, reasoning and explaining to explore in LLMs and establish a benchmark containin"},"verification_status":{"content_addressed":true,"pith_receipt":true,"author_attested":false,"weak_author_claims":0,"strong_author_claims":0,"externally_anchored":false,"storage_verified":false,"citation_signatures":0,"replication_records":0,"graph_snapshot":true,"references_resolved":false,"formal_links_present":false},"canonical_record":{"source":{"id":"2305.18365","kind":"arxiv","version":3},"metadata":{"license":"http://arxiv.org/licenses/nonexclusive-distrib/1.0/","primary_cat":"cs.CL","submitted_at":"2023-05-27T14:17:33Z","cross_cats_sorted":["cs.AI"],"title_canon_sha256":"229d1b1325bc40663bc6173851bebb667a4a3e13871d409add2fc542ca75de13","abstract_canon_sha256":"1d0620d2ca042183bb9ee83f70a65413670835ea5c8dea9f85db5431aec4b70e"},"schema_version":"1.0"},"receipt":{"kind":"pith_receipt","key_id":"pith-v1-2026-05","algorithm":"ed25519","signed_at":"2026-07-05T07:28:28.324790Z","signature_b64":"xvx2Z/LU8+SxgIANNwG12cgJ5tMNWU9FCBhBgjTMGXDnxZ6eFmzuvI0QQ0buzub3NfAS7oVcVVmYEQn6LngJCg==","signed_message":"canonical_sha256_bytes","builder_version":"pith-number-builder-2026-05-17-v1","receipt_version":"0.3","canonical_sha256":"28f02b2cc457f52397f5731ae4c563bcc1775ffd3a95abc7cf9dfa359caf9516","last_reissued_at":"2026-07-05T07:28:28.324234Z","signature_status":"signed_v1","first_computed_at":"2026-07-05T07:28:28.324234Z","public_key_fingerprint":"8d4b5ee74e4693bcd1df2446408b0d54"},"graph_snapshot":{"paper":{"title":"What can Large Language Models do in chemistry? A comprehensive benchmark on eight tasks","license":"http://arxiv.org/licenses/nonexclusive-distrib/1.0/","headline":"","cross_cats":["cs.AI"],"primary_cat":"cs.CL","authors_text":"Bozhao Nan, Kehan Guo, Nitesh V. Chawla, Olaf Wiest, Taicheng Guo, Xiangliang Zhang, Zhenwen Liang, Zhichun Guo","submitted_at":"2023-05-27T14:17:33Z","abstract_excerpt":"Large Language Models (LLMs) with strong abilities in natural language processing tasks have emerged and have been applied in various kinds of areas such as science, finance and software engineering. However, the capability of LLMs to advance the field of chemistry remains unclear. In this paper, rather than pursuing state-of-the-art performance, we aim to evaluate capabilities of LLMs in a wide range of tasks across the chemistry domain. We identify three key chemistry-related capabilities including understanding, reasoning and explaining to explore in LLMs and establish a benchmark containin"},"claims":{"count":0,"items":[],"snapshot_sha256":"258153158e38e3291e3d48162225fcdb2d5a3ed65a07baac614ab91432fd4f57"},"source":{"id":"2305.18365","kind":"arxiv","version":3},"verdict":{"id":null,"model_set":{},"created_at":null,"strongest_claim":"","one_line_summary":"","pipeline_version":null,"weakest_assumption":"","pith_extraction_headline":""},"integrity":{"clean":true,"summary":{"advisory":0,"critical":0,"by_detector":{},"informational":0},"endpoint":"/pith/2305.18365/integrity.json","findings":[],"available":true,"detectors_run":[],"snapshot_sha256":"c28c3603d3b5d939e8dc4c7e95fa8dfce3d595e45f758748cecf8e644a296938"},"references":{"count":0,"sample":[],"resolved_work":0,"snapshot_sha256":"258153158e38e3291e3d48162225fcdb2d5a3ed65a07baac614ab91432fd4f57","internal_anchors":0},"formal_canon":{"evidence_count":0,"snapshot_sha256":"258153158e38e3291e3d48162225fcdb2d5a3ed65a07baac614ab91432fd4f57"},"author_claims":{"count":0,"strong_count":0,"snapshot_sha256":"258153158e38e3291e3d48162225fcdb2d5a3ed65a07baac614ab91432fd4f57"},"builder_version":"pith-number-builder-2026-05-17-v1"},"aliases":[{"alias_kind":"arxiv","alias_value":"2305.18365","created_at":"2026-07-05T07:28:28.324304+00:00"},{"alias_kind":"arxiv_version","alias_value":"2305.18365v3","created_at":"2026-07-05T07:28:28.324304+00:00"},{"alias_kind":"doi","alias_value":"10.48550/arxiv.2305.18365","created_at":"2026-07-05T07:28:28.324304+00:00"},{"alias_kind":"pith_short_12","alias_value":"FDYCWLGEK72S","created_at":"2026-07-05T07:28:28.324304+00:00"},{"alias_kind":"pith_short_16","alias_value":"FDYCWLGEK72SHF7V","created_at":"2026-07-05T07:28:28.324304+00:00"},{"alias_kind":"pith_short_8","alias_value":"FDYCWLGE","created_at":"2026-07-05T07:28:28.324304+00:00"}],"events":[],"event_summary":{},"paper_claims":[],"inbound_citations":{"count":6,"internal_anchor_count":0,"sample":[{"citing_arxiv_id":"2507.21035","citing_title":"GenoMAS: A Multi-Agent Framework for Scientific Discovery via Code-Driven Gene Expression Analysis","ref_index":38,"is_internal_anchor":false},{"citing_arxiv_id":"2604.26498","citing_title":"Do Larger Models Really Win in Drug Discovery? A Benchmark Assessment of Model Scaling in AI-Driven Molecular Property and Activity Prediction","ref_index":30,"is_internal_anchor":false},{"citing_arxiv_id":"2605.12784","citing_title":"ToolMol: Evolutionary Agentic Framework for Multi-objective Drug Discovery","ref_index":15,"is_internal_anchor":false},{"citing_arxiv_id":"2605.12784","citing_title":"ToolMol: Evolutionary Agentic Framework for Multi-objective Drug Discovery","ref_index":15,"is_internal_anchor":false},{"citing_arxiv_id":"2604.26498","citing_title":"Do Larger Models Really Win in Drug Discovery? A Benchmark Assessment of Model Scaling in AI-Driven Molecular Property and Activity Prediction","ref_index":30,"is_internal_anchor":false},{"citing_arxiv_id":"2402.01680","citing_title":"Large Language Model based Multi-Agents: A Survey of Progress and Challenges","ref_index":21,"is_internal_anchor":false}]},"formal_canon":{"evidence_count":0,"sample":[],"anchors":[]},"links":{"html":"https://pith.science/pith/FDYCWLGEK72SHF7VOMNOJRLDXT","json":"https://pith.science/pith/FDYCWLGEK72SHF7VOMNOJRLDXT.json","graph_json":"https://pith.science/api/pith-number/FDYCWLGEK72SHF7VOMNOJRLDXT/graph.json","events_json":"https://pith.science/api/pith-number/FDYCWLGEK72SHF7VOMNOJRLDXT/events.json","paper":"https://pith.science/paper/FDYCWLGE"},"agent_actions":{"view_html":"https://pith.science/pith/FDYCWLGEK72SHF7VOMNOJRLDXT","download_json":"https://pith.science/pith/FDYCWLGEK72SHF7VOMNOJRLDXT.json","view_paper":"https://pith.science/paper/FDYCWLGE","resolve_alias":"https://pith.science/api/pith-number/resolve?arxiv=2305.18365&json=true","fetch_graph":"https://pith.science/api/pith-number/FDYCWLGEK72SHF7VOMNOJRLDXT/graph.json","fetch_events":"https://pith.science/api/pith-number/FDYCWLGEK72SHF7VOMNOJRLDXT/events.json","actions":{"anchor_timestamp":"https://pith.science/pith/FDYCWLGEK72SHF7VOMNOJRLDXT/action/timestamp_anchor","attest_storage":"https://pith.science/pith/FDYCWLGEK72SHF7VOMNOJRLDXT/action/storage_attestation","attest_author":"https://pith.science/pith/FDYCWLGEK72SHF7VOMNOJRLDXT/action/author_attestation","sign_citation":"https://pith.science/pith/FDYCWLGEK72SHF7VOMNOJRLDXT/action/citation_signature","submit_replication":"https://pith.science/pith/FDYCWLGEK72SHF7VOMNOJRLDXT/action/replication_record"}},"created_at":"2026-07-05T07:28:28.324304+00:00","updated_at":"2026-07-05T07:28:28.324304+00:00"}