Citation notice #4619 · 2026-07-11 03:19:08.018010+00:00
The Validity Gap in Health AI Evaluation: A Cross-Sectional Analysis of Benchmark Composition
Correction
Crossref
Open
cites Large lan- guage models encode clinical knowledge,, which carries a correction notice dated 2023-07-27. One-hop deterministic notice: the citation edge exists in the Pith bibliography graph; no model judged whether the citation was load-bearing.
Citing paper Event page Original DOI Notice DOI File a formal challenge All reference changes
01Evidence
Raw extraction · bibliography line · bibliography index 10
Karan Singhal, Shekoofeh Azizi, Tao Tu, S. Sara Mahdavi, Jason Wei, Hyung Won Chung, Nathan Scales, Ajay Tanwani, Heather Cole-Lewis, Stephen Pfohl, Perry Payne, Martin Senevi- ratne, Paul Gamble, Chris Kelly, Abubakr Babiker, Nathanael Sch¨ arli, Aakanksha Chowdh- ery, Philip Mansfield, Dina Demner-Fushman, Blaise Ag¨ uera y Arcas, Dale Webster, Greg S. Corrado, Yossi Matias, Katherine Chou, Juraj Gottweis, Nenad Tomasev, Yun Liu, Alvin Rajkomar, Joelle Barral, Christopher Semturs, Alan Karthikesalingam, and Vivek Natarajan. Large language models encode clinical knowledge.Nature, 620(7972):172–180, August 2023. ISSN 1476-4687. doi: 10.1038/s41586-023-06291-2
02Event
- Type
- Correction
- Source
- Crossref
- Original DOI
- 10.1038/s41586-023-06291-2
- Notice DOI
- 10.1038/s41586-023-06455-0
- Date
- 2023-07-27
- Title
- Publisher Correction: Large language models encode clinical knowledge
- Reasons
- ['Correction']
- Work
- Large lan- guage models encode clinical knowledge, (2023) Nature
03Dispute this notice
If this citation does not depend on the flagged claim, or the event is wrong, say so. Disputes are public. For a signed challenge against the paper itself, use the formal challenge form.