Reference change · event page
Reference changes · DOI
Nature Medicine 28, 924–933
Published notice on a work cited in the Pith corpus. Exact quotes below. No model judges whether any citation was load-bearing.
This page records that a citing paper's bibliography includes a work with a published notice. It is not a judgment on the citing paper.
Correction
Crossref
5 open · 5 total · 0 disputed
- Event date
- 2022-08-12
01One-hop citing occurrences
Correction
Open
Frontier Lag: A Bibliometric Audit of Capability Misrepresentation in Academic AI Evaluation
ref [5] ·
2605.04135
· notice #5628
· dispute
Raw extraction · bibliography line
URL https://www.aisi.gov.uk/frontier-ai-trends-report. First public evidence-based assessment aggregating two years of AISI’s frontier model testing (November 2023 through October 2025); cited for the frontier-trajectory reframe of capability evaluation. Baptiste Vasey, Myura Nagendran, others, and DECIDE-AI Expert Group. Reporting guideline for the early-stage clinical evaluation of decision support systems driven by artificial intelligence: DECIDE-AI.Nature Medicine, 28(5):924–933, 2022. doi: 10.1038/s41591-022-01772-9. Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc Le, Ed Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou. Self-consistency improves chain of thought reasoning in language models. In International Conference on Learning Representations (ICLR), 2023. Self-consistency gains of+6.4– +17.9pp on math / reasoning benchmarks; used to calibrate the sampling-axis chip conservatively for SWE-Bench-Verified pass@1. Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, and Denny Zhou. Chain-of-thought prompting elicits reasoning in large language models. In Advances in Neural Information Processing Systems (NeurIPS), 2022. Companio
Correction
Open
Frontier Lag: A Bibliometric Audit of Capability Misrepresentation in Academic AI Evaluation
ref [5] ·
2605.04135
· notice #5630
· dispute
Raw extraction · bibliography line
URL https://www.aisi.gov.uk/frontier-ai-trends-report. First public evidence-based assessment aggregating two years of AISI’s frontier model testing (November 2023 through October 2025); cited for the frontier-trajectory reframe of capability evaluation. Baptiste Vasey, Myura Nagendran, others, and DECIDE-AI Expert Group. Reporting guideline for the early-stage clinical evaluation of decision support systems driven by artificial intelligence: DECIDE-AI.Nature Medicine, 28(5):924–933, 2022. doi: 10.1038/s41591-022-01772-9. Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc Le, Ed Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou. Self-consistency improves chain of thought reasoning in language models. In 43 International Conference on Learning Representations (ICLR), 2023. Self-consistency gains of+6.4– +17.9pp on math / reasoning benchmarks; used to calibrate the sampling-axis chip conservatively for SWE-Bench-Verified pass@1. Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, and Denny Zhou. Chain-of-thought prompting elicits reasoning in large language models. In Advances in Neural Information Processing Systems (NeurIPS), 2022. Compa
Correction
Open
The Open-Box Fallacy: Why AI Deployment Needs a Calibrated Verification Regime
ref [37] ·
2605.10601
· notice #5629
· dispute
Raw extraction · bibliography line
Baptiste Vasey, Myura Nagendran, Bruce Campbell, David A. Clifton, Gary S. Collins, Spiros Denaxas, Alastair K. Denniston, et al. Reporting guideline for the early-stage clinical evaluation of decision support systems driven by artificial intelligence: DECIDE-AI.Nature Medicine, 28:924–933, 2022. doi: 10.1038/s41591-022-01772-9
Correction
Open
LLMs in the Real World: Evaluating "AI" in Emergency Contexts
ref [80] ·
2607.00019
· notice #5631
· dispute
Raw extraction · bibliography line
Baptiste Vasey, Myura Nagendran, Bruce Campbell, David A Clifton, Gary S Collins, Spiros Denaxas, Alastair K Denniston, Livia Faes, Bart Geerts, Mudathir Ibrahim, Xiaoxuan Liu, Bilal A Mateen, Piyush Mathur, Melissa D McCradden, Lauren Morgan, Johan Ordish, Chris Rogers, Suchi Saria, Daniel Shu Wei Ting, and 4 others. 2022. https://doi.org/10.1038/s41591-022-01772-9 Reporting guideline for the early-stage clinical evaluation of decision support systems driven by artificial intelligence: DECIDE-AI . Nature Medicine, 28(5):924--933
Correction
Open
Deep Learning for Semen Analysis in Male Infertility: Computer Vision, Multimodal Fusion, and Clinical Translation
ref [109] ·
2607.05311
· notice #5632
· dispute
Raw extraction · bibliography line
Reporting guideline for the early-stage clinical evaluation of decision support systems driven by artificial intelligence: Decide-ai. Nature Medicine 28, 924–933. doi:10.1038/s41591-022-01772-9