Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T09:30:47.726480Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 0 inbound Pith citation observations for arXiv:2606.12291.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T09:30:47.726480Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
51 of 51 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0c261a26-5299-4181-9e32-924ec78a814a · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context The evaluation illusion of large language models in medicine.npj Digital Medicine, 8(1):600, 2025
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b80a1a4e-76e0-4561-956d-3f49a38a89a9 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Introducing Claude Sonnet 4.6, February 2026
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5a1742e-396c-4ff5-8470-868ed1648851 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Comparing physician and artificial intelligence chatbot responses to patient questions posted to a public social media forum.JAMA internal medicine, 183(6):589–596, 2023
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67e647d4-03be-4e73-b46b-8dba7ca2f750 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Fries, Michael Wornow, Akshay Swami- nathan, Lisa Soleymani Lehmann, Hyo Jung Hong, Mehr Kashyap, Akash R
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation be36c7f0-cd1a-4761-99c8-606b73022289 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context and Saba, Luca and Hadamitzky, Martin and Kather, Jakob Nikolas and Truhn, Daniel and Cuocolo, Renato and Adams, Lisa C
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7e605dd6-e4a8-4331-bce0-cb9334308c6b · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context MedBench: A Large-Scale Chinese Benchmark for Evaluating Medical Large Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d577bdeb-c2e2-452a-bf3e-cec76a00e335 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Singer, Xuguang Ai, Po-Ting Lai, Zhizheng Wang, et al
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1e241908-2353-4149-87e4-a074900d5322 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context M ed R isk E val: Medical Risk Evaluation Benchmark of Language Models, On the Importance of User Perspectives in Healthcare Settings
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 719ec87d-f4b3-4372-be84-e1568a1de664 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Nour, Seth Spielman, Samuel F
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5351bece-dab5-4883-a7cc-6ef534825137 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Openseeker: Democratizing frontier search agents by fully open-sourcing training data
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 540d2000-a428-44b9-ab3d-16b6a5dfff03 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Gemini 3.1 Flash-Lite model card, March 2026
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5699b5a-6319-465d-be68-7fab4520da9c · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Gemini 3.1 Pro model card, February 2026
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27ac4c67-ad69-4624-9585-ed5d35d10b72 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Gemma 4 model card, April 2026
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a4ed9e4-0b01-4fca-b6dc-362ca410d226 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context InProceedings of the 16th ACM Workshop on Artificial Intelligence and Security (AISec @ CCS 2023)
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 32c75ced-a565-4b9f-9ea8-459d041a706b · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Bressem, Jakob Niko- las Kather, and Daniel Truhn
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1c4768a3-0fc8-408c-9191-6c9e5ae83fe9 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context MedAgentBench: A Realistic Virtual EHR Environment to Benchmark Medical LLM Agents
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4a51a8e8-4057-411b-a30a-9ff8d7640f71 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context What disease does this patient have? A large-scale open domain question answering dataset from medical exams.Applied Sciences, 11(14):6421
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 58b18d2e-4f74-41c8-aaf5-84660a3342ab · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context H., et al
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 89a3a225-a350-4a6e-a4bc-33a2f0d0ab4e · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Evaluating clinical competencies of large language models with a general practice benchmark.Nature Communications, 2026
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1a051535-a6c0-4d00-84d3-dc346cc526ad · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9e798e51-8f1c-4fb7-9a37-9cf94bc46b4f · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Chen, Yining Hua, Peilin Zhou, Junling Liu, Chengfeng Mao, Chenyu You, Xian Wu, Yefeng Zheng, Lei Clifton, Zheng Li, Jiebo Luo, and David A
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3b68e594-aaad-45d0-851a-2c5d26fc447e · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Benchmarking large language models on CMExam - A comprehensive chinese medical exam dataset
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38bcec30-d012-4b2e-a991-839b04c20f97 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Carrero, Xiaofeng Jiang, Dyke Ferber, Georg Wölflein, Li Zhang, Sanddhya Jayabalan, Tim Lenz, Zhouguang Hui, et al
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b64375e8-78c7-4159-b8ff-7db14240cf13 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Beyond the Leaderboard: Rethinking Medical Benchmarks for Large Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 66d22251-a305-46e8-86ef-e56236d08886 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Persuasion with Large Language Models: A Survey of Empirical Evidence, Study Methodologies, and Ethical Implications
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1d8fef7f-8bf1-4cd9-8e4e-b071d2fb5167 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Capabilities of GPT-4 on Medical Challenge Problems
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d88b5189-5ad4-4e53-9185-0cd2cf898024 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Wieler, Alexander W
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5148f2fd-5fa3-4acf-8a4f-09af51f4093f · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Introducing HealthBench, May 2025
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7a54846-05e7-43af-ab70-764dbaacd8e6 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Introducing ChatGPT health, January 2026
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 941e6f5f-006e-42ee-abf5-c930b369fb1a · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Introducing GPT-5.4, March 2026
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49d70b12-958a-434a-ab60-021c82fb2059 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8402bc5-06b9-45e1-b2f0-a77ca98e549d · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Bowman, Amanda Askell, Roger Grosse, Danny Hernandez, Deep Ganguli, Evan Hubinger, Nicholas Schiefer, and Jared Kaplan
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0ab6ef48-f502-416c-9af1-027bd2548c5f · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context A benchmark of expert-level academic questions to assess ai capabilities.Nature, 649(8099):1139–1146, 2026
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 544ca31a-b033-459b-9e6a-345ab9c37178 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Qwen3.6-35B-A3B: Agentic coding power, now open to all, April 2026
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d825c699-3e00-43b2-b048-d8e047d718f7 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Rao, Kaiz P
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b4f2261e-e04c-4420-869d-b162aede1ea0 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 75acb012-7971-4e53-ac1e-1ca640badf4b · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context MedGemma Technical Report
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 88941836-0bfd-4bf4-97ac-daacae0e5e87 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context MedGemma Technical Report
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6bad8ac1-4d39-4408-94d6-2d1db9c1f87d · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Bowman, Esin Durmus, Zac Hatfield-Dodds, Scott R
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e86c6f93-bb8b-420f-adee-b5903f9810c6 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Large Language Models Encode Clinical Knowledge
Reference 40
Source-reported events for the cited work
correction dated 2023-07-27. Source: crossref record 10.1038/s41586-023-06455-0->10.1038/s41586-023-06291-2:correction, observed 2026-07-11T03:08:19.417011+00:00. This notice travels one citation hop only.
Observation 1f9089c2-00cb-4e96-81e9-ee9bd8b32776 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Toward expert-level medical question answering with large language models.Nature medicine, 31(3):943–950, 2025
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06564744-66fb-43d8-9e4a-068065fa7b05 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context J., Ting, D
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2b451404-2d67-4e8f-8c24-2c9344267af2 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Scientists invented a fake disease
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56da7f00-9808-4575-ae31-8c6dedc8d968 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Sara Mahdavi, Christoph er Semturs, Juraj Gottweis, Joelle Barral, Katherine Chou, Greg S
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 246cfae3-a8c3-4950-b0f6-6006f339bc29 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context A novel evaluation benchmark for medical llms illuminating safety and effectiveness in clinical domains.npj Digital Medicine, 2025
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c145b93-6c3a-40ee-81fb-08fbd133268c · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Understanding the infodemic and misinformation in the fight against COVID-19, 2026
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3f29295-7fea-479c-ab2e-e5a41d1454a2 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context Medjourney: Benchmark and evaluation of large language models over patient clinical journey.Advances in Neural Information Processing Systems, 37:87621–87646, 2024
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02ea82fb-5be2-413e-813c-a31f49c1b078 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context npj Health Systems 2025 2:1 2:2-
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 375f9888-66cb-4d10-8354-53ba2a5d62de · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context React: Synergizing reasoning and acting in language models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1476ca49-19fc-46c6-930f-65b89229fa44 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context PoisonedRAG: Knowledge corruption attacks to Retrieval-Augmented generation of large language models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85ae5e85-f44d-490e-8969-4645295422f7 · outbound
Measuring Epistemic Resilience of LLMs Under Misleading Medical Context MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
No inbound Pith citation observations are available.