Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:57:11.870689Z
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 1 inbound Pith citation observation for arXiv:2509.10843.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:57:11.870689Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-18T00:23:21.332115Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-18T00:25:32.869703Z
19 of 19 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation bcb98488-40dd-41ff-9e6d-49e7f509fad8 · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering Guidelines & Statements Search , 2025
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 30dbe4df-1018-41f6-9a25-655f54c99ebc · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering Search Cochrane Library , 2025
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 70ef8dc9-8e59-4431-b005-c8eacf219689 · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering Adams, Felix Busch, Conor Fallon, Marc Huppertz, Robert Siepmann, Philipp Prucker, Nadine Bayerl, Daniel Truhn, Marcus Makowski, Alexander Löser, and Keno K
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2745a473-7938-4cbc-9820-b7ebdbd5b8f8 · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering PubMedQA: A Dataset for Biomedical Research Question Answering
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eaa6e95b-cbb6-4059-98f2-78b3ee41b012 · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering Evaluating Open-Domain Question Answering in the Era of Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f46bea3-b453-41d1-b4c3-da0e09238c22 · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering Gpt versus resident physicians—a benchmark based on official board scores
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 0d44aa29-7f2b-43f6-856b-476481a5bf32 · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering BioASQ - QA : A manually curated corpus for Biomedical Question Answering
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f046df8d-a272-42c6-97cb-7f0d37ff4dc3 · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering MedGUIDE: Benchmarking Clinical Decision-Making in Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d028c16-2f4b-4767-ac8c-8af172114e6d · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering Kragen: a knowledge graph-enhanced rag framework for biomedical problem solving using large language models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 38004b69-c1a4-422f-b45d-332e23d7fb2a · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering Can Large Language Models Match the Conclusions of Systematic Reviews?
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5730e682-1ccd-497d-99f9-070caef9d06d · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering It’s time to bench the medical exam benchmark, 2025
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef0b8d0b-1cf0-4530-af33-dc4b39ccbafe · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering Pfohl, Heather Cole-Lewis, Darlene Neal, Qazi Mamunur Rashid, Mike Schaekermann, Amy Wang, Dev Dash, Jonathan H
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c7ff269-1589-460e-a44c-c63a88ffbaf3 · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering Generating explanations in medical question-answering by expectation maximization inference over evidence
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation cc5da50a-e1b7-4687-92f4-4d904113113d · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering HealthFC: Verifying Health Claims with Evidence-Based Medical Fact-Checking
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 042b0fa7-1595-4ce4-a6d6-74ee3cc69bd5 · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering MedREQAL: Examining Medical Knowledge Recall of Large Language Models via Question Answering
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b3a15392-64d2-4fc6-9ece-2b41cd517868 · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering What Evidence Do Language Models Find Convincing?
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3eb83bf-ca2e-4428-9af7-9bb8448cca01 · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering Woolf, Richard Grol, Allen Hutchinson, Martin Eccles, and Jeremy Grimshaw
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation d5141b22-2827-434e-96c4-42d1c9265c09 · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering Benchmarking retrieval-augmented generation for medicine
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e946e8db-9f49-47b7-a1d3-335d06a97acf · outbound
Evaluating Large Language Models for Evidence-Based Clinical Question Answering MIRIAD: Augmenting LLMs with millions of medical query-response pairs
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60b077cb-e0be-471b-a74e-db5cbfe18b89 · inbound
Contradictions in Context: Challenges for Retrieval-Augmented Generation in Healthcare Evaluating Large Language Models for Evidence-Based Clinical Question Answering
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.