Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-08T03:29:57.083691Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 2 inbound Pith citation observations for arXiv:2604.24710.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-08T03:29:57.083691Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T13:17:16.965678Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-06-30T12:24:39.879150Z
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation df6a07d0-e3a2-45a7-a91c-1fb9e442ba0b · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters Monitoring performance of clinical artificial intelligence in health care: a scoping review
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1218eebf-ae67-4269-ae0a-2df5f14d1a93 · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters State of Clinical AI Report 2026
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d52e4466-ebd9-4d92-96c2-3936aa81e22c · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters Evaluating clinical AI summaries with large language models as judges
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1d63a446-306d-457d-b556-d8825fbc8135 · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters A framework for human evaluation of large language models in healthcare derived from literature review.npj Digital Medicine
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b6f7ba66-13c0-4c15-9c19-970be585c320 · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters The 9-Item PDQI-9 score is not useful in evaluating EMR note quality in Emergency Medicine.Applied Clinical Informatics
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fbee87b9-f9bb-4688-80e3-597eed91ee70 · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters Assessing the Assessment–Developing a Novel Tool for Eval- uating Clinical Notes’ Diagnostic Assessment Quality.J Gen Intern Med
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e39b0ac2-fa8d-43e0-ab37-6cc74e76d7d5 · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters The impact of inconsistent human annotations on AI driven clinical decision making
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2baa0a62-aab6-485e-b086-8abb02576483 · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters Introducing HealthBench
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 241f6b53-e63a-47fd-af2b-891197c89396 · outbound
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 36a7e73d-e9fd-4682-80ba-695d35b6b16c · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters Sequential Diagnosis with Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3d8cd92d-de23-45b0-8137-2eb238658871 · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters AI-based Clinical Decision Support for Primary Care: A Real-World Study
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 89ed87b0-74ba-4b45-b0ab-544eca6938de · outbound
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9c4a7649-c215-419f-a777-38efe989e0cb · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters canvas-hyperscribe
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 85f1dc4e-9201-440f-a623-0cd186c95eea · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters Canvas SDK Commands
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 28a8a3f3-5245-43c1-862a-2c0f67b9b93a · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters End-to-End Evaluation and Governance of Hyperscribe, an EHR-Embedded Clinical AI Agent
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7eed98ca-1527-44a7-aeb7-7c7acd12d320 · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters ScribeBench: Dataset Usage
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a633164e-b7c8-4fa2-a5d9-dde506938c90 · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters ScribeBench: A Benchmark Dataset for Evaluating AI-Generated Medical Documentation PhysioNet
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 31b36906-5322-461f-b422-9b6fd34a1c08 · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters A new measure of rank correlation.Biometrika
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cd016ca9-8739-47b7-9a07-c8fd5e8383f3 · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters Development and validation of the provider documentation summarization quality instrument for large language models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 869b051d-416e-4ac7-be74-cf89c7aedfa0 · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters npj Digital Medicine8, 274 (2025) https://doi.org/10.1038/s41746-025-01670-7
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d422c0db-f00c-45ef-9826-8b1ef59374a9 · outbound
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters Benchmarking and datasets for ambient clinical documentation: a scoping review.medRxiv
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 65a1575c-f123-43b8-bdfb-c281e70bba87 · inbound
When Reasoning Hurts: Source-Aware Evaluation of Frontier LLMs for Clinical SOAP Note Generation Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fc2325a6-8a90-4925-aec0-e0054a498af5 · inbound
When Reasoning Hurts: Source-Aware Evaluation of Frontier LLMs for Clinical SOAP Note Generation Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.