Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T00:23:11.924161Z
Paper Citation Record · LEDGER
As of 2 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 1 inbound Pith citation observation for arXiv:2604.20441.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T00:23:11.924161Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-01T06:32:01.292127+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-14T14:13:05.342181Z
A source-named dated measurement, never combined with another source.
Source: cited_works
23 of 23 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2847d94c-c931-4c37-a0d7-30ecd0f1aec1 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation f58c9f30-9fa4-49ea-aafa-a2f8016efb60 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Available: https://arxiv.org/abs/2603.04448
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 387dfcc1-819c-40dd-a578-28cfb4369da3 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Artificial hallucinations in ChatGPT: implications in scientific writ- ing.Cureus
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation bdad854f-bcd1-457c-845f-26ebfa7c95e0 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Evaluating large language models and agents in healthcare: key challenges in clinical applications.Intelligent Medicine
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 9408c621-496b-4684-bbb0-a26fdaac6da7 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Survey of hallucination in natural language generation.ACM Computing Surveys
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation ad98ef44-2e97-4ca7-a20d-7c4655b1de5b · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Performance of ChatGPT on USMLE: potential for AI-assisted medical education using large language models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation daf00cea-982c-4247-b2c7-5a2a64badf19 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Capabilities of GPT-4 on Medical Challenge Problems
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 0c63dbd3-3003-4cd2-ab30-81e2ace8e169 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Toward expert-level medical question answering with large language models.Nature Medicine
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation fd1967b1-fdc5-4b47-bb65-25ad8f361079 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills A novel evaluation benchmark for medical LLMs illuminating safety and effectiveness in clinical domains.npj Digital Medicine
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 005e241e-5a16-42f2-bf25-2288a53cdda0 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Large language model agents for biomedicine: a comprehensive review of methods, evaluations, challenges, and future directions.Information
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 4d08f078-c778-4bac-a497-96a05758cf77 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills MedAgentBench: a virtual EHR environment to benchmark medical LLM agents.NEJM AI
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation a22f7ee9-8407-44b0-9ea0-80cc2561f871 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills The clinicians’ guide to large language models: a general perspective with a focus on hallucinations.Interactive Journal of Medical Research
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 565ed65d-653a-4a3a-b728-8822b37d76af · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Human researchers are superior to large language models in writing a medical systematic review in a comparative multitask assessment.Scientific Reports
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 53a6b705-4d3a-4897-b309-2cb9beaf5c3e · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Citation integrity in the age of AI: evaluating the risks of reference hallucination in maxillofacial literature.Journal of Cranio-Maxillofacial Surgery
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 025bbcfd-554f-446a-bb02-500322b85120 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills 2025;12:e80371
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 561e8192-e414-458d-9353-44f1a726abf8 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Systems and software engineering — Systems and software Quality Re- quirements and Evaluation (SQuaRE) — System and software quality models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 4da631c9-e38e-4451-b8c3-71e12503e2cc · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Data structures for statistical computing in Python.Proceedings of the 9th Python in Science Conference
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 63a37b86-0588-457a-aff9-f5ad6307b955 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Pingouin: statistics in Python.Journal of Open Source Software
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation a1fecdbd-0f00-43f0-b644-3fdd6af9e029 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills SciPy 1.0: fundamental algorithms for scientific computing in Python.Nature Methods
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 5b40463b-23a6-4d72-b3a0-629a133577a9 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Scikit-learn: Machine Learning in Python.Journal of Machine Learning Research
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation b95c98ec-d528-42ba-bfa1-afe077dfc496 · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills A guideline of selecting and reporting intraclass correlation coefficients for reliability research.Journal of Chiropractic Medicine
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 5408ba97-481c-4d44-9529-a39d5275c84b · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Weighted kappa: nominal scale agreement with provision for scaled disagreement or partial credit.Psychological Bulletin
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 11af07d3-5b69-4f4b-a94a-f244269f03fe · outbound
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills Statistical methods for assessing agreement between two methods of clinical measurement.The Lancet
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 976039e1-885a-41ad-9f78-d3851b1bacd5 · inbound
Dynamic Agent Skills: A Lifecycle Survey and Taxonomy of Evolving Skill Libraries MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.