Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:42:53.123407Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 76 of 76 outbound references and 3 inbound Pith citation observations for arXiv:2505.24040.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:42:53.123407Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-26T17:02:30.696295Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T04:19:34.994254Z
76 of 76 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 896e37bd-4420-440b-b58a-955634184988 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Can We Use Large Language Models to Fill Relevance Judgment Holes?
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ba8765b-c6d5-4427-a016-69206bdb2628 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Evaluating Correctness and Faithfulness of Instruction-Following Models for Question Answer- ing.Transactions of the Association for Computational Linguistics, 12:681–699, 2024
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation edb08fed-c0a0-4444-a431-21e84c342fa6 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Prompt-Reverse Inconsistency: LLM Self-Inconsistency Beyond Generative Randomness and Prompt Paraphrasing
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcc6e987-7e02-414b-b536-0475620f76c7 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering LLM Stability: A detailed analysis with some surprises.CoRR, January 2024
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a7088955-a11b-40fc-b0f9-5cbba743a2c1 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Does the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team Performance
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5c5091e5-d6b4-4eb1-9bbc-097a5aa693bf · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering LLMs with Chain-of-Thought Are Non-Causal Reasoners.CoRR, January 2024
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c962acac-362b-44e9-a52c-ab7da4651209 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2dab7938-f13c-4679-b569-71c0f1da83aa · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 605ec00f-5428-4b24-8de1-0fb383a0b60b · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Clinical Reasoning of a Generative Artificial Intelligence Model Compared With Physicians.JAMA Internal Medicine, 184(5):581–583, May 2024
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0022eabe-eae4-4e43-a1d3-c13a3cc5b603 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Barnhill, Mar Llamas-Velasco, Gabriela Poch, Sören Korsing, Wiebke Sondermann, Frank Friedrich Gellrich, Markus V
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1aa9cc7f-e276-443d-bc5e-9ac402ac6630 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Benchmarking Large Language Models on Answering and Explaining Challenging Medical Questions
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 60895be4-10d9-4b8e-81bf-5fa59db03965 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Reasoning Models Don’t Always Say What They Think
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a333b942-e8db-48f9-a0b5-f255f65bab64 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3c51a8e8-fc08-426b-8379-074619e98521 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering What is Your Data Worth to GPT? LLM-Scale Data Valuation with Influence Functions
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bebac39-b22f-4dc8-b55c-b2c82906396f · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Identifying Key Terms in Prompts for Relevance Evaluation with GPT Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e558dbca-b791-4367-8d0d-00a90e38525e · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering SelfCite: Self-Supervised Alignment for Context Attribution in Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92323745-e460-41c4-859c-9603845f493c · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Learning to Attribute with Attention
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbbf625c-d841-4454-9cb7-1168e8422158 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering ContextCite: Attributing Model Generation to Context
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 25c4cf0f-5e78-460e-a01f-5b4faf721827 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Current and future state of evaluation of large language models for medical summarization tasks.npj Health Systems, 2(1):1–13, February 2025
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5cd13f48-69b2-4551-8d11-b8913902161d · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Skinner, Ariel Dora Stern, and David Wennberg
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5f85e1c6-8d58-4b93-af8a-3d4ccc12c41c · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8646b9b0-66ce-46f4-a2ce-7eab95d78398 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering RAGAs: Automated Evaluation of Retrieval Augmented Generation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ebedf70f-cde7-4367-bdd8-4953545d8098 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Detecting hallucinations in large language models using semantic entropy.Nature, 630(8017):625–630, June 2024
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 956e18ff-2a35-4c26-9997-0ba048f0648e · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering CiteBench: A Benchmark for Scientific Citation Text Generation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d2fafe29-dff1-4c5e-8cf7-779d624989c9 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Enabling Large Language Models to Generate Text with Citations
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7ac6e3ac-3e43-45eb-9f47-5d1ed96a87bc · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Koch, Matthias F
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 933f896e-63d6-4172-a219-43e18b2ff4bb · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering The Llama 3 Herd of Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e4e0f94-6992-457e-849c-f11294198349 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Large Language Models lack essential metacognition for reliable medical reasoning.Nature Communications, 16(1):642, January 2025
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 858e417c-5da7-4210-8a11-adbd5dc72bad · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering A Survey on LLM-as-a-Judge
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 414ad4e3-3197-4a1c-9a74-43f0ede88883 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering McKone, Daniel K
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 31db672f-c43e-4c3f-88fc-749c9f84ae2d · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Measuring Massive Multitask Language Understanding
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49b3ef4f-48b4-4ec8-89a4-4895b76b0d5b · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Spurious
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 554300d1-91e4-4681-a9c6-7bc73458aaec · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering RJUA-MedDQA: A Multimodal Benchmark for Medical Document Question Answering and Clinical Reasoning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 00d8c488-aa8c-4d0a-9c97-98bf7547a73a · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering What Disease Does This Patient Have? A Large-Scale Open Domain Question Answering Dataset from Medical Exams.Applied Sciences, 11(14):6421, January 2021
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e5109b91-6a33-4bdf-baf7-1efaf4f3b2da · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering PubMedQA: A Dataset for Biomedical Research Question Answering
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3e55d466-c684-44d0-8684-7aa9cee9bae2 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Effective Context Selection in LLM- Based Leaderboard Generation: An Empirical Study
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e2f9176e-86cf-4e67-948e-bbbf3a2d91a8 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering GPT versus Resident Physicians — A Benchmark Based on Official Board Scores.NEJM AI, 1(5):AIdbp2300192, April 2024
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1f555e5e-8ec6-48f4-a962-e52dba343d5a · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Baleen: robust multi-hop reasoning at scale via condensed retrieval
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f6434b9f-01aa-484f-a74c-930e59a325b9 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Li, Vidhisha Balachandran, Shangbin Feng, Jonathan S
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1f67cda6-ea80-4875-82ee-58408f191d64 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering AttriBoT: A Bag of Tricks for Efficiently Approximating Leave-One-Out Context Attribution
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 56fd67af-3220-4adc-af55-53276c8e0353 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 64a4ff52-f181-4bd4-9cf0-d4020f078ffa · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Sara Mahdavi, Sushant Prakash, Anupam Pathak, Christopher Semturs, Shwetak Patel, Dale R
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 850b2af6-8012-430e-9571-07616fc036ad · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Context Example Selection for LLM Generated Relevance Assessments
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2a0765b8-280c-4ef4-b802-90f2dee564db · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering GPT-4o System Card
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6203c3c4-7570-4e33-a864-ccb56c174205 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering MedMCQA: A Large- scale Multi-Subject Multi-Choice Dataset for Medical domain Question Answering
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ecd5867d-7b9f-47f0-9424-6e4c838f586b · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Bowman, and Shi Feng
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 41b30a8b-73d9-463b-89dd-a2fee267dff3 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Unresolved cited work
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 572f7938-3816-440f-8d60-01db2dac20db · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Qwen2.5 Technical Report
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a2b63cc-2784-4256-b394-ab564a89868e · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Benchmarking Prompt Sensitivity in Large Language Models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9ad4cffd-c04b-41da-8b46-f9574415308d · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Towards Human-Centered Explainable AI: A Survey of User Studies for Model Explanations.IEEE Trans
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 03c0d856-2e1d-4243-ae7b-523cd8d01e8a · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Unresolved cited work
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b6a919f8-20ca-478d-b579-b69def8345e1 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Jung, Maria Zerlik, Waldemar Hahn, Martin Sedlmayr, and Brita Sedlmayr
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e97d18c-f69c-4a7f-a6a0-6b9cd9d21e2a · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Relevance of Unsupervised Metrics in Task-Oriented Dialogue for Evaluating Natural Language Generation
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 618559c9-d48e-4adc-84e1-0c84343e0e0d · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Judging the Judges: A Systematic Study of Position Bias in LLM-as-a-Judge, April 2025
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f00423a-8a68-4d61-805b-db539a1b02dc · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Pfohl, Heather Cole-Lewis, Darlene Neal, Qazi Mamunur Rashid, Mike Schaekermann, Amy Wang, Dev Dash, Jonathan H
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 00ce585b-331c-4895-a2de-ba017a47fc50 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Don’t Use LLMs to Make Relevance Judgments.Information Retrieval Research, 1(1):29–46, March 2025
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 332a4770-99a4-4ef3-a481-e6fccb35aeae · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering RadQA: A Question Answer- ing Dataset to Improve Comprehension of Radiology Reports
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 131b2e92-d6ca-4370-ae6e-8b32163ac763 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Unresolved cited work
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d75952b-8710-4d0a-be25-59fb2c34cfd8 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Clinical Camel: An Open Expert-Level Medical Language Model with Dialogue-Based Knowledge Encoding
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4718ff6a-cf92-411f-8310-476c7ff3d371 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Prompt engineering in consistency and reliability with the evidence-based guideline for LLMs
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9773ce0e-f307-480e-abf7-c1e95ab90dc7 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 012df2de-7e4e-421c-8f92-ba936bbb4b9f · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering An automated framework for assessing how well LLMs cite relevant medical references.Nature Communications, 16(1):3615, April 2025
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1758cd8c-4c30-4e3c-9b32-823590261821 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering CARES: A Comprehensive Benchmark of Trust- worthiness in Medical Vision Language Models.Advances in Neural Information Processing Systems, 37:140334–140365, December 2024
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fa70b999-29d9-4a59-a506-b046a9b0b448 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Harnessing Biomedical Literature to Calibrate Clinicians’ Trust in AI Decision Support Systems
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation edca9806-32b8-4907-b8a0-62b4c8b149a1 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering A survey of datasets in medicine for large language models.Intelligence & Robotics, 4(4):457–478, December 2024
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ed446d7e-86b8-46f0-bad2-1b772a81272b · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Meyer, and Steffen Eger
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 30d7ea68-3a86-48f9-8f68-d268b451553a · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Gonzalez, and Ion Stoica
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ee80f7e-a2d4-4c6b-a799-220692457244 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Melton, James Zou, and Rui Zhang
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f548da07-6d9c-438d-bc51-f2528ac72cc9 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bf35d46-e107-48e7-bb1b-8f8c404da2b8 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Unresolved cited work
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e67608e9-17db-4bdc-8e0d-51711bed76e4 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Unresolved cited work
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 753df1bd-8613-42a7-8c2a-9c5e9cbedbed · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Unresolved cited work
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 76b99462-dda6-4b93-ac42-1f74f8952046 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Her new job involves walking several miles daily across a large facility
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 12000119-02f6-4bc8-b0e2-c5e33039fb90 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Towards Digital Sustainability in Health Care: Developing Digital Health Products through Data-Driven User Insights
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d72d57a3-c8b8-4355-b7c5-2d08ac677448 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Attributed Question Answering: Evaluation and Modeling for Attributed Large Language Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2972f1e-6fff-451e-b275-05a5a6bfbc06 · outbound
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering Unresolved cited work
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2b5d90d-a592-4cfb-a612-656ac73f0cb7 · inbound
Treatment, evidence, imitation, and chat MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 754b740e-9581-46b2-9fe4-fbb513d5b8b2 · inbound
Chain of Risk: Safety Failures in Large Reasoning Models and Mitigation via Adaptive Multi-Principle Steering MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3911e0e9-f2f5-4471-a71b-d8a61893a70b · inbound
Automating SKILL.md Generation for Computer-Using Agents via Interaction Trajectory Mining MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.