Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:10:35.941849Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2504.13475.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:10:35.941849Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8b9d883c-279a-4465-8870-100ee51bf74b · outbound
LLM Sensitivity Evaluation Framework for Clinical Diagnosis Large Language Models Are State-of-the-Art Evaluators of Translation Quality
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8835d984-0a6f-499d-93c1-1cb5cab1c9a3 · outbound
LLM Sensitivity Evaluation Framework for Clinical Diagnosis Large Language Models Understand and Can be Enhanced by Emotional Stimuli
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8fc6580-3551-410a-859a-a66eb33dba92 · outbound
LLM Sensitivity Evaluation Framework for Clinical Diagnosis GPT-4 Technical Report
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf92fdf1-6516-4668-aff5-daa45834ef69 · outbound
LLM Sensitivity Evaluation Framework for Clinical Diagnosis Large Language Models Sensitivity to The Order of Options in Multiple-Choice Questions
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d301c822-3210-47fa-8aeb-80c9c8dd3cb0 · outbound
LLM Sensitivity Evaluation Framework for Clinical Diagnosis Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47840c07-0c71-44c8-9abe-b9e03cdc7329 · outbound
LLM Sensitivity Evaluation Framework for Clinical Diagnosis The Earth is Flat because...: Investigating LLMs' Belief towards Misinformation via Persuasive Conversation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fc91d06-11cc-44f1-ba2b-2dc64ddfc8bd · outbound
LLM Sensitivity Evaluation Framework for Clinical Diagnosis Enhancing Small Medical Learners with Privacy-preserving Contextual Prompting
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 04aa36d6-dadf-418f-adf7-55bfaf4a8f37 · outbound
LLM Sensitivity Evaluation Framework for Clinical Diagnosis Large Language Models Are Not Robust Multiple Choice Selectors
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad76e021-1aea-42e8-8e20-4a992a0b2f88 · outbound
LLM Sensitivity Evaluation Framework for Clinical Diagnosis MultifacetEval: Multifaceted Evaluation to Probe LLMs in Mastering Medical Knowledge
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1706e4f7-6afc-40ad-b785-950f7bf90a1a · outbound
LLM Sensitivity Evaluation Framework for Clinical Diagnosis Language Models are Few-Shot Learners
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad5baca5-e8af-4451-9bbd-b265eb7ef7ad · outbound
LLM Sensitivity Evaluation Framework for Clinical Diagnosis Training language models to follow instructions with human feedback
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3337559f-3d8d-4f38-aa07-03494a23c98b · outbound
LLM Sensitivity Evaluation Framework for Clinical Diagnosis MedBench: A Large-Scale Chinese Benchmark for Evaluating Medical Large Language Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c398c02a-4530-4f0f-b024-57d4ca3cf920 · outbound
LLM Sensitivity Evaluation Framework for Clinical Diagnosis Principled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.