Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T14:21:57.784836Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 0 inbound Pith citation observations for arXiv:2508.21430.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T14:21:57.784836Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
20 of 20 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 38e85c1a-bdd7-4236-85b9-cbcae452f34a · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c319738c-c5eb-41f6-af4c-93a91875a08b · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 776a21ff-b22c-4a04-8626-15bd2e718400 · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 67f35a9b-fae1-42e3-9332-c2d97c799206 · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9d13ea83-753d-468b-896d-04d4127abcf0 · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 80d65daf-bf30-4c19-a3bb-6c5f5417cd2a · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models MM-Eval: A Multilingual Meta-Evaluation Benchmark for LLM-as-a-Judge and Reward Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da040793-6afc-4da9-a976-96a115cf4ac3 · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa498442-b8c9-46d0-9379-96303932e058 · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models Evaluating Judges as Evaluators: The JETTS Benchmark of LLM-as-Judges as Test-Time Scaling Evaluators
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 548f9eee-b684-45d2-80d9-790aefc92624 · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models Avoid any position biases and ensure that the order in which the responses were presented does not influence your decision
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ecf46c95-0526-423c-bcef-c21d657f7c6f · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models immediate PET-CT for a suspected pulmonary nodule,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation aba0a1e2-a3c7-44fd-9f42-d2956b6cc78f · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2bcceeaf-4caa-4493-97e0-180d924d1cab · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 79d69425-8446-4115-b84c-99b5157d7159 · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d3b7a27a-046c-480a-a762-fbfb440a9ae0 · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models The most likely cause is... I recommend
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6e29a972-f3a3-461c-a627-10b8fdb0f1e4 · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models Additional Notes - All medical images are de-identified from real clinical cases
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 527bb06d-1583-48f6-a375-a6a5466e66eb · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models Fenglin Liu, Tingting Zhu, Xian Wu, Bang Yang, Chenyu You, Chenyang Wang, Lei Lu, Zhangdai- hong Liu, Yefeng Zheng, Xu Sun, et al
Reference 1654
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 33c3338e-64c2-4b7c-a62f-2420b1d25284 · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models Judge Anything: MLLM as a Judge Across Any Modality
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f89be8b3-0511-4bd6-85c4-d878df93c4c4 · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models VLRMBench: A Comprehensive and Challenging Benchmark for Vision-Language Reward Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce46be49-8ef7-4a6a-b82b-5cca7ba006c3 · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models GSCo: Towards Generalizable AI in Medicine via Generalist-Specialist Collaboration
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 776d64b9-c31a-422a-b910-27d5a4a555fc · outbound
Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models Qwen2.5-VL Technical Report
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.