Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:52:16.027375Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2505.11887.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:52:16.027375Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
39 of 39 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation fe99e6b2-9ad4-4833-9ce6-5a156b5afac0 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 736cacf7-1b7a-4eea-961a-98d75003af34 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation A Survey on Evaluation of Large Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ae1315f-b2e0-4ea2-be39-cdbecc09e727 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Introspective Tips: Large Language Model for In-Context Decision Making
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3a4a363-d82e-462d-9b14-f15d2b378a85 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 063b6d11-0eb9-4b03-9afa-a12882c2c04f · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d7314dc-5800-4c7e-9114-d2d66514e84f · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Fu, Stefano Ermon, Atri Rudra, and Christopher R \'e
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36d53d3c-ce5d-4da7-9ce9-5c06e55ddab4 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ae0853d-ebca-4618-beea-738a08a39a98 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation SimCSE: Simple Contrastive Learning of Sentence Embeddings
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0ff8caf-385a-48e6-b8e6-99532332d0a0 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Toward a Formal Model of Cognitive Synergy
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation bfa08233-90d0-4f6c-b2fe-e0a910ef035a · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89f45ce1-1142-4618-b202-8b989b4a4413 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcc36fad-f0dd-4f5d-92a2-66b04e719431 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 75244dc7-19e0-4da3-833e-598a2453cda8 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Generative Judge for Evaluating Alignment
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0546bd6-e589-410c-8fa1-fd9973c40731 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60031fa8-d1b2-4d9e-ac79-746eb84c8c87 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Leveraging Large Language Models for NLG Evaluation: Advances and Challenges
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ce9db86-1999-470d-8155-96053a3c25ec · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cec2532-cd5a-4414-8135-6becd8038079 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85bb0d07-65c8-4179-8b85-d7dbabe780bd · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 99621ce5-1fcb-42a1-be4e-5ba4f055cfb8 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 27a5002f-8dd5-471b-99bd-417cee32d824 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Capabilities of GPT-4 on Medical Challenge Problems
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bf04891-3e13-4240-be05-c998aae1a61a · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec796b18-65fd-4cc5-a77a-8f7563445835 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fd885281-6373-4f5d-9fe6-2571e34e8ebb · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation ZeRO: Memory Optimizations Toward Training Trillion Parameter Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26999ed6-402f-43ee-8b98-d0ecba9448f0 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec019236-0aeb-45d0-9fbb-c2e864adf274 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d52e11df-71b6-4140-b3e9-f58d670e1e82 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Towards Expert-Level Medical Question Answering with Large Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c502c932-4b3f-4e1b-8c9b-c0d10e69fae9 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Towards Generalist Biomedical AI
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62e19b43-25cc-4929-bff4-926201a354b8 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Is ChatGPT a Good NLG Evaluator? A Preliminary Study
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31e83056-c601-4f4b-9e4c-0b4f6bc2ea90 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning Optimization
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ed5c82e-5a2f-4981-85b3-0e38ec60b5f2 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Metacognitive Prompting Improves Understanding in Large Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6a6421d-a7c4-4106-9d8f-69d28feb8bf5 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation PMC-LLaMA: Towards Building Open-source Language Models for Medicine
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44085475-2b97-4358-981d-f60b41e82d89 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation DoctorGLM: Fine-tuning your Chinese Doctor is not a Herculean Task
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f264726-f235-421e-a510-d611fd71af91 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Baize: An Open-Source Chat Model with Parameter-Efficient Tuning on Self-Chat Data
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bcb0c7c-1f6a-42b9-8737-f99deb595d32 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation MedGPTEval: A Dataset and Benchmark to Evaluate Responses of Large Language Models in Medicine
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a442ca2-7af7-4e24-b452-500945c5925b · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3741cb3-ea52-4b82-ac10-7fd2b015fbae · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation BERTScore: Evaluating Text Generation with BERT
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a460ea7c-0bef-4bbc-860a-30cb27ad795b · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9f921f0-261d-437f-a04e-24c482c895c0 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation online" 'onlinestring :=
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1da77e8c-2915-4da3-9140-e64a5116ad75 · outbound
AutoMedEval: Harnessing Language Models for Automatic Medical Capability Evaluation write newline
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.