Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:02:47.526084Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 2 inbound Pith citation observations for arXiv:2506.00777.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:02:47.526084Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T07:20:33.617934Z
A source-named dated measurement, never combined with another source.
Source: cited_works
55 of 55 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cbc6415e-83f8-4705-a460-ea733bb3876e · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7106c982-2c18-40da-b7fc-c50c17e8645b · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge GPT-4 Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77d10d3b-a33b-47cb-84f2-7a33f045acd9 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge PaLM 2 Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8b80b04-3649-4f7b-a8e1-866fb883057d · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2b483b88-b595-4566-a055-62d9d73d343a · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae6a5b49-c271-4d04-bbc5-f107955debdb · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Do, Yan Xu, and Pascale Fung
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14591c23-b534-4fe1-a311-27e02139c2e3 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbe47d2d-c4e9-4e09-94f4-5926e32a1963 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Lessons from the Trenches on Reproducible Evaluation of Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdc7ff7e-8cd7-4932-bb5f-d44032f25b2e · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e46f93ed-c5e8-4406-8463-fbfd8a3d8f3f · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Tiny Titans: Can Smaller Large Language Models Punch Above Their Weight in the Real World for Meeting Summarization?
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ec81de2-7bac-4235-9bde-7f605eb4f692 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c16909de-1bcd-45b2-a1ba-703ef69ca796 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge The Llama 3 Herd of Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50cf506a-f7c1-4a37-a1bf-2367ccb2b5fb · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge A Survey on LLM-as-a-Judge
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c4fc371-c7ee-459a-a98d-c2389ef9501d · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Textbooks Are All You Need
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 273682e8-0e02-4749-83a5-286b9146a2da · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cbae3a98-ca16-4dd6-bbdd-44dca82ce701 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 318fbf0a-8b9e-46ba-b30e-1bc1ba27c0f1 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0acd8fa1-ae60-42a3-a3ab-4e4fcb1837a8 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 750697ef-7b14-416b-8c6e-a9e46d38f0a4 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a37e638d-10f7-4dbd-a7d6-7fc50ad630ae · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Mistral 7B
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5c456ea-1544-4d02-b8f0-e6aad9adf624 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9435f877-6007-4322-a81e-c5810a4c9458 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 81544950-e252-4a74-9359-af0383cc8fb3 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 26eb1209-bb7b-4000-a877-b69af6dcc843 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2560fe2a-3da9-4493-8ab7-659330931aed · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 298d0589-20bc-4e25-85c1-96296dd9b862 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06bf94d4-9c2f-48d7-9bff-76396c9df471 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4a2ef83a-53bf-4b90-80a8-e7d056c11299 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning?
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5cadf5ed-49d7-46e9-bd80-4030446042c0 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation aff110c5-0e2c-4e02-a339-daba80ea5fbf · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83333015-ee06-4a8d-b60c-dfcf78e74003 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 42e80012-10ba-4cc7-8113-db1af6a125ff · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2095bc52-8ce1-4095-8456-d2006553892e · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 92117b81-12ff-469f-8f40-7013e7db6479 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ce3edb10-ee22-421e-b374-a70e25b9e77d · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Large Language Models: A Survey
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97fa6cba-6fe9-4112-b633-bd51a6587bc9 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 833e2177-030b-4af2-bc95-7632073d094d · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 26328e55-814b-46e0-b17c-7f2f17807602 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Capabilities of Gemini Models in Medicine
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1eb2724-5824-4238-ae81-2151f47b44b0 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a1c00059-8291-4ba1-8775-412fa7dbda72 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Gemini: A Family of Highly Capable Multimodal Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae4979be-f501-4ceb-9281-b6eb49557a47 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 314fb2e5-19f0-4af3-9c87-50d91fa7f7fd · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge LLaMA: Open and Efficient Foundation Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e742cba8-03c6-46fa-b04f-7a509c554f3c · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68bca6ec-97b6-4e23-84de-f71a635099d1 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a894cab8-c0f5-4d16-a157-c7161b72cab6 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4663eb2b-2686-445f-a5d7-a8a7b0df7032 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 30c32547-5fa8-4a28-bec9-fa48dd432be1 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Qwen2 Technical Report
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95adc132-2783-4092-a826-74124f853fc3 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Qwen2.5 Technical Report
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58ebfe33-87a9-4cf2-acb7-27659c5495b8 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04287b94-8cbe-4d13-a478-4cbb1f7b122d · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82a3b566-caff-4747-837c-9698baec91c2 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge A Survey of Large Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3160186-47ab-43f5-aa75-3981d65cb121 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29b03e43-fffa-4182-a8a2-54e8643a8ea0 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge Unresolved cited work
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fa99093f-eff0-48e5-8bd8-ad20e2821bd3 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge online" 'onlinestring :=
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65b484b4-22af-470f-9f3e-ec5db528a1a3 · outbound
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge write newline
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2697847-5dea-482a-82da-5a066fb655a6 · inbound
Enjoy Your Talk: A Human-Centered Benchmark for Multi-Turn Dialogue with Decoupled User Simulation, Target Modeling, and Judging Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d01d9e4-23ce-496d-8e1e-f52f7e92bc81 · inbound
Enjoy Your Talk: A Human-Centered Benchmark for Multi-Turn Dialogue with Decoupled User Simulation, Target Modeling, and Judging Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.