Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2406.07545.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:41:40.029820Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T10:17:57.221151Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation cbee2a8d-6a73-4c03-bd79-2bfedb1e788f · inbound
VoiceBench: Benchmarking LLM-Based Voice Assistants Open-LLM-Leaderboard: From Multi-choice to Open-style Questions for LLMs Evaluation, Benchmark, and Arena
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c79cca48-3ef8-47d3-b589-8fb08bb09684 · inbound
LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods Open-LLM-Leaderboard: From Multi-choice to Open-style Questions for LLMs Evaluation, Benchmark, and Arena
Reference 166
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d031f4f1-005b-4baf-8be1-021ff3e3a4fa · inbound
Human-Centric Evaluation for Foundation Models Open-LLM-Leaderboard: From Multi-choice to Open-style Questions for LLMs Evaluation, Benchmark, and Arena
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47cbe7f8-93d0-48a8-82ed-ba153745657c · inbound
DRAG: Distilling RAG for SLMs from LLMs to Transfer Knowledge and Mitigate Hallucination via Evidence and Graph-based Distillation Open-LLM-Leaderboard: From Multi-choice to Open-style Questions for LLMs Evaluation, Benchmark, and Arena
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a45c7a3-b8cd-497e-8811-e2f50673d835 · inbound
Psycholinguistic Word Features: a New Approach for the Evaluation of LLMs Alignment with Humans Open-LLM-Leaderboard: From Multi-choice to Open-style Questions for LLMs Evaluation, Benchmark, and Arena
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fe4b995-1c8b-407f-b102-12b70a3744c7 · inbound
SLM-Bench: A Comprehensive Benchmark of Small Language Models on Environmental Impacts--Extended Version Open-LLM-Leaderboard: From Multi-choice to Open-style Questions for LLMs Evaluation, Benchmark, and Arena
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7727e9bb-03d0-4a6f-908e-27a44bdca87b · inbound
Position: AI Evaluations Should be Grounded on a Theory of Capability Open-LLM-Leaderboard: From Multi-choice to Open-style Questions for LLMs Evaluation, Benchmark, and Arena
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 437054b5-47cb-4991-b811-c6b8ac66e482 · inbound
Efficient Evaluation of LLM Performance with Statistical Guarantees Open-LLM-Leaderboard: From Multi-choice to Open-style Questions for LLMs Evaluation, Benchmark, and Arena
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9abee72f-ed8b-4cc7-bed1-4b05ec0a9710 · inbound
KNIGHT: Knowledge Graph-Driven Multiple-Choice Question Generation with Adaptive Hardness Calibration Open-LLM-Leaderboard: From Multi-choice to Open-style Questions for LLMs Evaluation, Benchmark, and Arena
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9726ac4-c6f7-42ef-9a95-0595c74e9aae · inbound
Safety Under Scaffolding: How Evaluation Conditions Shape Measured Safety Open-LLM-Leaderboard: From Multi-choice to Open-style Questions for LLMs Evaluation, Benchmark, and Arena
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a9037bc-3af2-4654-baba-8e4918c7f508 · inbound
HumorRank: A Tournament-Based Leaderboard for Evaluating Humor Generation in Large Language Models Open-LLM-Leaderboard: From Multi-choice to Open-style Questions for LLMs Evaluation, Benchmark, and Arena
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9cd40c27-7468-4b36-8464-767b41362652 · inbound
HumorRank: A Tournament-Based Leaderboard for Evaluating Humor Generation in Large Language Models Open-LLM-Leaderboard: From Multi-choice to Open-style Questions for LLMs Evaluation, Benchmark, and Arena
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d54dfb1-535a-410c-bd23-840e74d9d914 · inbound
Improving Cross-Format Robustness in Language Models with Multi-Format Training Open-LLM-Leaderboard: From Multi-choice to Open-style Questions for LLMs Evaluation, Benchmark, and Arena
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.