Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2401.16745.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:55:42.704868Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 372e1cc8-c5a1-4a4b-b150-799272fb5cfa · inbound
Towards Efficient and Effective Alignment of Large Language Models MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6371c8b8-2223-47bb-ae69-5fa74cfed51e · inbound
ConsistencyChecker: Tree-based Evaluation of LLM Generalization Capabilities MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a0fa028-b01f-4103-82d8-16698ef36b85 · inbound
A Metamorphic Testing Perspective on Knowledge Distillation for Language Models of Code: Does the Student Deeply Mimic the Teacher? MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 90453220-c5b7-46d2-a3d6-291a51294cbe · inbound
CompliBench: Benchmarking LLM Judges for Compliance Violation Detection in Dialogue Systems MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation adbfc25d-0dd6-496e-91d5-c9389cf933ec · inbound
EditPropBench: Measuring Factual Edit Propagation in Scientific Manuscripts MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bea8f247-7084-4e4f-9d98-954c2932bc65 · inbound
SCICONVBENCH: Benchmarking LLMs on Multi-Turn Clarification for Task Formulation in Computational Science MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c12f2aa1-818a-48a9-82aa-19ae31de65da · inbound
DecisionBench: A Benchmark for Emergent Delegation in Long-Horizon Agentic Workflows MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 92bfd61c-b505-45b1-9217-309de1853398 · inbound
Does Capability Transfer to Subjective Behavior -- and Would Our Instruments Tell Us? A Self-Evolving, Trust-by-Construction Evaluation Paradigm MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.