Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2401.16745.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T16:53:23.265499Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation e09085b0-df98-48c3-b9dc-4cc05a50a673 · inbound
LLM-as-an-Interviewer: Beyond Static Testing Through Dynamic LLM Evaluation MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83aea4ad-7624-47e5-8d3b-e0ca6f8fcdae · inbound
MORTAR: Multi-turn Metamorphic Testing for LLM-based Dialogue Systems MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd26b6f7-8aad-4975-9346-9dd74318f6ac · inbound
Evaluating and Enhancing LLMs for Multi-turn Text-to-SQL with Multiple Question Types MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db642df7-edfb-4da8-bde4-4a837a70e2f1 · inbound
PPTAgent: Generating and Evaluating Presentations Beyond Text-to-Slides MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 570ba48a-6af6-47d8-85df-0e53f3d8e0f4 · inbound
MDEval: Evaluating and Enhancing Markdown Awareness in Large Language Models MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 372e1cc8-c5a1-4a4b-b150-799272fb5cfa · inbound
Towards Efficient and Effective Alignment of Large Language Models MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6371c8b8-2223-47bb-ae69-5fa74cfed51e · inbound
ConsistencyChecker: Tree-based Evaluation of LLM Generalization Capabilities MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4950a1a8-6d54-4e7b-9100-741d92b2e643 · inbound
ReSURE: Regularizing Supervision Unreliability for Multi-turn Dialogue Fine-tuning MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f84cbc8-05ef-4b6a-be83-5eaf92c4356d · inbound
GEM-Bench: A Benchmark for Ad-Injected Response Generation within Generative Engine Marketing MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a0fa028-b01f-4103-82d8-16698ef36b85 · inbound
A Metamorphic Testing Perspective on Knowledge Distillation for Language Models of Code: Does the Student Deeply Mimic the Teacher? MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 90453220-c5b7-46d2-a3d6-291a51294cbe · inbound
CompliBench: Benchmarking LLM Judges for Compliance Violation Detection in Dialogue Systems MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation adbfc25d-0dd6-496e-91d5-c9389cf933ec · inbound
EditPropBench: Measuring Factual Edit Propagation in Scientific Manuscripts MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation bea8f247-7084-4e4f-9d98-954c2932bc65 · inbound
SCICONVBENCH: Benchmarking LLMs on Multi-Turn Clarification for Task Formulation in Computational Science MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c12f2aa1-818a-48a9-82aa-19ae31de65da · inbound
DecisionBench: A Benchmark for Emergent Delegation in Long-Horizon Agentic Workflows MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 92bfd61c-b505-45b1-9217-309de1853398 · inbound
Does Capability Transfer to Subjective Behavior -- and Would Our Instruments Tell Us? A Self-Evolving, Trust-by-Construction Evaluation Paradigm MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.