Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2312.14033.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:17:46.338385Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T17:18:43.767493Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 44993b3d-a8d2-4b37-848d-8145c10d175b · inbound
InternLM2 Technical Report T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
Reference 193
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f8ecf101-dcae-4aee-8ff4-887f44a8337a · inbound
A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6512ffbd-b7dd-481e-88e9-49d88c10eb49 · inbound
DICE-BENCH: Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df1edc00-06ac-4a00-ba81-7c9cc172fbae · inbound
Teaching a Language Model to Speak the Language of Tools T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e8428c8-01d7-4f9a-880f-66f929af1677 · inbound
A Survey of Context Engineering for Large Language Models T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
Reference 161
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9c1b31bb-935b-4b46-8506-b73c01c8e05b · inbound
A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation af3d1478-45bd-4d65-b7e9-7e7899f8da05 · inbound
Evaluation and Benchmarking of LLM Agents: A Survey T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0068ad15-fda3-4d1b-a81c-c670d82594d5 · inbound
Don't Start What You Can't Finish: A Counterfactual Audit of Support-State Triage in LLM Agents T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 144effff-7c90-45c6-bbfd-91d7e78da2a5 · inbound
Consistency as a Testable Property: Statistical Methods to Evaluate AI Agent Reliability T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 69380883-da7f-4587-9e82-7b4c290fbf28 · inbound
Holistic Evaluation and Failure Diagnosis of AI Agents T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a56ed124-4fff-42de-b142-8e02ced8ff85 · inbound
Synthesize and Reward -- Reinforcement Learning for Multi-Step Tool Use in Live Environments T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 326e0d88-bee7-4345-99f9-7c3999f2fdd0 · inbound
Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
Reference 227
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7b596bd0-182b-46a4-af88-d51e1343a432 · inbound
Qwen-Audio-VAE Technical Report T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
Reference 234
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee8046a5-274a-4651-92ce-9272532307be · inbound
WorkSurface-Bench: Benchmarking Enterprise Agents on Multi-Surface Knowledge Routing T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.