Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2406.20015.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T10:12:48.107868Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-29T16:53:40.615942Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation a162cb5d-ab0e-473a-9ff9-4827dec9ee5e · inbound
Reducing Tool Hallucination via Reliability Alignment ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79e100d3-f173-4ecf-8cad-bceb313c8fa4 · inbound
When2Call: When (not) to Call Tools ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a802829-5b63-4fef-9498-339fe3536c9e · inbound
Beyond the Final Answer: Evaluating the Reasoning Trajectories of Tool-Augmented Agents ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 29b5b9a9-4885-482b-a389-0d078725eb55 · inbound
Beyond the Final Answer: Evaluating the Reasoning Trajectories of Tool-Augmented Agents ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c08a245-735d-493f-8bf9-cc39efbb1b4c · inbound
Your LLM Agents are Temporally Blind: The Misalignment Between Tool Use Decisions and Human Time Perception ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3baf782e-911a-4674-a71f-5abbd8aec7c6 · inbound
Model-Adaptive Tool Necessity Reveals the Knowing-Doing Gap in LLM Tool Use ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation a24ddf88-4a77-4ff5-babe-b378d297ab64 · inbound
Model-Adaptive Tool Necessity Reveals the Knowing-Doing Gap in LLM Tool Use ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 27a4c51f-cec4-441a-9980-a3141ab48d2a · inbound
Learning to Act under Noise: Enhancing Agent Robustness via Noisy Environments ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 39fad04c-9c0d-4f00-8d94-09b7a528e154 · inbound
SAAG: Structured Agent Assessment and Grounding ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.