Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2311.18760.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:02:29.265475Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T21:18:59.738705Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 06084d52-bd1f-4843-ba0b-22d00827e86e · inbound
What Limits Virtual Agent Application? OmniBench: A Scalable Multi-Dimensional Benchmark for Essential Virtual Agent Capabilities TaskBench: Benchmarking Large Language Models for Task Automation
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a54ff64-1f1d-4520-a5f0-b99007a8a92c · inbound
CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scenarios TaskBench: Benchmarking Large Language Models for Task Automation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab839f1b-689b-47f7-a359-3fff8177c7b7 · inbound
DICE-BENCH: Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues TaskBench: Benchmarking Large Language Models for Task Automation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ff91483-f9bd-427e-b52e-bc1f71953158 · inbound
Butterfly Effects in Toolchains: A Comprehensive Analysis of Failed Parameter Filling in LLM Tool-Agent Systems TaskBench: Benchmarking Large Language Models for Task Automation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fd32775-a877-4f2f-8338-4d5d8e800dc2 · inbound
Evaluation and Benchmarking of LLM Agents: A Survey TaskBench: Benchmarking Large Language Models for Task Automation
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fd0df3b-8120-4a59-9dc2-31261a1047b8 · inbound
ToolMATH: A Diagnostic Benchmark for Long-Horizon Tool Use under Systematic Tool-Catalog Constraints TaskBench: Benchmarking Large Language Models for Task Automation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0f88185b-6004-4a3b-9ec1-e81652714fd7 · inbound
From Intent to Execution: Composing Agentic Workflows with Agent Recommendation TaskBench: Benchmarking Large Language Models for Task Automation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e3d1ed72-7b21-47d1-989a-8ca379ac76df · inbound
A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications TaskBench: Benchmarking Large Language Models for Task Automation
Reference 131
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c36dc98e-b038-476c-a0da-8d5223a49814 · inbound
A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications TaskBench: Benchmarking Large Language Models for Task Automation
Reference 133
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f00eb9e9-5c6c-4fa8-840f-e52a854f5302 · inbound
A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications TaskBench: Benchmarking Large Language Models for Task Automation
Reference 125
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ba5974ba-b04f-4dc3-92ec-390f9993e4d4 · inbound
The Scaling Laws of Skills in LLM Agent Systems TaskBench: Benchmarking Large Language Models for Task Automation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5ca47bc0-0474-4ff5-a8e9-43e394bf85e4 · inbound
ArtifactLinker: Linking Scientific Artifacts for Automatic State-of-the-Art Discovery TaskBench: Benchmarking Large Language Models for Task Automation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0c073145-80da-41e9-b334-c9fe81fe015e · inbound
Tool-Call Dependency Structure is Linearly Decodable in LLM Agent Residual Streams TaskBench: Benchmarking Large Language Models for Task Automation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2f4cddb0-fea0-4544-bd36-1835165cd92e · inbound
Skill-Augmented AI Agents for Medical Research Analysis: An Exploratory Multi-Model Human Evaluation in an NSCLC Transcriptomic Biomarker Task TaskBench: Benchmarking Large Language Models for Task Automation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 416983a5-6e59-45f4-81fd-e3a9e03315dc · inbound
Compositional Skill Routing for LLM Agents: Decompose, Retrieve, and Compose TaskBench: Benchmarking Large Language Models for Task Automation
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0d6b121e-3485-4c8d-903d-5a29cc237600 · inbound
ReGRPO: Reflection-Augmented Policy Optimization for Tool-Using Agents TaskBench: Benchmarking Large Language Models for Task Automation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 741ecbfb-e0ae-454d-9095-65ef8cbf5832 · inbound
Identifying Implicit Bias in LLM-based Chat AI Toward People with Intellectual Disabilities TaskBench: Benchmarking Large Language Models for Task Automation
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.