Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2508.09124.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T05:13:41.995563Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-11T01:37:42.801023Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 65d346f4-7218-450b-a8d3-1038b655014f · inbound
Finch: Benchmarking Finance & Accounting across Spreadsheet-Centric Enterprise Workflows OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5072d6c0-9b4a-4c27-8feb-cd934da8508e · inbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c62324e5-15e4-4783-944e-7ea13ed7d6c7 · inbound
RetailBench: Evaluating Long-Horizon Autonomous Decision-Making and Strategy Stability of LLM Agents in Realistic Retail Environments OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24c38b31-9cb7-4baa-9c58-338a028a99f0 · inbound
Opal: Private Memory for Personal AI OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 239
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6b0e752a-c0bd-4e4b-ae6e-4ebeb9b0e139 · inbound
GTA-2: Benchmarking General Tool Agents from Atomic Tool-Use to Open-Ended Workflows OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1bea0ef8-cb73-4380-9fa3-c3f46052a504 · inbound
CUJBench: Benchmarking LLM-Agent on Cross-Modal Failure Diagnosis from Browser to Backend OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4fa2cec6-8f8b-4a61-b35c-6e1a2595bf5d · inbound
Tools as Continuous Flow for Evolving Agentic Reasoning OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4d5e23ff-a394-4099-9135-f654742b73de · inbound
WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f5a64f86-eeb9-4c68-ba41-61067e440367 · inbound
EnergyAgentBench: Benchmarking LLM Agents on Live Energy Infrastructure Data OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6e0420ee-eb3c-4c33-8f04-924fbe58b065 · inbound
EnergyAgentBench: Benchmarking LLM Agents on Live Energy Infrastructure Data OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d295bdd-2f3e-46ed-8708-8acc3d463185 · inbound
SentinelBench: A Benchmark for Long-Running Monitoring Agents OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2fb866c7-3830-4bb4-b3fb-a9ecb1dd1d5a · inbound
Benchmarking Open-Ended Multi-Agent Coordination in Language Agents OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a4eb0f97-5ceb-46e3-8957-5d90ae170888 · inbound
Running the Gauntlet: Re-evaluating the Capabilities of Agents Beyond Familiar Environments OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3ab79eeb-82cd-4ef1-80d5-1a74c4fe99e4 · inbound
CEO-Bench: Can Agents Play the Long Game? OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f8e0eeee-4f78-4a5a-94ab-21a3e66e5da5 · inbound
CEO-Bench: Can Agents Play the Long Game? OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad13d3d4-f532-4450-96d8-0ce87595a72a · inbound
ChainWorld: Composing Long-Horizon Desktop Workloads from Atomic OSWorld Tasks OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 561a7d5e-68da-4aaa-9250-13b69716b9d0 · inbound
How Do Tool-Augmented LLM Agents Perform on Real-World Energy Analytics Tasks? OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0adcedff-cf5d-4e65-ac83-237afa52fb7e · inbound
Office Comprehension Benchmark OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 662aa882-8fc7-40db-979c-6a923384bff0 · inbound
PolyWorkBench: Benchmarking Multilingual Long-Horizon LLM Agents OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3d6e37f4-3f48-4ee5-9232-60d5934d89f5 · inbound
PolyWorkBench: Benchmarking Multilingual Long-Horizon LLM Agents OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 35539d15-0a15-41c9-9662-f7ecf4bb764c · inbound
OmegaUse-OfficeVal: Benchmarking LLM Agents on Long-Horizon Office-Suite Tasks with Economic Grounding OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.