Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T17:12:27.550904Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2411.12828.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T17:12:27.550904Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
24 of 24 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 75114879-3e1b-4b89-b7f7-eaa6cc6f5099 · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction online" 'onlinestring :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73434c25-2d58-4c6c-a3de-de822d4a6fce · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 620ae20c-b4ac-4889-8848-373f19591c09 · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction On the Measure of Intelligence
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b575b6ef-4675-4025-bea6-114e200c6a41 · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd32d7fe-6786-494d-9a52-773dd83e9a5c · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction MLAgentBench: Evaluating Language Agents on Machine Learning Experimentation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bccaa92b-f338-4d66-93e9-632f02e32831 · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d7438d53-4b2c-46fa-870c-ac653dcaf1b3 · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Evaluating Language-Model Agents on Realistic Autonomous Tasks
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a56b226-388d-45cf-bcbd-59b0b9336549 · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ba4cb79-1ea7-476d-9609-3b74bdbacfc0 · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction AgentBench: Evaluating LLMs as Agents
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b8ee843-340a-48f2-bb6f-cae1120dd06a · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction GAIA: a benchmark for General AI Assistants
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b73335bd-bd8d-49b6-b04a-8cc03d6a1e25 · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3aa019c1-0eec-4857-b712-115005caf43a · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a27ad8a0-b62a-4074-bb5f-577f43b96f2f · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 935984f7-7696-4ea4-9199-afc23baa82aa · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc3a9381-3c80-4aa0-8251-de8aec5cf89f · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Voyager: An Open-Ended Embodied Agent with Large Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28e004a6-1049-4d22-86c9-00aa6c1e1683 · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction MobileAgentBench: An Efficient and User-Friendly Benchmark for Mobile LLM Agents
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1369f742-cd63-4b19-bc59-5ad9cf5bbf7c · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction ScienceWorld: Is your Agent Smarter than a 5th Grader?
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85c4f3d1-6e8e-4392-8c47-c897e499d6fd · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a1212a9-3faa-4bcb-bde0-a3a9a6fc2689 · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction SmartPlay: A Benchmark for LLMs as Intelligent Agents
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd349079-c987-4354-ae06-32ed7d96a2a8 · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction The Rise and Potential of Large Language Model Based Agents: A Survey
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d22546ac-ecce-4062-ad32-11d7e5ab53b4 · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Do Large Language Models Latently Perform Multi-Hop Reasoning?
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94a06b83-d13f-4796-8956-11ec95cceb99 · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4165a0e-adfd-4a3e-8e87-4dd78ab8d0f2 · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction LlamaTouch: A Faithful and Scalable Testbed for Mobile UI Task Automation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec4f73e1-eef1-4e10-81a8-04a1d6061082 · outbound
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction WebArena: A Realistic Web Environment for Building Autonomous Agents
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.