Pith. sign in

Paper Citation Record · LEDGER

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction

As of 23 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2411.12828.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.12828 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T17:12:27.550904Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 75114879-3e1b-4b89-b7f7-eaa6cc6f5099 · outbound

This paper cites online" 'onlinestring :=.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.441501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.441501Z digest=sha256:fff446839d2848dee4e9942466cbd2c15ac8892c014fd4d310c597f0eeb0c5e4

Observation 73434c25-2d58-4c6c-a3de-de822d4a6fce · outbound

This paper cites write newline.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.448492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.448492Z digest=sha256:d5b8b2173f85cc08a80757f79e40416cf8bc4f3922953722a5030bd3d408015f

Observation 620ae20c-b4ac-4889-8848-373f19591c09 · outbound

This paper cites On the Measure of Intelligence.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction On the Measure of Intelligence

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.456162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.456162Z digest=sha256:af3e7bbb1ab353ac6381b372b8e5623aae4c2d2a93ffd0c03a1ac5334cfe5272

Observation b575b6ef-4675-4025-bea6-114e200c6a41 · outbound

This paper cites Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.462350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.462350Z digest=sha256:4cd72ddc1c8ee5057b60741db38e636dd9bcb3b09198dadecd153f996349a3ca

Observation dd32d7fe-6786-494d-9a52-773dd83e9a5c · outbound

This paper cites MLAgentBench: Evaluating Language Agents on Machine Learning Experimentation.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction MLAgentBench: Evaluating Language Agents on Machine Learning Experimentation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.467241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.467241Z digest=sha256:fcc9b8ef43be1c0834a9a41ca36f7ac250dde4a3b6d9185dd3aee0a3f995100a

Observation bccaa92b-f338-4d66-93e9-632f02e32831 · outbound

This paper cites an unresolved cited work.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:12:27.873208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T17:12:27.475054Z digest=sha256:e60016c6c589212dc64ab93e0ce0d0894a1780cdf6c98f6b9728834f3f43dead

Observation d7438d53-4b2c-46fa-870c-ac653dcaf1b3 · outbound

This paper cites Evaluating Language-Model Agents on Realistic Autonomous Tasks.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Evaluating Language-Model Agents on Realistic Autonomous Tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.480564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.480564Z digest=sha256:8d27c5890776bf896b22180f82af6882f26e282eab14ae4d374ed97bee3e54e5

Observation 0a56b226-388d-45cf-bcbd-59b0b9336549 · outbound

This paper cites an unresolved cited work.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.485332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.485332Z digest=sha256:a893f3579ca6a29a44cc0a128e884705cd2c477bbb38a7aa352bafe6fa428cf0

Observation 9ba4cb79-1ea7-476d-9609-3b74bdbacfc0 · outbound

This paper cites AgentBench: Evaluating LLMs as Agents.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction AgentBench: Evaluating LLMs as Agents

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.489930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.489930Z digest=sha256:0b82a7238b9f9d845d18a4afa7ab21659a7d6c68b21a64609b2ce1db3614cbce

Observation 6b8ee843-340a-48f2-bb6f-cae1120dd06a · outbound

This paper cites GAIA: a benchmark for General AI Assistants.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction GAIA: a benchmark for General AI Assistants

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.495003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.495003Z digest=sha256:daaba51e849b59a9c6f138aada5b12f702b623c066785636929965475ef415b3

Observation b73335bd-bd8d-49b6-b04a-8cc03d6a1e25 · outbound

This paper cites an unresolved cited work.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.500025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.500025Z digest=sha256:e7908947a882cbde060a671fa38f1f1b194bcd1dbf63a6365a35098f96ce5e61

Observation 3aa019c1-0eec-4857-b712-115005caf43a · outbound

This paper cites an unresolved cited work.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:12:27.847999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T17:12:27.504414Z digest=sha256:a6efc066444e3d7315238a05d71651ac6f9212ab48038bcfa196736cc16da50a

Observation a27ad8a0-b62a-4074-bb5f-577f43b96f2f · outbound

This paper cites Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.508094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.508094Z digest=sha256:b24e21c0f70fb177a3501ece9c1dd690e4dd6ee659b874c7563f43e65e69e777

Observation 935984f7-7696-4ea4-9199-afc23baa82aa · outbound

This paper cites NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.512222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.512222Z digest=sha256:c87c6b308691b6b1d49256d57aa0627653ba06d3ac4235d91ce1795b7d33104f

Observation cc3a9381-3c80-4aa0-8251-de8aec5cf89f · outbound

This paper cites Voyager: An Open-Ended Embodied Agent with Large Language Models.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Voyager: An Open-Ended Embodied Agent with Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.516108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.516108Z digest=sha256:ba36cc702ff1e51820be537f3b9e10b07bed8b63baa60b91a13186a17a804e2f

Observation 28e004a6-1049-4d22-86c9-00aa6c1e1683 · outbound

This paper cites MobileAgentBench: An Efficient and User-Friendly Benchmark for Mobile LLM Agents.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction MobileAgentBench: An Efficient and User-Friendly Benchmark for Mobile LLM Agents

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.519744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.519744Z digest=sha256:a15916956e0dfa441bdf601259974bf02c7fca3f1c50e7c2880bc346fe54dbd9

Observation 1369f742-cd63-4b19-bc59-5ad9cf5bbf7c · outbound

This paper cites ScienceWorld: Is your Agent Smarter than a 5th Grader?.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction ScienceWorld: Is your Agent Smarter than a 5th Grader?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.523178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.523178Z digest=sha256:5dd85fa9f986ebc56a068f78e03eda25c5727f43839aff03801b43f502d9df01

Observation 85c4f3d1-6e8e-4392-8c47-c897e499d6fd · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.527467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.527467Z digest=sha256:b59be3e6d99438fab682d50fd3805318fa1f69dadec838595583a83e26d4c223

Observation 9a1212a9-3faa-4bcb-bde0-a3a9a6fc2689 · outbound

This paper cites SmartPlay: A Benchmark for LLMs as Intelligent Agents.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction SmartPlay: A Benchmark for LLMs as Intelligent Agents

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.531912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.531912Z digest=sha256:fa0bd30695bb588829a27ae36c870bbd0242e27097170280266cc5b07b95e7fb

Observation fd349079-c987-4354-ae06-32ed7d96a2a8 · outbound

This paper cites The Rise and Potential of Large Language Model Based Agents: A Survey.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction The Rise and Potential of Large Language Model Based Agents: A Survey

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.535964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.535964Z digest=sha256:19dac44fbd485a4fe4581791e1651b117fd238c2b38d44cdf203dc1214daa3cb

Observation d22546ac-ecce-4062-ad32-11d7e5ab53b4 · outbound

This paper cites Do Large Language Models Latently Perform Multi-Hop Reasoning?.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction Do Large Language Models Latently Perform Multi-Hop Reasoning?

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.539468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.539468Z digest=sha256:d702cc3923f3a87cbee6041db8d90970f5c2a4c1f0cd1fe3bbb8ae50e7c64bb5

Observation 94a06b83-d13f-4796-8956-11ec95cceb99 · outbound

This paper cites WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.543427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.543427Z digest=sha256:c04955f01cd90ec5d8adbee5330f397567badff2646ddd648fca1a802c42c35d

Observation e4165a0e-adfd-4a3e-8e87-4dd78ab8d0f2 · outbound

This paper cites LlamaTouch: A Faithful and Scalable Testbed for Mobile UI Task Automation.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction LlamaTouch: A Faithful and Scalable Testbed for Mobile UI Task Automation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.547029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.547029Z digest=sha256:2c193b33a3de36a3ff96e76af17521143e7b3d38d0c35662ec2d6c1a169b44b5

Observation ec4f73e1-eef1-4e10-81a8-04a1d6061082 · outbound

This paper cites WebArena: A Realistic Web Environment for Building Autonomous Agents.

Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction WebArena: A Realistic Web Environment for Building Autonomous Agents

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T17:12:27.550904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:12:27.550904Z digest=sha256:5e78a5745fa35105f6e4b0a559f62e9f3eb5dded6bcd2ba7093aae65db5779b8

Pith citing papers

No inbound Pith citation observations are available.