Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-11T17:59:28.201166Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2607.05458.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-11T17:59:28.201166Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T10:10:33.975463Z
A source-named dated measurement, never combined with another source.
Source: cited_works
43 of 43 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0664bf45-a77b-44b0-ba9e-f7d28663e9d4 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning AgentBench: Evaluating
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d85fda4f-3fb2-4a33-a705-3b4376e299de · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8c08bd2-c78e-414a-95aa-8a6f9e3245ce · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning AFlow: Automating Agentic Workflow Generation
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd97c3a0-7275-4ffe-8e6e-69d66e708e9d · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Evolving Agents in the Dark: Retrospective Harness Optimization via Self-Preference
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db988c2e-77fd-4071-a898-22e0db6a142b · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79fd85ee-e8f0-4d04-9d82-25d92345f025 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7391420d-66ef-4463-99d9-4665a979af72 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7eb3cbbc-378a-42f3-9e15-2c484d1e2dc3 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning GAIA: a benchmark for General AI Assistants
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b30ec862-e2b3-4a2c-abb2-dcb3f2c35cf9 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning and Yang, John and Wettig, Alexander and Yao, Shunyu and Pei, Kexin and Press, Ofir and Narasimhan, Karthik , booktitle =
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dac36de7-4b10-44f4-8340-15af432625a9 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9dd8f26-f589-4edb-b453-8644e5853a87 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning 2024 , eprint=
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16311f42-fe58-4cca-b588-7264ef1e10bc · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 965b7abe-e470-4f3f-8c51-caaa2b0c4a0e · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Learning Representations (ICLR) , year =
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1dd99511-4e5b-4599-80e8-8e3b8c21116d · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Learning Representations (ICLR) , year =
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 579d0376-ad82-489b-8841-3056180fa992 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning and Moazam, Hanna and Miller, Heather and Zaharia, Matei and Potts, Christopher , journal =
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69b2e575-6ffb-4b0f-8c7d-6413bcb540d8 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Learning Representations , year=
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be06200b-c01b-4495-9329-a5a3241a9d1e · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 374d94e8-e629-4f0e-9d54-e2c0482f5b50 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fd18a8a-d4ea-4131-a9f3-42ddf78a99d3 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c84888e-4a4e-48b1-8e9b-a84521382fc5 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning and Tan, Shangyin and Soylu, Dilara and Ziems, Noah and Khare, Rishi and Opsahl-Ong, Krista and Singhvi, Arnav and Shandilya, Herumb and Ryan, Michael J
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2db7ddec-8599-40e7-ac24-c7caa78640a6 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Agentic Harness Engineering: Observability-Driven Automatic Evolution of Coding-Agent Harnesses
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7ccc6a2-8e8a-46d2-aab5-f391477ae2aa · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning 2026 , eprint=
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ab18ca7-1297-4e8f-854f-8d1349979ba1 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning 2023 , eprint=
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 263ab113-d5b8-4670-a433-a8eaf0c9c612 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e74298f1-c114-47f4-8ceb-5fb670473c2b · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Machine Learning (ICML) , year =
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85f8103a-5f56-47ae-9b28-2507652597d0 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Advances in Neural Information Processing Systems , volume=
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d24bdd2d-8fd5-4080-914e-0ce71f54ae7d · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07a8c811-b25a-4fb1-92cf-14ebfed10e9a · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d034320-5f84-4b5f-8d2f-670dbfff140c · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Learning Representations (ICLR) , year =
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2902dd34-2650-4591-be05-c170f2cd5e50 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Offline Reinforcement Learning with Implicit
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40c34489-5da4-4485-9e9f-a95d7b932e25 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning , booktitle =
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6714bb8d-e124-4347-85bc-8973152a9fa2 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Artificial Intelligence and Statistics (AISTATS) , year =
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e286e28-7c73-4dac-9ec3-482c70ed8a2e · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Learning Representations , year=
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68445256-30dd-40f5-84fb-ea640b39c02e · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Learning Representations , year=
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5864aca-8b9a-43d9-a6b2-c4df8ce89a47 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning 2024 , eprint=
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23cad747-53fd-4180-890e-d236b2445f05 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0487d67-c7c4-4579-9746-ab5413ffe7d2 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Process vs
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1764b52f-8495-41a5-b98f-04678a925e76 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning arXiv preprint arXiv:2510.25694 , year =
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 731e91b7-33f9-4833-ad9c-e5baf3fd2c9e · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b19d764-e7c5-46cf-9491-d5b0cc7e16e2 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Strained Coherence: A Pre-Failure Signal in Coding Agent Execution Trajectories
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 494390d9-45b7-4740-8962-c85c1d0279c5 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning and Salakhutdinov, Ruslan and Manning, Christopher D
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2541c560-0321-466b-b738-c607d3cc6b80 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dd3a443-0263-44d1-812a-a60676cb0b03 · outbound
Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Proceedings of the 16th International Conference on Machine Learning (ICML) , pages =
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b92da23f-7233-4967-8ac6-620726359d9d · inbound
Harness-R1: Learning to Edit Executable Runtime Harnesses from Agent Failure Trajectories Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.