Pith. sign in

Paper Citation Record · LEDGER

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning

As of 11 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2607.05458.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.05458 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-11T17:59:28.201166Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T10:10:33.975463Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0664bf45-a77b-44b0-ba9e-f7d28663e9d4 · outbound

This paper cites AgentBench: Evaluating.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning AgentBench: Evaluating

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:40d33aa6530be8dfd547b418f916bc0e6cb1b422c00189c0c745cff7eadd0dd2

Observation d85fda4f-3fb2-4a33-a705-3b4376e299de · outbound

This paper cites SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:87a9fd0d1b28852693ac305ac01693230db71a4164f73e8aeccf950652662c4e

Observation f8c08bd2-c78e-414a-95aa-8a6f9e3245ce · outbound

This paper cites AFlow: Automating Agentic Workflow Generation.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning AFlow: Automating Agentic Workflow Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:1fa1e31c874376a444434e83fb11a54a22d2e790ceb793d714258e39578299fe

Observation cd97c3a0-7275-4ffe-8e6e-69d66e708e9d · outbound

This paper cites Evolving Agents in the Dark: Retrospective Harness Optimization via Self-Preference.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Evolving Agents in the Dark: Retrospective Harness Optimization via Self-Preference

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:6395f908e26b69e1f7b2d580ba8ff10bd03750a1b6cbeb2830948a7da889dce8

Observation db988c2e-77fd-4071-a898-22e0db6a142b · outbound

This paper cites an unresolved cited work.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:f423e94dcecc368adb6a0d02c3ca51b90c9a2a02d2697a9b37a09ca83f65a9a2

Observation 79fd85ee-e8f0-4d04-9d82-25d92345f025 · outbound

This paper cites an unresolved cited work.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:faf282fcda3f856928b03a35f475c2b4743d3c52abded6e6b24dbe4b54ad2b22

Observation 7391420d-66ef-4463-99d9-4665a979af72 · outbound

This paper cites an unresolved cited work.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:b4a83e8a85f2cfc08d6b0009436a46f73f9b24a6e2789eefe6ee81eb64b702fd

Observation 7eb3cbbc-378a-42f3-9e15-2c484d1e2dc3 · outbound

This paper cites GAIA: a benchmark for General AI Assistants.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning GAIA: a benchmark for General AI Assistants

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:60133e24c7ea0b695b79f15d7b0289b312e34f9131d5616ee4dd45325d1ba71a

Observation b30ec862-e2b3-4a2c-abb2-dcb3f2c35cf9 · outbound

This paper cites and Yang, John and Wettig, Alexander and Yao, Shunyu and Pei, Kexin and Press, Ofir and Narasimhan, Karthik , booktitle =.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning and Yang, John and Wettig, Alexander and Yao, Shunyu and Pei, Kexin and Press, Ofir and Narasimhan, Karthik , booktitle =

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:65117305de0b0ff3be640948d9f1026f1b706b02b7a121d7aae08025b9efc754

Observation dac36de7-4b10-44f4-8340-15af432625a9 · outbound

This paper cites an unresolved cited work.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:c991b818bbf6fde9ea39345f8c20dc03183a38a75a48b5c9311000d65d8853aa

Observation d9dd8f26-f589-4edb-b453-8644e5853a87 · outbound

This paper cites 2024 , eprint=.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning 2024 , eprint=

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:29e493017d373e8937af2a34d7285ffacd2ae4d9b47d8cd5314a5988ab7a169a

Observation 16311f42-fe58-4cca-b588-7264ef1e10bc · outbound

This paper cites an unresolved cited work.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:3156e83003ec3239e61e93f4ac7ecf9e1fa8aeceb515753ecb420bee46550127

Observation 965b7abe-e470-4f3f-8c51-caaa2b0c4a0e · outbound

This paper cites International Conference on Learning Representations (ICLR) , year =.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Learning Representations (ICLR) , year =

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:6aa3ffbaa7caf140e47a77c71e5f3524fa5098f58d5ac24b7f0dac307a252428

Observation 1dd99511-4e5b-4599-80e8-8e3b8c21116d · outbound

This paper cites International Conference on Learning Representations (ICLR) , year =.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Learning Representations (ICLR) , year =

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:f7518006e41b9608952939ca04a972fe6c0e25e09f5fb8057531452ec7c1445a

Observation 579d0376-ad82-489b-8841-3056180fa992 · outbound

This paper cites and Moazam, Hanna and Miller, Heather and Zaharia, Matei and Potts, Christopher , journal =.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning and Moazam, Hanna and Miller, Heather and Zaharia, Matei and Potts, Christopher , journal =

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:a15aac8b227cec6666340ef89b375ae5c297db1bdd0fc38dbfb9075c624a6255

Observation 69b2e575-6ffb-4b0f-8c7d-6413bcb540d8 · outbound

This paper cites International Conference on Learning Representations , year=.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Learning Representations , year=

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:ecfe26d68b3e54f5b31c001e4f78243ccbf363bfc5a96db0aee50fc68dd6be3d

Observation be06200b-c01b-4495-9329-a5a3241a9d1e · outbound

This paper cites an unresolved cited work.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:4f6523a59645152a681fd91f7cd08a7df4b83c3c24a35d9819124737313cc552

Observation 374d94e8-e629-4f0e-9d54-e2c0482f5b50 · outbound

This paper cites an unresolved cited work.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:c0e22215b524d79b914143caee7b2b07ec24bcbc0e7a0c559c377e849e0d054f

Observation 6fd18a8a-d4ea-4131-a9f3-42ddf78a99d3 · outbound

This paper cites an unresolved cited work.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:5bb90e5b09d81b3646dd477c256c8399e8972a1821f029fd6dd46e66672a19b6

Observation 3c84888e-4a4e-48b1-8e9b-a84521382fc5 · outbound

This paper cites and Tan, Shangyin and Soylu, Dilara and Ziems, Noah and Khare, Rishi and Opsahl-Ong, Krista and Singhvi, Arnav and Shandilya, Herumb and Ryan, Michael J.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning and Tan, Shangyin and Soylu, Dilara and Ziems, Noah and Khare, Rishi and Opsahl-Ong, Krista and Singhvi, Arnav and Shandilya, Herumb and Ryan, Michael J

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:eca4ccb4114ca0527eac4ad58860417cfdccd7731990ee68875276b5beccb7bf

Observation 2db7ddec-8599-40e7-ac24-c7caa78640a6 · outbound

This paper cites Agentic Harness Engineering: Observability-Driven Automatic Evolution of Coding-Agent Harnesses.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Agentic Harness Engineering: Observability-Driven Automatic Evolution of Coding-Agent Harnesses

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:73cccc23e74785ede5e31f681bf37a3e61262bcf651d038a896fa4736a28a384

Observation b7ccc6a2-8e8a-46d2-aab5-f391477ae2aa · outbound

This paper cites 2026 , eprint=.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning 2026 , eprint=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:87a1215aa50e3ba674a89bd468856bb7a15f9f6e6bbc44ca8539dd6245f7edb4

Observation 7ab18ca7-1297-4e8f-854f-8d1349979ba1 · outbound

This paper cites 2023 , eprint=.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning 2023 , eprint=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:c46f654abe34c88d62552b1e304a94cf1d37069533e4f829c0ad4d90d5cb801d

Observation 263ab113-d5b8-4670-a433-a8eaf0c9c612 · outbound

This paper cites an unresolved cited work.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:6549190e5158371911f1025bd68702168b3d3933a5d34858bbd5621ade63824e

Observation e74298f1-c114-47f4-8ceb-5fb670473c2b · outbound

This paper cites International Conference on Machine Learning (ICML) , year =.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Machine Learning (ICML) , year =

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:de26b0f650da53b1c53bc64b809ad1dfcb3b5c56179c2767620850e01fa37b79

Observation 85f8103a-5f56-47ae-9b28-2507652597d0 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Advances in Neural Information Processing Systems , volume=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:04853778399674c11c5ce93de26f40884b4d920f8aaa2b98d459c01a6fa679c8

Observation d24bdd2d-8fd5-4080-914e-0ce71f54ae7d · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:50ab143462d59ed7c3b85fe0bb6a545c798f30fa06878c12ec90c40b493d562b

Observation 07a8c811-b25a-4fb1-92cf-14ebfed10e9a · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:c2c736308c94907d8a920909c771b154e9a5c8dada4c81c36582fe23443e57c9

Observation 5d034320-5f84-4b5f-8d2f-670dbfff140c · outbound

This paper cites International Conference on Learning Representations (ICLR) , year =.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Learning Representations (ICLR) , year =

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:2bc5866e40165d3ea225ddd094c401b11df49bc5418ff7bd589936a7f86c1c0c

Observation 2902dd34-2650-4591-be05-c170f2cd5e50 · outbound

This paper cites Offline Reinforcement Learning with Implicit.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Offline Reinforcement Learning with Implicit

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:adf86b2d6ff5d4eb02ed5a23b1dd01146a389b0b07663edbeaec0e7357b233f3

Observation 40c34489-5da4-4485-9e9f-a95d7b932e25 · outbound

This paper cites , booktitle =.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning , booktitle =

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:d4ac0230acf9c3d539fdd0ba93f640f7cf409490b80034e1a44365c08fdaa42e

Observation 6714bb8d-e124-4347-85bc-8973152a9fa2 · outbound

This paper cites International Conference on Artificial Intelligence and Statistics (AISTATS) , year =.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Artificial Intelligence and Statistics (AISTATS) , year =

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:d91c583827e0d26bb1e94cbb0d08f96b9e9461d8a2f76cd5c9c81e6ddf214502

Observation 0e286e28-7c73-4dac-9ec3-482c70ed8a2e · outbound

This paper cites International Conference on Learning Representations , year=.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Learning Representations , year=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:93dfb750b83500baaf988bcb9854e1de8c2c2f8ec38fa7488d196767f94fc41e

Observation 68445256-30dd-40f5-84fb-ea640b39c02e · outbound

This paper cites International Conference on Learning Representations , year=.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning International Conference on Learning Representations , year=

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:43ec6e5fd301f0c9e949d95b747b25c9afabe3280fd7f031b9eebb3c5f5c5e20

Observation d5864aca-8b9a-43d9-a6b2-c4df8ce89a47 · outbound

This paper cites 2024 , eprint=.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning 2024 , eprint=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:37914d977bbe5be4144663fb56b6bbb5023cee9b1a25f1a04cdb3949b2a4cd8a

Observation 23cad747-53fd-4180-890e-d236b2445f05 · outbound

This paper cites an unresolved cited work.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:4e623779f20c1af8da12f667f8f410b55f0cb206bd549713d6eaf08994b013dc

Observation d0487d67-c7c4-4579-9746-ab5413ffe7d2 · outbound

This paper cites Process vs.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Process vs

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:eea8c300a652b714444c07a7e99ddbe29d3ceb038de1fe7d2e458db1ca86f482

Observation 1764b52f-8495-41a5-b98f-04678a925e76 · outbound

This paper cites arXiv preprint arXiv:2510.25694 , year =.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning arXiv preprint arXiv:2510.25694 , year =

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:c63c663ff21b22624a87014c63d2b6a212c1081e231da317250922a1ea23804d

Observation 731e91b7-33f9-4833-ad9c-e5baf3fd2c9e · outbound

This paper cites an unresolved cited work.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:baddeb1fcbc4e2f45ff82fe2202f5c8d8a921bda0628c8638f0b875d677e59fc

Observation 4b19d764-e7c5-46cf-9491-d5b0cc7e16e2 · outbound

This paper cites Strained Coherence: A Pre-Failure Signal in Coding Agent Execution Trajectories.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Strained Coherence: A Pre-Failure Signal in Coding Agent Execution Trajectories

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:7413d9e5ac8015aed8b1ccb5990dd163ed2f43706a059803f539f35f6e51153a

Observation 494390d9-45b7-4740-8962-c85c1d0279c5 · outbound

This paper cites and Salakhutdinov, Ruslan and Manning, Christopher D.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning and Salakhutdinov, Ruslan and Manning, Christopher D

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:eb47d9959ec047c783e74dab197724b69bdaaf14134baf1b4f9c3abae69fd9bb

Observation 2541c560-0321-466b-b738-c607d3cc6b80 · outbound

This paper cites an unresolved cited work.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:0796ce5cb4f21aa27d351f8b975346d9fc53a2a382f93524ce019523fb5e8398

Observation 3dd3a443-0263-44d1-812a-a60676cb0b03 · outbound

This paper cites Proceedings of the 16th International Conference on Machine Learning (ICML) , pages =.

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning Proceedings of the 16th International Conference on Machine Learning (ICML) , pages =

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-11T17:59:28.201166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T17:59:28.201166Z digest=sha256:0e50c452f0ada4f4ae9441ec14207a3163ff5a33166488cc7432ce8df6c927b4

Pith citing papers

Observation b92da23f-7233-4967-8ac6-620726359d9d · inbound

Harness-R1: Learning to Edit Executable Runtime Harnesses from Agent Failure Trajectories cites this paper.

Harness-R1: Learning to Edit Executable Runtime Harnesses from Agent Failure Trajectories Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T10:10:33.975463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:10:33.975463Z digest=sha256:eb46564971867932350559774cb9e433b1f9dd1ee46ed4d87b17c7ea7355b085