Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T14:25:08.052988Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 100 of 108 outbound references and 2 inbound Pith citation observations for arXiv:2606.02373.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T14:25:08.052988Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T13:01:09.120785Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-08T13:01:09.639756Z
100 of 108 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation faaf47b4-b9a7-456b-b0f9-5d2237fb7076 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Retrieval-augmented generation for knowledge-intensive NLP tasks , year =
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6fef523-bb81-4122-b6e4-dacaa672b1ae · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2026 , eprint=
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14b4ee63-034b-4aa6-bafb-8197b5d324a1 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2023 , eprint=
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9aee9ca1-3c41-4882-98ac-80bf534ecd5a · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses , title =
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd96cb76-b4b9-49d1-8de5-1f06bede42e4 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , eprint=
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1b5044e-4eab-4864-82a2-ea1396a654ab · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , eprint=
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc67782a-c8ec-4908-be80-f32375ad1f84 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , eprint=
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 974089d0-cc3a-4108-a6a4-a526bed7e476 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , eprint=
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6a8aa56-5bdd-416d-af94-b7d50da49d64 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing , pages=
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07eb4697-9bee-4728-ae3a-b2ecbca802da · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2026 , eprint=
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea5da7b7-c14f-4925-a054-637f2e2ef79f · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , eprint=
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab5fc597-2609-48ad-b699-9ebbe832778a · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , eprint=
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82170d76-4bd4-400a-8196-4d18aab4bf03 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , eprint=
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fb828bd-bb3b-46f4-84b2-baa233ab37cf · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Proceedings of the ACM on Software Engineering , volume=
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a64c220c-6bfd-4a3c-ab2f-9fccc255d497 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2023 , eprint=
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73cdefd5-7bbe-4233-94f3-e02c544fd523 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , eprint=
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90ef5f55-4865-4e10-8604-3c62ecaea3e0 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses DeepResearcher: Scaling deep research via reinforcement learning in real-world environments
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9d0d9951-7e2a-42e7-b7ef-ee02dd81484e · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses ToolRL: Reward is All Tool Learning Needs
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 44ca8b46-9ffa-474a-b182-31fccb01ca3b · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses doi: 10.1038/s41586-025-09422-z
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1672d9ce-d0d3-483b-a9f9-955dc422ce67 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses and Finn, Chelsea , title =
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 820128cc-a57b-4dac-941f-4315aa191981 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2017 , eprint=
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a03fd3c8-bc55-4ce4-a325-74dc08a96354 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , eprint=
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ad7da53-1399-4c48-9a72-d5d2eba40d9e · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , eprint=
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 680c1318-8bbf-47c9-bb60-5a00f6c1da6f · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Search-o1: Agentic search-enhanced large reasoning models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0f0272b9-6774-4a6c-8fd8-a1dd59a00dc0 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2026 , eprint=
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba465f0e-8055-4ef4-95fa-ae9c92a2f338 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , eprint=
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9c2c0d7-1231-4f3e-9b97-a748f8cea014 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2026 , publisher=
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 598e6057-9c12-4ca5-bd50-7728a2ccdd0f · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2026 , month = apr, howpublished =
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ac9816d-70f7-4bcd-994c-8077152f2065 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses OpenAI engineering note , year=
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de7eb09f-15ed-4891-b4f9-7195aa71aeae · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses and Lewis, Mike , editor =
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 04a5b79e-9644-428b-9ddb-070c8c94db61 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Proceedings of the 61st annual meeting of the association for computational linguistics (volume 1: long papers) , pages=
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50acf5a5-e59a-4ca3-9d38-41133a0e7dfb · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Proceedings of the 2023 conference on empirical methods in natural language processing , pages=
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 202dd11b-51f0-4819-90df-49c442df819c · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses s3: You Don ' t Need That Much Data to Train a Search Agent via RL
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 18aadace-7964-4ecb-83f8-e8e932b99e46 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , eprint=
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1ec520d-3f9e-40cc-9758-a9029499aff1 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Unresolved cited work
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 376456ff-6727-4c61-8ec0-68d55ee6e2fa · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2026 , month =
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10afc3e3-81cd-47fd-b2b4-ae52cb6b0efa · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , eprint=
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9c50e40-787e-4598-bf95-40b8ab12aea8 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2024 , eprint=
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 393e8439-1d66-492c-b3bc-8afdbbf1944c · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2026 , eprint=
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95cf4839-87eb-4455-a5ed-e5453ef37209 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Recursive Language Models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3c57019a-e1ad-473e-b727-ae55b8191216 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Transactions of the association for computational linguistics , volume=
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 044f3f83-a83f-49c5-af15-225530819512 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , month =
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51fb6772-1f23-463a-ba0e-a6f09e28480a · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Advances in neural information processing systems , volume=
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 696e710a-a10b-4ac9-b5b6-3c8120ee0aae · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , eprint=
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7a2d72d-6cb4-4c78-be6c-c37f8e103e17 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Advances in Neural Information Processing Systems , volume=
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c7a007d-afce-4250-aec2-1ff09fe83036 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ff8f9599-756d-443a-a876-228636be9ae4 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2025 , eprint=
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea2e96af-9f8b-43c3-b39e-15f55e582cc6 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Unresolved cited work
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95dda421-5533-4bc5-8db6-20a3444b6c9f · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Proceedings of the 2018 conference on empirical methods in natural language processing , pages=
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7824055c-6fb7-4d32-9f88-6e4283fff520 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Humanity's Last Exam
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 49ada973-5248-441c-8a8d-5fe0a1bb938a · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses 2026 , eprint=
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 261b2d3f-9db6-4989-8988-fd73daeca72a · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing , pages=
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53252afc-447f-46f4-9a74-b59d115fec7b · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Introducing swe-grep and swe-grep-mini: Rl for multi-turn, fast context retrieval
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e90eb068-5c46-497d-b2e3-bb3f8e597f8a · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Chroma context-1: Training a self-editing search agent
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2dbb7668-e0ee-4cf6-b1c9-42c231482f3b · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f5bdb31a-6718-4b47-96b0-31b22de33cf4 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Apex-searcher: Augmenting llms' search capabilities through agentic planning and execution, 2026
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e19d15c-320b-4d9f-b8b3-4b7fe1a27b37 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Pan, Wen Zhang, Huajun Chen, Fan Yang, Zenan Zhou, and Weipeng Chen
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb28519c-26e8-4992-b9c1-a6e4e1848bab · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Browsecomp-plus: A more fair and transparent evaluation benchmark of deep-research agent, 2025
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d91d287-ce9d-4c80-b54e-9265ada45de7 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Atom-searcher: Enhancing agentic deep research via fine-grained atomic thought reward, 2025
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccb3a424-6fa6-46ec-8606-1c2dd4ea0192 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Beyond ten turns: Unlocking long-horizon agentic search with large-scale asynchronous rl, 2025
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ef8a40e-e995-4aa0-a92a-f22e92635786 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Unresolved cited work
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90bb7c06-5fea-4142-8148-131658aab4d2 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Michael Alvarez
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5088f26f-45b4-4378-b598-62e47ab24c1b · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Steer2adapt: Dynamically composing steering vectors elicits efficient adaptation of llms, 2026
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 606aea60-8241-4a48-8d45-c641cc76035d · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Context rot: How increasing input tokens impacts llm performance
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fca8cb7c-d90c-4a6b-b03a-6764e04680d5 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Reinforce++: Stabilizing critic-free policy optimization with global advantage normalization, 2025
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dc89a95-3121-49e1-9d73-4dd2d5ec0194 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses DeepRetrieval: Hacking Real Search Engines and Retrievers with Large Language Models via Reinforcement Learning
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0ce5068c-c9d8-4d70-aecf-eb16f5a1b9ba · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Adaptation of agentic ai: A survey of post-training, memory, and skills, 2026
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65bba525-7322-40b2-98d9-b0ca6abdb92f · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses s3: You don ' t need that much data to train a search agent via RL
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0117e6c8-9bdb-4257-aa37-27501e422d23 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Active retrieval augmented generation
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05f1bd12-fbc7-4f75-a5a9-78d3ff1bff30 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 07dfba9e-7213-48a9-85fb-1f6ec103d405 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Dhillon, David Brandfonbrener, and Rishabh Agarwal
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c451c1d2-92f8-44a3-89dc-540e6a51b254 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Fact, fetch, and reason: A unified evaluation of retrieval-augmented generation
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e73a119-ce43-4baf-a078-2749bc1b9f06 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Bowman, and Ethan Perez
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a33fcb3-8902-4d72-a2d9-8da7cde33649 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Meta-Harness: End-to-End Optimization of Model Harnesses
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 75c7b63f-e457-4902-9550-33718825213e · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Websailor: Navigating super-human reasoning for web agent, 2025
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1261f824-9203-4a7d-9928-7df4b56c3349 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Search-o1: Agentic search-enhanced large reasoning models
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3812861e-4d54-4462-8061-d19b4477cc3a · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Webexplorer: Explore and evolve for training long-horizon web agents, 2025
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c9dbb0f-e733-4fce-84e7-b519faf08016 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Harness engineering: leveraging codex in an agent-first world, 2026
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 146124d1-3134-4816-af85-01bcbf1184cf · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Unresolved cited work
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 614dafa1-23f0-4716-ad27-4b46d60cccc1 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Scaling managed agents: Decoupling the brain from the hands
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7329441-923c-4052-91b1-38c968fcb6fc · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Unresolved cited work
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dec39293-a115-4666-a911-4ac85505732e · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Patil, Ion Stoica, and Joseph E
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10b7a040-8f93-47e6-b853-2e4fe98f5f78 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses SealQA: Raising the Bar for Reasoning in Search-Augmented Language Models
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 62e1da60-e49f-4a7c-80c4-3faf73bcbd48 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Measuring and narrowing the compositionality gap in language models
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 462c1a5d-011d-4306-a9cc-a73b8c306257 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Toolllm: Facilitating large language models to master 16000+ real-world apis, 2023
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06f2951b-f1c4-412e-a596-aeec62a96566 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Proximal policy optimization algorithms, 2017
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d122814-3995-47d7-b56d-bc40d2235f33 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses HARBOR: Automated Harness Optimization
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9d062346-98d9-4a07-8446-fe46801c7d1d · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning
Reference 101
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a140e10e-b646-4609-89d6-70343a74e450 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Toolorchestra: Elevating intelligence via efficient model and tool orchestration, 2025
Reference 102
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1877b91a-498e-476b-9c4f-73a458cbcb14 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Decouplesearch: Decouple planning and search via hierarchical reward modeling
Reference 103
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51bf1f54-a9ce-474a-bb9f-1536d6341df6 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Tongyi DeepResearch Technical Report
Reference 104
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a4672d20-21c6-4a52-b30c-45c481ac8d7a · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Interleaving retrieval with chain-of-thought reasoning for knowledge-intensive multi-step questions
Reference 105
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3652ee58-b3fb-4925-bd75-fa1edcb1cf17 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Unresolved cited work
Reference 106
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45a5daac-0171-4fc5-bf89-01c6816d5d5d · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses USPTO Open Data Portal
Reference 107
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73c2e6de-5fcd-4e89-99de-8efdfbd566e7 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Securities and Exchange Commission
Reference 108
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9f02781-c1f8-4734-b987-c1863f1b9afa · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution
Reference 109
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 80832a9b-e435-4a00-afdf-f06b668aa714 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses StepSearch: Igniting LLMs Search Ability via Step-Wise Proximal Policy Optimization
Reference 110
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 25ad6795-52da-49e7-9bb6-490bee8d30b3 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Webdancer: Towards autonomous information seeking agency, 2025
Reference 111
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64a7d282-92b5-4bde-bcfa-50f2ff44876d · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Agentless: Demystifying LLM-based Software Engineering Agents
Reference 112
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4d548640-d3db-4ad3-87f0-14703cf2f595 · outbound
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Demystifying llm-based software engineering agents
Reference 113
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6ef2b82-acfa-47ce-804d-a6049effec00 · inbound
Fetch-then-Explore: Decoupling Selection from Extraction over a Persistent Workspace for Search Agents Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 431a4610-419f-4600-8b9e-a9e9f7df540a · inbound
EvoHarness-RL: Learning Self-Evolving Runtime Harness for Long-Horizon LLM Agents Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.