Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:40:59.083259Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 17 inbound Pith citation observations for arXiv:2506.01716.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:40:59.083259Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:34:45.225752Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T20:40:08.275113Z
63 of 63 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 44438846-9183-4610-90f0-7e593061c643 · outbound
Self-Challenging Language Model Agents Digirl: Training in-the-wild device-control agents with autonomous reinforcement learning,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4fcd1709-dd65-44fe-9aaf-d2c5432967d9 · outbound
Self-Challenging Language Model Agents Digi-q: Learning q-value functions for training device-control agents, 2025
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bbfc647f-825b-4944-83b2-87e8202bbfb5 · outbound
Self-Challenging Language Model Agents Active learning of inverse models with intrinsically motivated goal exploration in robots.Robotics and Autonomous Systems, 61(1):49–73, January
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 46d9604d-cec9-4ee1-a125-379e12100f72 · outbound
Self-Challenging Language Model Agents Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55271e5d-fbaa-47c6-af00-754928eef049 · outbound
Self-Challenging Language Model Agents Augmenting Autotelic Agents with Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e308654-6358-4618-b627-6424c6b70c04 · outbound
Self-Challenging Language Model Agents STP: Self-play LLM Theorem Provers with Iterative Conjecturing and Proving
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2e57cc7-d1e9-4b17-8f2a-79768333aacb · outbound
Self-Challenging Language Model Agents OMNI-EPIC: Open-endedness via Models of human Notions of Interestingness with Environments Programmed in Code
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84c72a2d-a0d9-4de3-813c-6a0277e8d042 · outbound
Self-Challenging Language Model Agents StableToolBench: Towards Stable Large-Scale Benchmarking on Tool Learning of Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4140923c-5a01-4a12-8a70-8c059f3d40e8 · outbound
Self-Challenging Language Model Agents WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55458e07-f4f4-41b1-98ed-6a97190d5b90 · outbound
Self-Challenging Language Model Agents OpenWebVoyager: Building Multimodal Web Agents via Iterative Real-World Exploration, Feedback and Optimization
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7075c19d-3928-4696-92d9-216f700d1d63 · outbound
Self-Challenging Language Model Agents AgentGen: Enhancing Planning Abilities for Large Language Model based Agent via Environment and Task Generation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5012d17-4c79-439e-902d-1f3c5defdb41 · outbound
Self-Challenging Language Model Agents Language Models as Zero-Shot Planners: Extracting Actionable Knowledge for Embodied Agents
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 410f2c01-01f4-4f77-a7ff-0e1c58e280a3 · outbound
Self-Challenging Language Model Agents VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0edaf56-ac88-469d-ab56-cbf610a6870a · outbound
Self-Challenging Language Model Agents SWE-bench: Can Language Models Resolve Real-World GitHub Issues?
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b43d03bc-307d-44df-902d-a0d13e9eb2c3 · outbound
Self-Challenging Language Model Agents Code as Policies: Language Model Programs for Embodied Control
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab90699e-f496-42b9-bff5-e614f6818198 · outbound
Self-Challenging Language Model Agents AgentBench: Evaluating LLMs as Agents
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8efe4362-c272-49f6-b02d-572875a7f6a5 · outbound
Self-Challenging Language Model Agents Toolverifier: Generalization to new tools via self-verification,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8509d31a-19ea-4514-b398-4b1ab50c565b · outbound
Self-Challenging Language Model Agents BAGEL: Bootstrapping Agents by Guiding Exploration with Language
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35fe6451-5425-4df8-8a1d-3c4c4b110d75 · outbound
Self-Challenging Language Model Agents NNetNav: Unsupervised Learning of Browser Agents Through Environment Interaction in the Wild
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 069c1569-803a-4f1a-985f-f97a6115f015 · outbound
Self-Challenging Language Model Agents TOOLVERIFIER: Generalization to New Tools via Self-Verification
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7a30c5e-1ac8-484e-8257-9319e44a419a · outbound
Self-Challenging Language Model Agents Asymmetric self-play for automatic goal discovery in robotic manipulation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbc14c05-cbd6-4278-9d8e-95b3dd20e67a · outbound
Self-Challenging Language Model Agents Training Software Engineering Agents and Verifiers with SWE-Gym
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e72e425-b329-461c-9199-90b445ef3a6c · outbound
Self-Challenging Language Model Agents GPT-4 Technical Report
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcba1132-d734-494e-be0a-fc5a787f1ed5 · outbound
Self-Challenging Language Model Agents WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd9b4def-ad0f-4d5b-a55b-888658788f96 · outbound
Self-Challenging Language Model Agents ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dc74701-3f40-4066-a660-f5562c4d9d39 · outbound
Self-Challenging Language Model Agents Autonomous Evaluation and Refinement of Digital Agents
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28c6de98-65d2-48c1-919e-d9871dddb61f · outbound
Self-Challenging Language Model Agents Toolformer: Language Models Can Teach Themselves to Use Tools
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 423fe5aa-3fd5-4986-aa7a-15cc66178918 · outbound
Self-Challenging Language Model Agents Proximal Policy Optimization Algorithms
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89d3f2f5-b7e7-477f-a571-cb0b6ec08b64 · outbound
Self-Challenging Language Model Agents Manning, and Chelsea Finn
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d9150a1-90bd-4f29-a1de-30872947537f · outbound
Self-Challenging Language Model Agents Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f77b4a4f-c468-4e2b-a958-4a503f82f586 · outbound
Self-Challenging Language Model Agents Beyond Browsing: API-Based Web Agents
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba78255a-5eb4-47a1-95e4-33852fa53471 · outbound
Self-Challenging Language Model Agents Learn-by-interact: A Data-Centric Framework for Self-Adaptive Agents in Realistic Environments
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dfb5f3c-33f7-400d-935a-6330fb4c0154 · outbound
Self-Challenging Language Model Agents DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bbadf7a-eb55-403f-ade2-fb13aff91af2 · outbound
Self-Challenging Language Model Agents Tool Learning in the Wild: Empowering Language Models as Automatic Tool Agents
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c56968f5-fefb-43bb-b764-1f379a84f427 · outbound
Self-Challenging Language Model Agents AppWorld: A Controllable World of Apps and People for Benchmarking Interactive Coding Agents
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 668c004a-9259-4962-93a0-545038760051 · outbound
Self-Challenging Language Model Agents DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d5913fc-de68-4b39-ab35-09770eacc699 · outbound
Self-Challenging Language Model Agents Intrinsic Motivation and Automatic Curricula via Asymmetric Self-Play
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7944699a-559c-4c4d-bafc-ae724a4602e5 · outbound
Self-Challenging Language Model Agents The llama 3 herd of models, 2024
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2e468fff-96e1-4eef-b693-be3b9530f842 · outbound
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1b6eaf50-65fe-4888-aeed-bbb2b5a13b0d · outbound
Self-Challenging Language Model Agents TravelPlanner: A Benchmark for Real-World Planning with Language Agents
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cd74399-0b7b-453b-aa7f-aef601ca5057 · outbound
Self-Challenging Language Model Agents Executable code actions elicit better llm agents, 2024
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 54afcc56-60f7-46b8-b56a-0a1e403c9161 · outbound
Self-Challenging Language Model Agents Self-Instruct: Aligning Language Models with Self-Generated Instructions
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d4610b9-1bc1-40d5-a38a-0a7037f0988b · outbound
Self-Challenging Language Model Agents TheAgentCompany: Benchmarking LLM Agents on Consequential Real World Tasks
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73a172da-00aa-4749-86ca-5381f9f6f256 · outbound
Self-Challenging Language Model Agents Fung, Sha Li, Zixuan Huang, Xu Cao, Xingyao Wang, Yiquan Wang, Heng Ji, and Chengxiang Zhai
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e54719fc-d061-4af5-b4c3-0d10d14eea7c · outbound
Self-Challenging Language Model Agents OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a663a407-f076-41a0-aa8d-4a3156ce8dff · outbound
Self-Challenging Language Model Agents Building Math Agents with Multi-Turn Iterative Preference Learning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b24f184-baf4-4673-852c-2dfb7273c58f · outbound
Self-Challenging Language Model Agents Self-Rewarding Language Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10aa5b6b-5496-4ae2-9d7a-4ee6d4e62532 · outbound
Self-Challenging Language Model Agents OMNI: Open-endedness via Models of human Notions of Interestingness
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 152db4ab-e76c-4598-830a-565b6c63cde2 · outbound
Self-Challenging Language Model Agents If LLM Is the Wizard, Then Code Is the Wand: A Survey on How Code Empowers Large Language Models to Serve as Intelligent Agents
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9200b938-1689-4bd8-9392-2c18140e93c8 · outbound
Self-Challenging Language Model Agents ReAct: Synergizing Reasoning and Acting in Language Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8266c9a7-a113-45a3-ae5c-aa8a09a141f8 · outbound
Self-Challenging Language Model Agents $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe81ee85-bce5-4ea0-9666-744b80f639c5 · outbound
Self-Challenging Language Model Agents ArCHer: Training Language Model Agents via Hierarchical Multi-Turn RL
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17e33e8c-9b01-4cc9-a34c-5da8a28ac835 · outbound
Self-Challenging Language Model Agents SWEET-RL: Training Multi-Turn LLM Agents on Collaborative Reasoning Tasks
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6de11e5-b574-4e06-9e1a-5db9438d4c39 · outbound
Self-Challenging Language Model Agents Absolute Zero: Reinforced Self-play Reasoning with Zero Data
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04cc2d64-6626-4b19-9581-e619b5c30386 · outbound
Self-Challenging Language Model Agents WebArena: A Realistic Web Environment for Building Autonomous Agents
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b86ca1b9-160a-4040-a855-44640c0293b4 · outbound
Self-Challenging Language Model Agents Proposer-Agent-Evaluator(PAE): Autonomous Skill Discovery For Foundation Model Internet Agents
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce4b2e87-9041-4d54-8d72-75dbd4fe54dc · outbound
Self-Challenging Language Model Agents book_hotel
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ea84a94d-1271-4b97-9f62-71d004797953 · outbound
Self-Challenging Language Model Agents Unresolved cited work
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation de108500-ab37-433a-90e3-c454647d5249 · outbound
Self-Challenging Language Model Agents Unresolved cited work
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1ba7cca5-8bda-48e2-bed5-bc982210b34e · outbound
Self-Challenging Language Model Agents Unresolved cited work
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 868d2b48-4c99-48ab-b1dc-0a5d626a1325 · outbound
Self-Challenging Language Model Agents order by mistake
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 38b1a9f2-cdfd-41b1-aa91-b4e89a113c30 · outbound
Self-Challenging Language Model Agents doi: 10.1016/j.robot.2012.05.008
Reference 2013
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7c33ebd-f11e-44f7-a256-d7d3f4bfdf96 · outbound
Self-Challenging Language Model Agents DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c8e5dff-a1bb-4d5b-9570-67e5ef85cc2b · inbound
A Survey on Autonomy-Induced Security Risks in Large Model-Based Agents Self-Challenging Language Model Agents
Reference 162
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9b936a2-fdb3-4a97-83f4-d8fb2ad1bedd · inbound
On the Surprising Efficacy of LLMs for Penetration-Testing Self-Challenging Language Model Agents
Reference 123
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44909370-f083-47e1-ba86-69cc313916ef · inbound
EvoCurr: Self-evolving Curriculum with Behavior Code Generation for Complex Decision-making Self-Challenging Language Model Agents
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0919db2c-5c8f-4623-b730-30bdc51d60f7 · inbound
A global log for medical AI Self-Challenging Language Model Agents
Reference 124
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76755336-f839-4dc1-9ff6-32b58057e22b · inbound
Agent Learning via Early Experience Self-Challenging Language Model Agents
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b82e5b8-1b53-460d-a017-6896d798bfd9 · inbound
Help Without Being Asked: A Deployed Proactive Agent System for On-Call Support with Continuous Self-Improvement Self-Challenging Language Model Agents
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c40d748a-1a6a-45f1-b923-dca4511200a7 · inbound
Training LLM Agents for Spontaneous, Reward-Free Self-Evolution via World Knowledge Exploration Self-Challenging Language Model Agents
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ec11e069-6aa6-4de0-9353-d988fd8028ef · inbound
Bootstrapping Post-training Signals for Open-ended Tasks via Rubric-based Self-play on Pre-training Text Self-Challenging Language Model Agents
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d6a93fdf-ad16-42ef-be51-5363fbcb7baa · inbound
G-Zero: Self-Play for Open-Ended Generation from Zero Data Self-Challenging Language Model Agents
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bc6d68ac-dd6b-428c-8fd6-2d9f97ce6bfc · inbound
unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning Self-Challenging Language Model Agents
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 75264551-f090-40b5-8889-1034574942d0 · inbound
BenchEvolver: Frontier Task Synthesis via Solution-Centric Evolution Self-Challenging Language Model Agents
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bbab49ce-fdad-41c3-8d4e-f97b020e41b5 · inbound
SENTINEL: Failure-Driven Reinforcement Learning for Training Tool-Using Language Model Agents Self-Challenging Language Model Agents
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 690af5de-7de2-4896-8668-b5dddf254573 · inbound
PROTON: Prototype-Based Test-Time Online OOD Detection for Medical VLMs Self-Challenging Language Model Agents
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b43cfd81-e61e-4f3d-b85d-9e7af405a657 · inbound
Autodata: An agentic data scientist to create high quality synthetic data Self-Challenging Language Model Agents
Reference 125
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 63012396-d1c7-464f-ba22-ff404ea41de7 · inbound
Autodata: An agentic data scientist to create high quality synthetic data Self-Challenging Language Model Agents
Reference 125
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 876aefd1-24ed-4b36-9741-cf7d9a18fa55 · inbound
Autodata: An agentic data scientist to create high quality synthetic data Self-Challenging Language Model Agents
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27bf22bf-c004-4295-b46e-018c7c847b7b · inbound
Internalizing the Future: A Unified Agentic Training Paradigm for World Model Planning Self-Challenging Language Model Agents
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.