Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2505.07773.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T12:23:55.198073Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T16:29:57.263901Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 180f9b04-4a4c-4050-8278-f640fa91051b · inbound
AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82b74efe-b908-4be7-8a9d-f20ce2219763 · inbound
Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e85ad13-c4e7-40ce-8ed5-8bdbda4a1f96 · inbound
rStar2-Agent: Agentic Reasoning Technical Report Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bb1be83-9bc9-4c92-9ae7-b325b209dc4b · inbound
SimpleTIR: End-to-End Reinforcement Learning for Multi-Turn Tool-Integrated Reasoning Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7140552f-ec9d-448e-97bf-b4347eeb36f3 · inbound
Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 06159f94-f92e-435e-92db-076035518def · inbound
ReSeek: A Self-Correcting Framework for Search Agents with Instructive Rewards Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation cae61186-19c1-47b3-a9c8-511c1c15a7e5 · inbound
CostBench: Evaluating Multi-Turn Cost-Optimal Planning and Adaptation in Dynamic Environments for LLM Tool-Use Agents Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 988f592a-a2e1-4807-a6cf-a8a1c1f34491 · inbound
CostBench: Evaluating Multi-Turn Cost-Optimal Planning and Adaptation in Dynamic Environments for LLM Tool-Use Agents Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dcb72e0-1c50-4f99-8526-881157608874 · inbound
Intern-S1-MO: Long-horizon Reasoning Agent for Olympiad?Level Mathematical Problem Solving Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19d6a823-d3f4-43b7-b473-f5f0fbd076c8 · inbound
AutoTool: Dynamic Tool Selection and Integration for Agentic Reasoning Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b980d36e-f6be-41c2-9d23-73eeefa72421 · inbound
Making Expert Reasoning Learnable with Self-Distillation Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a63896ad-167b-4433-a866-9d3e277308e7 · inbound
Multi-Modal Multi-Agent Reinforcement Learning for Radiology Report Generation Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fb5d1019-2f23-4e62-9b69-c82919ee8576 · inbound
Walk the Talk: Bridging the Reasoning-Action Gap for Thinking with Images via Multimodal Agentic Policy Optimization Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b733163e-6620-46d7-815e-d53f3b665ac6 · inbound
Rethinking Agentic Reinforcement Learning In Large Language Models Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 410b48f9-75f4-4f9a-bc58-6b4240c966b2 · inbound
Rethinking Agentic Reinforcement Learning In Large Language Models Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fb9f6857-d435-4436-8ddb-fc5d03233f04 · inbound
Rethinking Agentic Reinforcement Learning In Large Language Models Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4c808a61-188d-4e55-9e86-c5dd66b4d398 · inbound
PruneTIR: Inference-Time Tool Call Pruning for Effective yet Efficient Tool-Integrated Reasoning Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e74073e5-3fd5-4790-919a-d63fab83ad43 · inbound
IAPO: Input Attribution-Aware Policy Optimization for Tool Use in Small Multimodal Agents Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f13c40ec-3a25-460e-8d45-abf23c9f8f8a · inbound
ReSum: Synergizing LLM Reasoning and Summarization with Reinforcement Learning Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8f44afc6-8f45-4f69-b31c-6d109f32a5c8 · inbound
ReSum: Synergizing LLM Reasoning and Summarization with Reinforcement Learning Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41f5a501-6ba6-4947-a74c-a9677759db78 · inbound
STARE: Surprisal-Guided Token-Level Advantage Reweighting for Policy Entropy Stability Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9dd4607a-78b4-4fc7-8563-ef806333cb08 · inbound
Latent Visual States for Efficient Multimodal Reasoning Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 65a78ef7-3e5b-49be-aeda-577874f59027 · inbound
PEARL: Solver-in-the-Loop Interactive Optimization Modeling from Natural Language Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.