Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:55:09.996272Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 6 inbound Pith citation observations for arXiv:2507.01489.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:55:09.996272Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T00:55:55.496339Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-29T13:53:28.611729Z
15 of 15 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2980a235-2774-4187-bc1b-f97f589fd2ef · outbound
Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning Brown, B
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 741dd78a-e701-4532-96fa-7b09cd18633c · outbound
Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b56633a1-1bb9-4945-9047-497a5ffb5378 · outbound
Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd883f29-6f39-45e4-a322-8bb6646e2d57 · outbound
Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 579394ad-e4ce-4fc6-8220-841b831c109b · outbound
Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning Measuring and Narrowing the Compositionality Gap in Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4058c2a-53ed-4ea3-a043-9ebed6af1788 · outbound
Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning Qwen2.5 Technical Report
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65766e1d-9f9c-413a-b2a4-8a57bdb89f94 · outbound
Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a78b56a-609d-4642-bbdd-b527cbff4f49 · outbound
Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning ReAct: Synergizing Reasoning and Acting in Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc12c887-380f-4ca6-a429-67fadbedd70f · outbound
Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning OpenResearcher: Unleashing AI for Accelerated Scientific Research
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3779de78-c0be-4671-ac6a-69f58a4549d4 · outbound
Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning DeepResearcher: Scaling Deep Research via Reinforcement Learning in Real-world Environments
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb878f04-f70a-467a-a495-7e8d242f156b · outbound
Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7748746-bc8d-4743-82b9-64a1f7610247 · outbound
Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning MuSiQue: Multihop Questions via Single-hop Question Composition
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc514193-ee30-4d29-9e02-81507033a662 · outbound
Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning GPT-4o System Card
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44c22c11-5e34-44c1-a8b0-410e413b38c8 · outbound
Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b64dbb8-5014-4415-9e30-908677aa8367 · outbound
Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning Synthetic Data Generation & Multi-Step RL for Reasoning & Tool Use
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88b47baf-aff7-4e85-bc8e-def515eb4e0b · inbound
Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning
Reference 228
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 56960e50-3558-4410-8bca-ed3bd3392f08 · inbound
GEAR: Granularity-Adaptive Advantage Reweighting for LLM Agents via Self-Distillation Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 91881407-ef2d-4594-a886-d4bd1467fc30 · inbound
GEAR: Granularity-Adaptive Advantage Reweighting for LLM Agents via Self-Distillation Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2d126fbc-da39-44de-b05f-8ef1dc014e1f · inbound
IdleSpec: Exploiting Idle Time via Speculative Planning for LLM Agents Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ae45dc3-c1ce-487e-9d3b-aa30ec874642 · inbound
Knowing When to Ask: Segment-Level Credit Assignment for LLM Tool Use Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4efbff9a-c058-4e2b-9631-ae4d821c1538 · inbound
OpenClaw and Ollama in Agentic AI: Toward Fully Autonomous and Scalable AI Agent Systems Agent-as-Tool: A Study on the Hierarchical Decision Making with Reinforcement Learning
Reference 108
Source-reported events for the cited work
Unavailable: canonical work link unavailable.