Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:49:40.262037Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2507.14897.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:49:40.262037Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
26 of 26 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 545a452f-77c9-49c0-90c0-06a9dce691eb · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Ahmadian, C
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 57a527e9-81cb-420d-8db1-5bef6d9e62c5 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b09e553c-2459-4e37-8782-65cd2a56899e · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Group-in-Group Policy Optimization for LLM Agent Training
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 814e8c07-510b-4f08-8a4e-724c1a893d1a · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 9a5b15d0-9306-4dfa-9baa-9793114ebc4f · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a1c42fc-8829-41df-bf51-1cf17397ab75 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation cb1bdac4-c8d3-4608-a7b1-6613022f3e80 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 812e6747-e5ca-4c02-9d20-4c9cf3192ae8 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Ouyang, J
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 40a437bd-051b-423d-9aeb-c2b5191c5802 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation a7e06540-0385-432d-8f66-c60d1a8d84be · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eff85625-5804-4d7f-b0f3-7f8dbcbdedeb · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Proximal Policy Optimization Algorithms
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf64fbaf-93c4-4dd1-b106-27eb9cf74d15 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 363b84b1-c703-4a4d-b97f-8712702a66d8 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents HybridFlow: A Flexible and Efficient RLHF Framework
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fe9480a-1ad1-4035-9382-8f48545101ce · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents ALFWorld: Aligning Text and Embodied Environments for Interactive Learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43ae38bb-e1e5-42a3-b986-1d616c41d714 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Rl-factory
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 07ada5ad-2663-45b8-bb3b-9e9a93878ac9 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Sumers, S
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2c791712-6f51-48c1-8d3f-971d617ffd7a · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation bcb806f7-9237-444c-8551-380929bed4e9 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents ScienceWorld: Is your Agent Smarter than a 5th Grader?
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b51cc60f-e16b-4cc8-a9f6-07c1fae16fc4 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation bf73ace5-97c3-44ff-bcec-7e739bb4cb97 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2ec57a07-5b94-4bf9-b66c-fcf93d52a08f · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ccacb417-bc29-4a6f-aea0-f5c336bc2c3d · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73a24159-f12e-439c-b31a-f3b636ab88f6 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f9a5efa7-819c-47cd-8a32-c85e981442a2 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b5938653-f283-4f96-9fe5-667ca10fd4d9 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents qa_f1_reward
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6da7cb92-1644-4bd1-bfca-f184dc506489 · outbound
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.