Pith. sign in

Paper Citation Record · LEDGER

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents

As of 21 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2507.14897.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.14897 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:49:40.262037Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

26 of 26 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 545a452f-77c9-49c0-90c0-06a9dce691eb · outbound

This paper cites Ahmadian, C.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Ahmadian, C

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:49:42.542688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:49:37.344151Z digest=sha256:f5df373a4a697b7e134fc1c9dab6af89d818863350be28df95f40f6cb74ebe11

Observation 57a527e9-81cb-420d-8db1-5bef6d9e62c5 · outbound

This paper cites an unresolved cited work.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:49:42.408889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:49:37.443399Z digest=sha256:562b92bcf215a756ba6ac32dfffc0e250efc25d2ce4fc0a74478e27aeb4d7c75

Observation b09e553c-2459-4e37-8782-65cd2a56899e · outbound

This paper cites Group-in-Group Policy Optimization for LLM Agent Training.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Group-in-Group Policy Optimization for LLM Agent Training

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:49:37.552030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:49:37.552030Z digest=sha256:6ac74337e8a97e10b75419fbca8a8ac4f43e093f51c1fdb49bf859500fba0e20

Observation 814e8c07-510b-4f08-8a4e-724c1a893d1a · outbound

This paper cites an unresolved cited work.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:49:42.291835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:49:37.693263Z digest=sha256:834cd3203738ee23a59083206e6c85058eef1f422fa638f1f4597afa7bcb495a

Observation 9a5b15d0-9306-4dfa-9baa-9793114ebc4f · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:49:37.805080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:49:37.805080Z digest=sha256:7adabdc139b51b231a71ac023e6e6995ed50bf6525deed169fba6a8784d563ea

Observation 0a1c42fc-8829-41df-bf51-1cf17397ab75 · outbound

This paper cites an unresolved cited work.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:49:42.170209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:49:37.945699Z digest=sha256:cffcc206095122d214ea1fc9d2c2c2b3275c3c7984bb98fc85c173f863c281ce

Observation cb1bdac4-c8d3-4608-a7b1-6613022f3e80 · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T15:49:38.134448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:49:38.134448Z digest=sha256:79488c75ad643955b8f037385a56b6886a5dccef018dbe2da26fdd0bee4fa676

Observation 812e6747-e5ca-4c02-9d20-4c9cf3192ae8 · outbound

This paper cites Ouyang, J.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Ouyang, J

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:49:42.044253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:49:38.242729Z digest=sha256:7e97a35a875b68c77e2fa28f731b03a26d16047c450b92858414b6ecf169d0ed

Observation 40a437bd-051b-423d-9aeb-c2b5191c5802 · outbound

This paper cites an unresolved cited work.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:49:41.944332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:49:38.326110Z digest=sha256:aece070e25a8d563144300ff8c6ebf39b8ec47ed79341be58ba7d966fcadad6e

Observation a7e06540-0385-432d-8f66-c60d1a8d84be · outbound

This paper cites an unresolved cited work.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:49:38.468957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:49:38.468957Z digest=sha256:5771e58813c698e0d1d5d7af38d00fbfc82250af3dcf35ba49e63b536970381c

Observation eff85625-5804-4d7f-b0f3-7f8dbcbdedeb · outbound

This paper cites Proximal Policy Optimization Algorithms.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Proximal Policy Optimization Algorithms

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:49:38.583346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:49:38.583346Z digest=sha256:1a8e0f59b90ae22bbd2af5765f8de35f3692db7a66ac597be30d8f835d2acedb

Observation bf64fbaf-93c4-4dd1-b106-27eb9cf74d15 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:49:38.725073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:49:38.725073Z digest=sha256:6313c5ed42253b22bf69cdfbc4051bf80f555e665924800a8a41de6cb35d500f

Observation 363b84b1-c703-4a4d-b97f-8712702a66d8 · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents HybridFlow: A Flexible and Efficient RLHF Framework

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:49:38.832390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:49:38.832390Z digest=sha256:5b437bc0b6fc6a1356091a1ce76123949fc8ae700e8bc943d92eabc0b54e89ba

Observation 0fe9480a-1ad1-4035-9382-8f48545101ce · outbound

This paper cites ALFWorld: Aligning Text and Embodied Environments for Interactive Learning.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents ALFWorld: Aligning Text and Embodied Environments for Interactive Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T15:49:38.967794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:49:38.967794Z digest=sha256:746044a81ff9f960f8a822e8afade3310a3d64dae3594480b3c66369ec88612a

Observation 43ae38bb-e1e5-42a3-b986-1d616c41d714 · outbound

This paper cites Rl-factory.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Rl-factory

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:49:41.722444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:49:39.069127Z digest=sha256:b0939a8e5b4687f3aa6ef00cfd45fac49d501d0b602e49b72b2908e2b9379177

Observation 07ada5ad-2663-45b8-bb3b-9e9a93878ac9 · outbound

This paper cites Sumers, S.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Sumers, S

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:49:41.591462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:49:39.181757Z digest=sha256:532e20424c7120a8b924a89e1b296c16b5bc34fe85bf12a4fff47f914d5e44ca

Observation 2c791712-6f51-48c1-8d3f-971d617ffd7a · outbound

This paper cites an unresolved cited work.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:49:41.449138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:49:39.322374Z digest=sha256:6fd832b131f8d0fc24a078ecf0b27c0d5c0802b488d73613abe9297199cdce42

Observation bcb806f7-9237-444c-8551-380929bed4e9 · outbound

This paper cites ScienceWorld: Is your Agent Smarter than a 5th Grader?.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents ScienceWorld: Is your Agent Smarter than a 5th Grader?

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T15:49:39.486542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:49:39.486542Z digest=sha256:e25a812e2ac975cf5bf1696950eb4322a0e266163f2d73f14195d27387698d9e

Observation b51cc60f-e16b-4cc8-a9f6-07c1fae16fc4 · outbound

This paper cites an unresolved cited work.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:49:41.323935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:49:39.635367Z digest=sha256:8d67bfd2a042b011a88548adf84045298aaf3c0ddbf43b8614234c7cf5d5f29e

Observation bf73ace5-97c3-44ff-bcec-7e739bb4cb97 · outbound

This paper cites an unresolved cited work.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:49:41.137535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:49:39.762963Z digest=sha256:7392bcc22980830b49ba88a9328e52ddef6968a987e48136514381a0cbcbf80a

Observation 2ec57a07-5b94-4bf9-b66c-fcf93d52a08f · outbound

This paper cites an unresolved cited work.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:49:40.952067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:49:39.924443Z digest=sha256:77ff83cf08ea8b3b3f1157c70951c711eab4530fb4b42d80e6c166e797efe254

Observation ccacb417-bc29-4a6f-aea0-f5c336bc2c3d · outbound

This paper cites WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:49:40.070569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:49:40.070569Z digest=sha256:6188c5d8410994ebcdac577b94e07cd2c5e67b039252e56218f2963881d172a4

Observation 73a24159-f12e-439c-b31a-f3b636ab88f6 · outbound

This paper cites an unresolved cited work.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:49:40.824110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:49:40.129587Z digest=sha256:5d02e952d348a9063c0e5eeb4e0455ad10e95a2c2207ad5560040b086e7f685a

Observation f9a5efa7-819c-47cd-8a32-c85e981442a2 · outbound

This paper cites an unresolved cited work.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:49:40.675842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:49:40.191632Z digest=sha256:9bd41fb90bdbf0db972c14191725fd51a28ce4867d60c982c8575c47a7e1c764

Observation b5938653-f283-4f96-9fe5-667ca10fd4d9 · outbound

This paper cites qa_f1_reward.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents qa_f1_reward

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:49:40.459525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:49:40.262037Z digest=sha256:e61c0263ff735d53e0c1955ea0c884a132e30d651a30e75b42e58894c3d45e6f

Observation 6da7cb92-1644-4bd1-bfca-f184dc506489 · outbound

This paper cites an unresolved cited work.

AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents Unresolved cited work

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T15:49:40.015259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:49:40.015259Z digest=sha256:c318e66fdec3c821042fae9149fd1f4f173bfe2e5b9d930105c0f3b4de72d2f6

Pith citing papers

No inbound Pith citation observations are available.