Pith. sign in

Paper Citation Record · LEDGER

FireAct: Toward Language Agent Fine-tuning

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 69 inbound Pith citation observations for arXiv:2310.05915.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.05915 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 69 of 69 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 69 of 69 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T20:57:48.380397Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

6
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b87ab206-5067-4d51-9903-e4393ce36fea · inbound

Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security cites this paper.

Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security FireAct: Toward Language Agent Fine-tuning

Reference 247

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:57:26.697334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T00:57:26.303195Z digest=sha256:d7466bd66e69bcd92a441193e0f97b975394d262c2b9dc120167300c2d777c19

Observation 646dde47-932f-4ed0-aa43-f7185d9a3607 · inbound

Efficient Multi-Agent System Training with Data Influence-Oriented Tree Search cites this paper.

Efficient Multi-Agent System Training with Data Influence-Oriented Tree Search FireAct: Toward Language Agent Fine-tuning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:07:30.575127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T04:06:23.521344Z digest=sha256:9dda7cc4bfa9d3d1cdc4018e7a9f1f8ed97972631c32ecbe33848dc63f478498

Observation 22b0482f-0f63-45dc-9219-3a0fb7b7248f · inbound

Digi-Q: Learning Q-Value Functions for Training Device-Control Agents cites this paper.

Digi-Q: Learning Q-Value Functions for Training Device-Control Agents FireAct: Toward Language Agent Fine-tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T20:57:48.380397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:57:48.380397Z digest=sha256:91344c310f81f60450e18f0f444130b9de9b84888fcff827144af901b574f93e

Observation 3a52da57-1d21-4abd-889f-50538f316495 · inbound

Effective Reinforcement Learning for Reasoning in Language Models cites this paper.

Effective Reinforcement Learning for Reasoning in Language Models FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:56:15.384955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:56:15.384955Z digest=sha256:de8c53a5cbc8f0bcc511c14fcff13d40ae293f091f918ab1a8f1dd245c2a45f4

Observation 9d59dde0-aa66-4377-b779-ff008ad7da81 · inbound

Large Language Models for Planning: A Comprehensive and Systematic Survey cites this paper.

Large Language Models for Planning: A Comprehensive and Systematic Survey FireAct: Toward Language Agent Fine-tuning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:11:51.825888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:11:51.825888Z digest=sha256:b6ace3e96ac8f3418e722b75bbdad5ce51e0ae462265b1e0f40fc86f7ed8cfb8

Observation 4e07b629-aa8e-4ecc-a5f4-db57fc4ab08d · inbound

Divide and Conquer: Grounding LLMs as Efficient Decision-Making Agents via Offline Hierarchical Reinforcement Learning cites this paper.

Divide and Conquer: Grounding LLMs as Efficient Decision-Making Agents via Offline Hierarchical Reinforcement Learning FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:10:02.785697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:10:02.785697Z digest=sha256:8911364197f746b89c9202da7272ab276c2040760cf14ac3e769445f7e262617

Observation 24ba30b1-3a1a-4744-9b32-ffb030550847 · inbound

Training LLM-Based Agents with Synthetic Self-Reflected Trajectories and Partial Masking cites this paper.

Training LLM-Based Agents with Synthetic Self-Reflected Trajectories and Partial Masking FireAct: Toward Language Agent Fine-tuning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:53.740452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:06:53.740452Z digest=sha256:e2057712362296205e16855500f5b0f33b32226ced008800ee5f3b6e643467e0

Observation acce814c-6101-497e-8130-b2be289ae051 · inbound

SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution cites this paper.

SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution FireAct: Toward Language Agent Fine-tuning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:02.795247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:02.795247Z digest=sha256:036757c2a0962ca35234b1e28a3ed8ade2ab319a1aeda47736dcd8a3b93bd0f7

Observation 34e23431-e069-445d-90ab-5951361985e9 · inbound

RRO: LLM Agent Optimization Through Rising Reward Trajectories cites this paper.

RRO: LLM Agent Optimization Through Rising Reward Trajectories FireAct: Toward Language Agent Fine-tuning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:51:41.263415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:51:41.263415Z digest=sha256:4e737343155943f17a590376a352b22a439e23fc2dd17d44307e94498c973efe

Observation e42a0d3b-5c02-40f9-8b20-fe1b05831de5 · inbound

Agent-Environment Alignment via Automated Interface Generation cites this paper.

Agent-Environment Alignment via Automated Interface Generation FireAct: Toward Language Agent Fine-tuning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:30.856512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:45:30.856512Z digest=sha256:250e846e2552200467d2a60e9e49ee543f1e42631b8266a9d8648b10217f72ba

Observation 05a885bf-085a-4741-89e4-742cb11aae9e · inbound

WebDancer: Towards Autonomous Information Seeking Agency cites this paper.

WebDancer: Towards Autonomous Information Seeking Agency FireAct: Toward Language Agent Fine-tuning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:59.724999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:59.724999Z digest=sha256:fd6b06e26cf6f71ceb7392afa4888d7d179176ffd21bd50684f46cdf761bb0c4

Observation da351dd0-93da-4672-b144-ced5533be68a · inbound

OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation cites this paper.

OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation FireAct: Toward Language Agent Fine-tuning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:31.904774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:31.904774Z digest=sha256:14bf3bed42c3f81942fe78869b8a3fab418fb12363fbae23771478b02453edfd

Observation 622a48e0-de72-472a-a0c6-f43cbd5a0d5a · inbound

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization cites this paper.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.567143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:00.567143Z digest=sha256:bd9c35c46a5589d9835a5b7e2da86c64d7873428c83fba802f90f86aa6ffd9da

Observation 7722ca1e-4be8-49b1-b8d0-65aaa2cb1e06 · inbound

LAM SIMULATOR: Advancing Data Generation for Large Action Model Training via Online Exploration and Trajectory Feedback cites this paper.

LAM SIMULATOR: Advancing Data Generation for Large Action Model Training via Online Exploration and Trajectory Feedback FireAct: Toward Language Agent Fine-tuning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:54.902423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:54.902423Z digest=sha256:ba81fa67802dae8c39b975c672527ab0243437c724f90370189a674d56ea7a96

Observation 5753f9fc-d847-4e26-a0c5-593e07aa9abe · inbound

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games cites this paper.

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games FireAct: Toward Language Agent Fine-tuning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-19T12:02:16.630032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T12:01:42.681135Z digest=sha256:852dc6039f0674d1dac6e2e7cd292fdfdfc6b3ca3df08f228515f8b0e7b6523d

Observation 1ac41ca7-a31c-4ac2-885e-2f056241fcbe · inbound

Automated Skill Discovery for Language Agents through Exploration and Iterative Feedback cites this paper.

Automated Skill Discovery for Language Agents through Exploration and Iterative Feedback FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:25.898272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:25.898272Z digest=sha256:29ec3d28cc7bee8fbfe3cb4c2f1601c0983f722548d0cebe8eec119b38f21b74

Observation cfbb0d56-e472-4344-ab30-6a7a09779e7c · inbound

Agent-RewardBench: Towards a Unified Benchmark for Reward Modeling across Perception, Planning, and Safety in Real-World Multimodal Agents cites this paper.

Agent-RewardBench: Towards a Unified Benchmark for Reward Modeling across Perception, Planning, and Safety in Real-World Multimodal Agents FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:42.146724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:36:42.146724Z digest=sha256:d499da9c595d6de0e6b96bc95ffdd3ff158854bf33df5407d7aa5356ea8d160a

Observation 57d5bf82-b0ed-40e6-a1f0-041237f1d47e · inbound

Knowledge Augmented Finetuning Matters in both RAG and Agent Based Dialog Systems cites this paper.

Knowledge Augmented Finetuning Matters in both RAG and Agent Based Dialog Systems FireAct: Toward Language Agent Fine-tuning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:59:59.720213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:59:59.720213Z digest=sha256:7fdc4881a1df9342b6ff3465a768b9b49c8c1d4146d8311f331218a7528a73b9

Observation 92f03b00-374b-4a14-9c7c-068dcff9429b · inbound

Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning cites this paper.

Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:00.459541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:55:00.459541Z digest=sha256:4a4f3eb7701bdf7d589f3af27d092f8ba91e993d8aef3dbed9d7ba8ff6106d1c

Observation dfcb3673-566d-4a51-bd5f-6caf64e3b635 · inbound

WebSailor: Navigating Super-human Reasoning for Web Agent cites this paper.

WebSailor: Navigating Super-human Reasoning for Web Agent FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:37:09.609424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T15:37:09.572241Z digest=sha256:249d2ff0f8327dc1d24545af97038be8c62bf82e797d8f6e79edc8cebea3677c

Observation 4ece5b4b-0d92-46bc-a299-ae61de169c11 · inbound

SAND: Boosting LLM Agents with Self-Taught Action Deliberation cites this paper.

SAND: Boosting LLM Agents with Self-Taught Action Deliberation FireAct: Toward Language Agent Fine-tuning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:47:19.417248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:47:19.417248Z digest=sha256:43ed7461bf4aae9885fa5a745af7d4b97c4386b068e6a9cf1c53ac03263e1ac7

Observation a6f50084-356e-42ed-b398-20cf7978c4cf · inbound

Initial Steps in Integrating Large Reasoning and Action Models for Service Composition cites this paper.

Initial Steps in Integrating Large Reasoning and Action Models for Service Composition FireAct: Toward Language Agent Fine-tuning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T14:33:53.491820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:33:53.491820Z digest=sha256:417cd6e9a6ac64f63f56bb6c131c07ad7f4f9096bca8b0cd639d4c7a6708dbf0

Observation 3baeeb8c-e4ef-4928-8cff-4e32bae54d62 · inbound

MMAT-1M: A Large Reasoning Dataset for Multimodal Agent Tuning cites this paper.

MMAT-1M: A Large Reasoning Dataset for Multimodal Agent Tuning FireAct: Toward Language Agent Fine-tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T12:18:31.887391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:18:31.887391Z digest=sha256:bbe8b1a10a941400984558fb59412df2aba9cba4337335a32b2923e225d33b27

Observation 7d149853-178e-4757-a71c-fdf5f94f7ea8 · inbound

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey cites this paper.

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey FireAct: Toward Language Agent Fine-tuning

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-18T19:21:48.752794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T19:19:36.427337Z digest=sha256:ccd4796a1d83c99871c44776729197ea2d8723217e6718596200e32751ced029

Observation 233947b3-af51-43fb-aa22-d60d4dd230e7 · inbound

Bridging the Capability Gap: Joint Alignment Tuning for Harmonizing LLM-based Multi-Agent Systems cites this paper.

Bridging the Capability Gap: Joint Alignment Tuning for Harmonizing LLM-based Multi-Agent Systems FireAct: Toward Language Agent Fine-tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T18:50:10.443190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T18:50:10.443190Z digest=sha256:0730842db7d236e9a33a340257b93f1b8f22f12cc8a1a5deab6f33b71780891b

Observation c363136c-8b56-418e-a591-d3f89be6daf9 · inbound

On-Device Fine-Tuning via Backprop-Free Zeroth-Order Optimization cites this paper.

On-Device Fine-Tuning via Backprop-Free Zeroth-Order Optimization FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:05:21.475359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T22:03:53.594703Z digest=sha256:1c6ebe3a028fa2fd44096d9ee0355581a9f992eaace1dbf9c7735cfc0ad48297

Observation a703daeb-00fa-40df-9da0-698878cdb6e4 · inbound

MemVerse: Multimodal Memory for Lifelong Learning Agents cites this paper.

MemVerse: Multimodal Memory for Lifelong Learning Agents FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T18:48:57.657515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:48:57.657515Z digest=sha256:51cbc9567bde0df4ea9e4c7accee1960427820e90dc4b651043cb1aaad0f920f

Observation 701c787b-a889-4547-acf4-0b55d6b846b8 · inbound

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models cites this paper.

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models FireAct: Toward Language Agent Fine-tuning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T11:44:05.540078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T11:44:05.540078Z digest=sha256:e49ba5c2b1e7beb73e847cb3af344041147abd5a4624264aca2000d695d81186

Observation 215a4d4b-b014-49a7-b769-75b812a26f03 · inbound

SoK: Agentic Skills -- Beyond Tool Use in LLM Agents cites this paper.

SoK: Agentic Skills -- Beyond Tool Use in LLM Agents FireAct: Toward Language Agent Fine-tuning

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:19:31.175043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T23:19:31.024268Z digest=sha256:b748e41493dc7fcc11c83338d2b9dba3f42bb89eeaca60222c8d584c78650783

Observation a2334525-f988-410c-bc05-b0a5cb16df15 · inbound

ForkKV: Scaling Multi-LoRA Agent Serving via Copy-on-Write Disaggregated KV Cache cites this paper.

ForkKV: Scaling Multi-LoRA Agent Serving via Copy-on-Write Disaggregated KV Cache FireAct: Toward Language Agent Fine-tuning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:10:54.920774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:16:49.292491Z digest=sha256:9097974f6b9c87679cffc19a2c4bbbde180ffcf7ef9fbe00116332feac223b35

Observation ebdc028a-c5d4-46cc-9c4e-61c78c258e34 · inbound

PASK: Toward Intent-Aware Proactive Agents with Long-Term Memory cites this paper.

PASK: Toward Intent-Aware Proactive Agents with Long-Term Memory FireAct: Toward Language Agent Fine-tuning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:30:59.754002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:05:40.242925Z digest=sha256:36512ff4bdd0ec53dfd13027df9a0cbc38ceef53d78e21247879d6d843cb5ee6

Observation b0613023-41f0-471c-990b-2b4ecff093f3 · inbound

From Coarse to Fine: Self-Adaptive Hierarchical Planning for LLM Agents cites this paper.

From Coarse to Fine: Self-Adaptive Hierarchical Planning for LLM Agents FireAct: Toward Language Agent Fine-tuning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:46:10.623340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T08:10:36.579810Z digest=sha256:b8e73db769e234483faf1ecc1da0fb469a692388543e6a3b42b71e482cbba6b2

Observation d76db5ba-0679-41f7-aad1-bd04ba976c6a · inbound

Rethinking Agentic Reinforcement Learning In Large Language Models cites this paper.

Rethinking Agentic Reinforcement Learning In Large Language Models FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:21:29.906759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T06:30:09.945371Z digest=sha256:f4aa84d8b07ee08d9ae5ea01da13c2e0e8b2f15e26a471937852f8faaa1a6866

Observation 0cfc38ba-4da7-40ec-b7a5-cd2a40a62e60 · inbound

Rethinking Agentic Reinforcement Learning In Large Language Models cites this paper.

Rethinking Agentic Reinforcement Learning In Large Language Models FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:11:17.132850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T03:12:19.414358Z digest=sha256:3191250cd3bf0d5d3a11cfeaf7048123280ea2100489948fbbd5c24be1c7270c

Observation 2816bf74-7e82-4ed4-bfc5-c5539c5ab24a · inbound

Rethinking Agentic Reinforcement Learning In Large Language Models cites this paper.

Rethinking Agentic Reinforcement Learning In Large Language Models FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-19T17:02:40.660252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T16:58:41.558250Z digest=sha256:7c8fb7738b5fe31f67aed3cadb752835787d7521ded5aaed4d5486bf7f35119c

Observation 414876f4-f394-4d5f-ae80-fb49c6d560d0 · inbound

SOD: Step-wise On-policy Distillation for Small Language Model Agents cites this paper.

SOD: Step-wise On-policy Distillation for Small Language Model Agents FireAct: Toward Language Agent Fine-tuning

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:40:54.506607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T02:25:59.056181Z digest=sha256:28d7335be3cc3f4e6ec0f6bc9c84ee8dc1ca6c2adc7574220543e745d4157f76

Observation a76083b1-83e4-42ea-9057-87b5a5a12829 · inbound

SOD: Step-wise On-policy Distillation for Small Language Model Agents cites this paper.

SOD: Step-wise On-policy Distillation for Small Language Model Agents FireAct: Toward Language Agent Fine-tuning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T05:20:45.134981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:20:45.134981Z digest=sha256:41e40296b5632f8246dc3788cf3b09b0273fe1748ec38d6cace513b2db0703aa

Observation c713480d-1da4-4114-941f-bc60dfa7ca11 · inbound

Learning CLI Agents with Structured Action Credit under Selective Observation cites this paper.

Learning CLI Agents with Structured Action Credit under Selective Observation FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:00:56.420074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T02:59:26.100818Z digest=sha256:ac0135d6c9ecf06c422a70d523eade4fb361a6f0cbc86838810aa30c930ca59f

Observation e71c9f61-b55d-40a9-ad30-20bcc08b828b · inbound

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation cites this paper.

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation FireAct: Toward Language Agent Fine-tuning

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:11:27.660465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T03:36:24.941205Z digest=sha256:242eb772a7d435c243017dd5eb12b16e754cb2a278891c42ea101289fae05edd

Observation 3701cabe-b463-4441-948b-0692179f4b81 · inbound

Verifiable Process Rewards for Agentic Reasoning cites this paper.

Verifiable Process Rewards for Agentic Reasoning FireAct: Toward Language Agent Fine-tuning

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:46:27.390166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T04:59:44.082011Z digest=sha256:cb013610cfdda92110f1a6c903dce40a5d014b2db8d39f993b4740d5a78f4016

Observation 972a20b1-7617-496f-a8cf-7bf7c87c0396 · inbound

Verifiable Process Rewards for Agentic Reasoning cites this paper.

Verifiable Process Rewards for Agentic Reasoning FireAct: Toward Language Agent Fine-tuning

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-30T22:45:07.256818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T22:43:50.317854Z digest=sha256:72520fccdd8fb50da45237bd1cebd310ddf7818c5ec1eedfbd1996ff04bd6df1

Observation 50bf99b1-d234-4603-9715-a2fd78b34f04 · inbound

SkillGen: Verified Inference-Time Agent Skill Synthesis cites this paper.

SkillGen: Verified Inference-Time Agent Skill Synthesis FireAct: Toward Language Agent Fine-tuning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:17:23.043506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T06:14:28.614825Z digest=sha256:52f71760de92875b55bc27f315bf7537921ed65665d79c099659898c2ce156f4

Observation d945513a-b9ff-4c3d-8ce0-3cf7ca4ee900 · inbound

ClawForge: Generating Executable Interactive Benchmarks for Command-Line Agents cites this paper.

ClawForge: Generating Executable Interactive Benchmarks for Command-Line Agents FireAct: Toward Language Agent Fine-tuning

Reference 87

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T05:09:45.136629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-15T05:08:15.750558Z digest=sha256:6ccee0a85a0033a4ef353bfc90037b5be0ed6c62b6281f1bbd425e5731120df4

Observation 03da1012-3696-4a93-90a2-e6fa14d1f4d9 · inbound

ClawForge: Generating Executable Interactive Benchmarks for Command-Line Agents cites this paper.

ClawForge: Generating Executable Interactive Benchmarks for Command-Line Agents FireAct: Toward Language Agent Fine-tuning

Reference 87

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T20:23:43.246128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T20:19:21.824216Z digest=sha256:d132d9801195fab45b7ad4ab3ef25a9262be030180d157674e41bf3afa773190

Observation 7c719705-040b-4d02-826f-e8d9c0595458 · inbound

TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination cites this paper.

TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination FireAct: Toward Language Agent Fine-tuning

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T18:02:42.288078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-19T18:01:06.649723Z digest=sha256:c34502d8a7a4806a75c18d35a5d1cbf044d391cb1cae0328911cc4fe683c4354

Observation 37d965ed-ee2c-4c4d-87e8-d9aac2c260eb · inbound

AMATA: Adaptive Multi-Agent Trajectory Alignment for Knowledge-Intensive Question Answering cites this paper.

AMATA: Adaptive Multi-Agent Trajectory Alignment for Knowledge-Intensive Question Answering FireAct: Toward Language Agent Fine-tuning

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:48:19.927143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T13:43:42.097447Z digest=sha256:d6a93f48e20d7626aeeb59f6076936f7b50b646cb946ae8e11234a7037e00709

Observation bf1dca57-c7ec-4bf1-a977-a4b09a47e521 · inbound

Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost cites this paper.

Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost FireAct: Toward Language Agent Fine-tuning

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T06:31:10.437580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-22T06:26:31.841138Z digest=sha256:64e965a91973799b2b7254b5d514c9dd8550d6acd13401507fbe2e1699cd60c3

Observation 87fd8886-651e-4c1b-9b21-1a64611edfbf · inbound

Test-Time Deep Thinking to Explore Implicit Rules cites this paper.

Test-Time Deep Thinking to Explore Implicit Rules FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:54:38.220238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T11:52:15.163893Z digest=sha256:5cd821e16f881dda64093c46c93d96559105c76b6afc06e82c2dfe1bb5520d67

Observation 38045e09-7dbd-448d-89d7-bdbc8e7d98c3 · inbound

Agent Explorative Policy Optimization for Multimodal Agentic Reasoning cites this paper.

Agent Explorative Policy Optimization for Multimodal Agentic Reasoning FireAct: Toward Language Agent Fine-tuning

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:23:23.735371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T12:22:39.655615Z digest=sha256:865118b0a72ed38260d19d15a80487112ffab566e43c1982f377031000b87806

Observation 64d64355-8c2a-4094-8ea1-bf6143d4dfe2 · inbound

COMAP: Co-Evolving World Models and Agent Policies for LLM Agents cites this paper.

COMAP: Co-Evolving World Models and Agent Policies for LLM Agents FireAct: Toward Language Agent Fine-tuning

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:16:24.848399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T14:27:50.260308Z digest=sha256:da6d22e8005ae7d18a5778174cf2189cbd8a1537011c48ed5d7ce730d9fcb3df

Observation 71ae13f4-84af-4eac-86d7-6e31862d9f7f · inbound

SaliMory: Orchestrating Cognitive Memory for Conversational Agents cites this paper.

SaliMory: Orchestrating Cognitive Memory for Conversational Agents FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:16:31.723007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T10:09:26.995556Z digest=sha256:3982719bb09729832da4fee973a31288912f631260ef2c3f16445a41f82b7709

Observation ee620fc9-0804-4521-bc91-bd3ab31fbd65 · inbound

Self-evolving LLM agents with in-distribution Optimization cites this paper.

Self-evolving LLM agents with in-distribution Optimization FireAct: Toward Language Agent Fine-tuning

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:57:09.540367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T22:18:27.021136Z digest=sha256:77826153568c0d132fdce71e8dedaf1adabec2786d6ab285f15d5432b3bb27a9

Observation 93edcb76-e681-48db-94cb-a54fbdee09f9 · inbound

Evoflux: Inference-Time Evolution of Executable Tool Workflows for Compact Agents cites this paper.

Evoflux: Inference-Time Evolution of Executable Tool Workflows for Compact Agents FireAct: Toward Language Agent Fine-tuning

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-06-27T09:50:48.183424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T09:46:59.954720Z digest=sha256:e2cfddda9fb14ddbb8f1aaac438f0f71c7d201b6ac46e248488ea72b834fe1c6

Observation d2f72c89-7104-4a4b-a26a-624139ce3b11 · inbound

Training the Orchestrator: A Supervised Approach to End-to-End PDDL Planning with LLM Agents cites this paper.

Training the Orchestrator: A Supervised Approach to End-to-End PDDL Planning with LLM Agents FireAct: Toward Language Agent Fine-tuning

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:59:38.128661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T13:56:51.914966Z digest=sha256:7eb915d9672437dae511297863b040794e62e459836dd2cacd3e41242be75fab

Observation 698a2d61-ed51-475a-b800-89780383ac57 · inbound

AgentOdyssey: Open-Ended Long-Horizon Text Game Generation for Test-Time Continual Learning Agents cites this paper.

AgentOdyssey: Open-Ended Long-Horizon Text Game Generation for Test-Time Continual Learning Agents FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:46:11.465290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T21:59:25.449570Z digest=sha256:32cc976032f72daed0a824066aae58562bd6918a5d12fae7c4a6c4a5bbfbbd4a

Observation b5074fd1-6551-4c79-8f50-a25b58c0258c · inbound

Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks cites this paper.

Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks FireAct: Toward Language Agent Fine-tuning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:49:52.879268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T05:51:01.446139Z digest=sha256:761890c7c2c89e8ed872a7821d911bedd98dff7318b9f84d0daac64ffa3917bd

Observation 8096747b-5fb9-4f05-8ddc-2319995977ba · inbound

Agentic-Ideation: Sample Efficient Agentic Trajectories Synthesis for Scientific Ideation Agents cites this paper.

Agentic-Ideation: Sample Efficient Agentic Trajectories Synthesis for Scientific Ideation Agents FireAct: Toward Language Agent Fine-tuning

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:05:41.381986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T05:49:52.979376Z digest=sha256:603d7b00384a5e87958f154be5a8d07e3316ba5b32a49c872fde7a63492b605b

Observation 2444a382-c0b1-40c2-900f-f5c24b58fa44 · inbound

Atomic Task Graph: A Unified Framework for Agentic Planning and Execution cites this paper.

Atomic Task Graph: A Unified Framework for Agentic Planning and Execution FireAct: Toward Language Agent Fine-tuning

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:08:21.353187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T13:59:31.697155Z digest=sha256:e319c2e9041e4b3a4ee214e3817754169333102ca5bfce0aca865e79569b76cb

Observation 56cecc91-1805-4e49-912f-9797b1ea6101 · inbound

STAPO: Selective Trajectory-Aware Policy Optimization for LLM Agent Training cites this paper.

STAPO: Selective Trajectory-Aware Policy Optimization for LLM Agent Training FireAct: Toward Language Agent Fine-tuning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-11T10:50:54.419477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T10:50:54.419477Z digest=sha256:ca0465ff6ae9e316d55b3e09871df9828c0fc4442017c77965608adfaeae58ab

Observation 0d8d235e-bea8-4cda-88cc-f30d8bc5c0a4 · inbound

Eluna: An Agentic LLM System for Automating Warehouse Operations with Reasoning and Task Execution cites this paper.

Eluna: An Agentic LLM System for Automating Warehouse Operations with Reasoning and Task Execution FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-13T05:30:10.497783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T05:30:10.497783Z digest=sha256:50b5312bc4691b4125b23634f31f2f67144979eb08dce3f3b01b1ebdbdf31069

Observation fadb3a24-6dea-449d-8890-a67f0292a037 · inbound

Agentic-DPO: From Imitation to Agentic Policy Optimization on Expert Trajectories cites this paper.

Agentic-DPO: From Imitation to Agentic Policy Optimization on Expert Trajectories FireAct: Toward Language Agent Fine-tuning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T10:33:54.851493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T10:33:54.851493Z digest=sha256:7157ff8bcd82dea76ffb9206f94ec925e0722b80681cad60229b9b52bf43ba14

Observation 0812fa76-c88c-4290-9848-71d20c002f2b · inbound

From Outcomes to Actions: Leveraging Hindsight for Long-Horizon Language Agent Training cites this paper.

From Outcomes to Actions: Leveraging Hindsight for Long-Horizon Language Agent Training FireAct: Toward Language Agent Fine-tuning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T09:45:49.534257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:45:49.534257Z digest=sha256:dede0bd8d96e25fe47ff3fe0edaaa20dbd79d75a128a0edd50fc516f54bbabf2

Observation 8650e116-c765-4a9c-bdf2-a1b5bc534758 · inbound

Procedural Knowledge Is Not Low-Rank: Why LoRA Fails to Internalize Multi-Step Procedures cites this paper.

Procedural Knowledge Is Not Low-Rank: Why LoRA Fails to Internalize Multi-Step Procedures FireAct: Toward Language Agent Fine-tuning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T13:17:10.198084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T13:17:10.198084Z digest=sha256:57cd91ea5099e28a581f762b44afe86bac3dd5809a4091913440f6a0262446a5

Observation 7bda8c0e-88b7-4aa2-b1f1-f9856096f671 · inbound

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG cites this paper.

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG FireAct: Toward Language Agent Fine-tuning

Reference 129

Resolution
unresolved
no resolver link, observed 2026-08-02T10:20:56.153236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T10:20:56.153236Z digest=sha256:b9747724b14d5ca3a30b4e7ba3c6761f5ac8934ad421567b8776960b6b36a93f

Observation 212faca4-28ac-4214-93c4-009490a9726b · inbound

ODYSSE: Episode-wise Policy Optimization for Personalized Agentic Reasoning cites this paper.

ODYSSE: Episode-wise Policy Optimization for Personalized Agentic Reasoning FireAct: Toward Language Agent Fine-tuning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T02:44:29.355941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:44:29.355941Z digest=sha256:9b4ee0a1d85b101593019da5e62aec7dfb40b8c7cd41e83609ff14e0fe8964c7

Observation 5d4c95c5-497b-4e2b-9371-b8e6e72dd8c4 · inbound

AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents cites this paper.

AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents FireAct: Toward Language Agent Fine-tuning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-30T14:30:12.562879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-30T14:30:12.562879Z digest=sha256:6051f0d3b30765de0b4de19e806b3a0132e60f2ce0699844281f42e6d8b14262

Observation 5191ba9d-6e22-40fb-8c00-8ee41facf7f3 · inbound

AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents cites this paper.

AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents FireAct: Toward Language Agent Fine-tuning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T04:29:02.267081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:29:02.267081Z digest=sha256:910441edb3127be41a94d669bc2c591e4461f3ed88d9e51eb7f55c4286fb179c

Observation b35807e4-fae5-419d-a086-825e7bb85ae3 · inbound

Leveraging Trajectory Graphs for Pre-Execution Error Diagnosis in Agentic LLM Systems cites this paper.

Leveraging Trajectory Graphs for Pre-Execution Error Diagnosis in Agentic LLM Systems FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-31T00:46:08.879772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T00:46:08.879772Z digest=sha256:b9a137c7568456c014e7112aa52286ff27903611a49ac65cff86014adf268ed6

Observation 6f20d570-efd7-4c93-ab90-7707693ad3a7 · inbound

MemHarness: Memory Is Reconstructed, Not Replayed cites this paper.

MemHarness: Memory Is Reconstructed, Not Replayed FireAct: Toward Language Agent Fine-tuning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-31T12:39:14.155269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T12:39:14.155269Z digest=sha256:51233f1b710678f1ec7031812fc668655668d364bce1b698a49f5373bcbc7595