Pith. sign in

Paper Citation Record · LEDGER

FireAct: Toward Language Agent Fine-tuning

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 86 inbound Pith citation observations for arXiv:2310.05915.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.05915 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 86 of 86 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:24:04.432770Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

6
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b87ab206-5067-4d51-9903-e4393ce36fea · inbound

Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security cites this paper.

Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security FireAct: Toward Language Agent Fine-tuning

Reference 247

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:57:26.697334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-17T00:57:26.303195Z digest=sha256:2ee0683b5c740ae7aea45252e9e82f77488dd721a03478dcc91bf315d3e2e383

Observation 3911caca-315b-48dc-ab0f-151a45476f58 · inbound

Training Agents with Weakly Supervised Feedback from Large Language Models cites this paper.

Training Agents with Weakly Supervised Feedback from Large Language Models FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T10:08:13.110427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T10:08:13.110427Z digest=sha256:c57d3d00b07cce6cca112571b28deb4579f3859ab15e393a75347becb89d771f

Observation 788502ce-303a-4407-b45f-274ef7c26a62 · inbound

Towards Adaptive Mechanism Activation in Language Agent cites this paper.

Towards Adaptive Mechanism Activation in Language Agent FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T05:09:47.707320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T05:09:47.707320Z digest=sha256:af63e0b0a5d3343557046d71f85f20b9cf75e092ba357d595ba0bc3f271fb0ad

Observation 85351a09-9f2b-4bdf-9146-2437a62c1183 · inbound

Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone cites this paper.

Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone FireAct: Toward Language Agent Fine-tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T19:28:37.798320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:28:37.798320Z digest=sha256:43be55520560f8c117ecddaaacec32c6c50ad36dc1ad5d135d976ac22777fd1b

Observation 73de4fac-6c3e-4ec7-9608-87cd9f5bfdd9 · inbound

Disentangling Reasoning Tokens and Boilerplate Tokens For Language Model Fine-tuning cites this paper.

Disentangling Reasoning Tokens and Boilerplate Tokens For Language Model Fine-tuning FireAct: Toward Language Agent Fine-tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T11:58:35.113146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:58:35.113146Z digest=sha256:4f7e5eb714c0136b6aae955a38d0648d518cfd6aacb21b38d8945356652257a0

Observation 640f5ca7-00db-45d7-a5df-331e3119812e · inbound

Think&Cite: Improving Attributed Text Generation with Self-Guided Tree Search and Progress Reward Modeling cites this paper.

Think&Cite: Improving Attributed Text Generation with Self-Guided Tree Search and Progress Reward Modeling FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T11:55:30.350423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:55:30.350423Z digest=sha256:e314603899f67e9f4fa5426a5a4c2f4ec6479fe1440915ad012e95ceed031ef3

Observation 5c29a3f8-f1d8-4bfc-9f81-68cd7864dc9e · inbound

Aviary: training language agents on challenging scientific tasks cites this paper.

Aviary: training language agents on challenging scientific tasks FireAct: Toward Language Agent Fine-tuning

Reference 113

Resolution
unresolved
no resolver link, observed 2026-08-10T23:07:33.921161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:07:33.921161Z digest=sha256:937d9ad3c1e2fcfe31ed41ea55f7cfd4df7846e75f827bc60d7f6e50001d784b

Observation 082b0d5e-bcf0-44bc-9c21-9346819e1817 · inbound

Learn-by-interact: A Data-Centric Framework for Self-Adaptive Agents in Realistic Environments cites this paper.

Learn-by-interact: A Data-Centric Framework for Self-Adaptive Agents in Realistic Environments FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T18:56:44.439851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T18:56:44.439851Z digest=sha256:1554a26b788926135b4c05e2b6754e58346f8a0419aef75f207ee484286b8d81

Observation 8e0aa012-322f-4fc4-8919-a50c581c0df7 · inbound

On Accelerating Edge AI: Optimizing Resource-Constrained Environments cites this paper.

On Accelerating Edge AI: Optimizing Resource-Constrained Environments FireAct: Toward Language Agent Fine-tuning

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-10T14:46:38.488034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T14:46:38.488034Z digest=sha256:6098749c937b762282eba13e62c88957521fcd96787f319b4d7969165fa84453

Observation 646dde47-932f-4ed0-aa43-f7185d9a3607 · inbound

Efficient Multi-Agent System Training with Data Influence-Oriented Tree Search cites this paper.

Efficient Multi-Agent System Training with Data Influence-Oriented Tree Search FireAct: Toward Language Agent Fine-tuning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:07:30.575127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-23T04:06:23.521344Z digest=sha256:4afa23f6cd11ebcd8439ec84bd3f048059c16e841bc1212daee25da8fa7ca7fc

Observation e48f99d2-9cf3-4d39-a050-74ea028dfc96 · inbound

Reinforcement Learning for Long-Horizon Interactive LLM Agents cites this paper.

Reinforcement Learning for Long-Horizon Interactive LLM Agents FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T14:56:00.193783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:56:00.193783Z digest=sha256:8a763ca0ba204beba016e523dab48c3b41edb7d1a198b74534048c09aa647074

Observation 6c105d71-7338-46c5-8849-cc86c692f7be · inbound

QLASS: Boosting Language Agent Inference via Q-Guided Stepwise Search cites this paper.

QLASS: Boosting Language Agent Inference via Q-Guided Stepwise Search FireAct: Toward Language Agent Fine-tuning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T11:47:31.226178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T11:47:31.226178Z digest=sha256:cbae27050e76b937da0b0616c0cf76b8a4439af5902131037586824ffb29bded

Observation 44a6eb19-835b-4422-8264-9a3c50ab297f · inbound

Learning Strategic Language Agents in the Werewolf Game with Iterative Latent Space Policy Optimization cites this paper.

Learning Strategic Language Agents in the Werewolf Game with Iterative Latent Space Policy Optimization FireAct: Toward Language Agent Fine-tuning

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-08T21:58:39.413359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:58:39.413359Z digest=sha256:cece11d21696bfa1efc558b5f8031697783bff0ba672625a13eac2a4da9a41c7

Observation 47799336-aff9-4f60-b8de-51c448c20082 · inbound

Hephaestus: Improving Fundamental Agent Capabilities of Large Language Models through Continual Pre-Training cites this paper.

Hephaestus: Improving Fundamental Agent Capabilities of Large Language Models through Continual Pre-Training FireAct: Toward Language Agent Fine-tuning

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:44.601401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:44.601401Z digest=sha256:5982d5dd60c688be17ab4590959c7dee06bd77fa5f9c58846d0030f532d33e27

Observation f718996a-ad8b-406c-8ffd-65ecf8b4eb47 · inbound

InSTA: Towards Internet-Scale Training For Agents cites this paper.

InSTA: Towards Internet-Scale Training For Agents FireAct: Toward Language Agent Fine-tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.313595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.313595Z digest=sha256:e4309b0a66fdcd10b4848e9ef052766cb899370356b3ece3893a87e814dd2d06

Observation 22b0482f-0f63-45dc-9219-3a0fb7b7248f · inbound

Digi-Q: Learning Q-Value Functions for Training Device-Control Agents cites this paper.

Digi-Q: Learning Q-Value Functions for Training Device-Control Agents FireAct: Toward Language Agent Fine-tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T20:57:48.380397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:57:48.380397Z digest=sha256:5bdd81ae12cdd01b22b4dad7eb90233fee2949a20e20dc321241b30389ceb4c5

Observation 0fbb9563-41b4-4cc8-96f9-c921b7047684 · inbound

InstructRAG: Leveraging Retrieval-Augmented Generation on Instruction Graphs for LLM-Based Task Planning cites this paper.

InstructRAG: Leveraging Retrieval-Augmented Generation on Instruction Graphs for LLM-Based Task Planning FireAct: Toward Language Agent Fine-tuning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T12:24:04.432770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:24:04.432770Z digest=sha256:dd4258f1266b78fbd6d180900876c000bca578ea6ca36df35d6d1c3b7056118b

Observation c5dd1a83-41b9-4a2f-97c9-555ac768a72b · inbound

Exploring Expert Failures Improves LLM Agent Tuning cites this paper.

Exploring Expert Failures Improves LLM Agent Tuning FireAct: Toward Language Agent Fine-tuning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-16T12:18:48.450541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:18:48.450541Z digest=sha256:966576e0ac9cb9ea68f5c74e2997ecab93422b6a69d9ede8c1c54a0e8bcf137e

Observation edcfd5ca-01dc-421b-9e57-e1fe127e13f3 · inbound

Optimization Problem Solving Can Transition to Evolutionary Agentic Workflows cites this paper.

Optimization Problem Solving Can Transition to Evolutionary Agentic Workflows FireAct: Toward Language Agent Fine-tuning

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-15T23:34:30.071160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:34:30.071160Z digest=sha256:99d07d0a6a27b9f9b61cec1004f975bcf4477ca404d7149db514f3e6c6ee2a74

Observation 3a52da57-1d21-4abd-889f-50538f316495 · inbound

Effective Reinforcement Learning for Reasoning in Language Models cites this paper.

Effective Reinforcement Learning for Reasoning in Language Models FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:56:15.384955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:56:15.384955Z digest=sha256:deaacc9d16e0a6a18747fc27c09f1860bded2a9d3ac945ced7430eb554c5e63a

Observation 9d59dde0-aa66-4377-b779-ff008ad7da81 · inbound

Large Language Models for Planning: A Comprehensive and Systematic Survey cites this paper.

Large Language Models for Planning: A Comprehensive and Systematic Survey FireAct: Toward Language Agent Fine-tuning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:11:51.825888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:11:51.825888Z digest=sha256:17186ead7fd98dc13b323293213b14e3debbb55673d1afb3ae066310e750b4b4

Observation 4e07b629-aa8e-4ecc-a5f4-db57fc4ab08d · inbound

Divide and Conquer: Grounding LLMs as Efficient Decision-Making Agents via Offline Hierarchical Reinforcement Learning cites this paper.

Divide and Conquer: Grounding LLMs as Efficient Decision-Making Agents via Offline Hierarchical Reinforcement Learning FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:10:02.785697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:10:02.785697Z digest=sha256:22705a4db4d355c1fe13f1d2e6761d3272edb44785d203bab6495ec9828e0ab0

Observation 24ba30b1-3a1a-4744-9b32-ffb030550847 · inbound

Training LLM-Based Agents with Synthetic Self-Reflected Trajectories and Partial Masking cites this paper.

Training LLM-Based Agents with Synthetic Self-Reflected Trajectories and Partial Masking FireAct: Toward Language Agent Fine-tuning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:53.740452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:06:53.740452Z digest=sha256:b9cedf13ec3e73e566fb97caadc48391a57f2b65774e1461daacb420f14f452d

Observation acce814c-6101-497e-8130-b2be289ae051 · inbound

SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution cites this paper.

SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution FireAct: Toward Language Agent Fine-tuning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:02.795247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:02.795247Z digest=sha256:eb541dccfda0f24425f86726cd3c8f0e2da35365a22128c8725d9402dacf79f1

Observation 34e23431-e069-445d-90ab-5951361985e9 · inbound

RRO: LLM Agent Optimization Through Rising Reward Trajectories cites this paper.

RRO: LLM Agent Optimization Through Rising Reward Trajectories FireAct: Toward Language Agent Fine-tuning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:51:41.263415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:51:41.263415Z digest=sha256:472625aa75c74cf8d731845e7671b9aa4c07c3a54b8f7edb312f7901fe5b09c4

Observation e42a0d3b-5c02-40f9-8b20-fe1b05831de5 · inbound

Agent-Environment Alignment via Automated Interface Generation cites this paper.

Agent-Environment Alignment via Automated Interface Generation FireAct: Toward Language Agent Fine-tuning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:30.856512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:45:30.856512Z digest=sha256:cb214f7cc58e048980f627335dd59364ace28a3475e5e36170b37c30db5e248b

Observation 05a885bf-085a-4741-89e4-742cb11aae9e · inbound

WebDancer: Towards Autonomous Information Seeking Agency cites this paper.

WebDancer: Towards Autonomous Information Seeking Agency FireAct: Toward Language Agent Fine-tuning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:59.724999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:59.724999Z digest=sha256:c84d2526bf3d68a22cd0a53ef945dcd6468d82600ff2247b9b8ea2d1399b5c2a

Observation da351dd0-93da-4672-b144-ced5533be68a · inbound

OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation cites this paper.

OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation FireAct: Toward Language Agent Fine-tuning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:31.904774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:31.904774Z digest=sha256:f9481d1684d489bd58484d4d9620dba833b713bb4608691ede6419fd88d16344

Observation 622a48e0-de72-472a-a0c6-f43cbd5a0d5a · inbound

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization cites this paper.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.567143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:00.567143Z digest=sha256:73f5d54d627e1defe2665c61ff4308a7b374ead5457e50e8211deefe1634b9a7

Observation 7722ca1e-4be8-49b1-b8d0-65aaa2cb1e06 · inbound

LAM SIMULATOR: Advancing Data Generation for Large Action Model Training via Online Exploration and Trajectory Feedback cites this paper.

LAM SIMULATOR: Advancing Data Generation for Large Action Model Training via Online Exploration and Trajectory Feedback FireAct: Toward Language Agent Fine-tuning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:54.902423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:54.902423Z digest=sha256:fae011799a243a0362a0492ca1250ba96872e43b8121d3e8b6f4828e057b35b6

Observation 5753f9fc-d847-4e26-a0c5-593e07aa9abe · inbound

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games cites this paper.

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games FireAct: Toward Language Agent Fine-tuning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-19T12:02:16.630032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T12:01:42.681135Z digest=sha256:80087ed7216fa052c808a658ac93b056fa3f6d75bde0e7bffe47354b84b8a28a

Observation 1ac41ca7-a31c-4ac2-885e-2f056241fcbe · inbound

Automated Skill Discovery for Language Agents through Exploration and Iterative Feedback cites this paper.

Automated Skill Discovery for Language Agents through Exploration and Iterative Feedback FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:25.898272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:25.898272Z digest=sha256:f5877bf1a3a83441420598f210febe5f6af3932bb65d1e2f552b1ccfa5390adf

Observation cfbb0d56-e472-4344-ab30-6a7a09779e7c · inbound

Agent-RewardBench: Towards a Unified Benchmark for Reward Modeling across Perception, Planning, and Safety in Real-World Multimodal Agents cites this paper.

Agent-RewardBench: Towards a Unified Benchmark for Reward Modeling across Perception, Planning, and Safety in Real-World Multimodal Agents FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T22:36:42.146724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:36:42.146724Z digest=sha256:e6be0af250174d3ea1e11329c3ba8fd02035de4f27f88170af7cab90772e4088

Observation 57d5bf82-b0ed-40e6-a1f0-041237f1d47e · inbound

Knowledge Augmented Finetuning Matters in both RAG and Agent Based Dialog Systems cites this paper.

Knowledge Augmented Finetuning Matters in both RAG and Agent Based Dialog Systems FireAct: Toward Language Agent Fine-tuning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:59:59.720213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:59:59.720213Z digest=sha256:b8aacbb1ae5ace7992edd03f9c86a30caf89135bcf3469e4412137b2326c4513

Observation 92f03b00-374b-4a14-9c7c-068dcff9429b · inbound

Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning cites this paper.

Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:00.459541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:55:00.459541Z digest=sha256:45a3008c3ed61dbc73ffbbd6ffca3b15ab1768d0b898a1338b69a6cb9910c87a

Observation dfcb3673-566d-4a51-bd5f-6caf64e3b635 · inbound

WebSailor: Navigating Super-human Reasoning for Web Agent cites this paper.

WebSailor: Navigating Super-human Reasoning for Web Agent FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:37:09.609424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-17T15:37:09.572241Z digest=sha256:65b454f132ea3a0d83b883a9ca0a8628343ba113540da335f5b8e363db4c343e

Observation 4ece5b4b-0d92-46bc-a299-ae61de169c11 · inbound

SAND: Boosting LLM Agents with Self-Taught Action Deliberation cites this paper.

SAND: Boosting LLM Agents with Self-Taught Action Deliberation FireAct: Toward Language Agent Fine-tuning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:47:19.417248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:47:19.417248Z digest=sha256:e33c03dad186def4bb50d9eb16bdbb875568cd5c93ff53d5323a57150fb1fd88

Observation a6f50084-356e-42ed-b398-20cf7978c4cf · inbound

Initial Steps in Integrating Large Reasoning and Action Models for Service Composition cites this paper.

Initial Steps in Integrating Large Reasoning and Action Models for Service Composition FireAct: Toward Language Agent Fine-tuning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T14:33:53.491820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:33:53.491820Z digest=sha256:bd80b3cdf0a7de4fc68620ec7c53775bb66906c02d9cd257e29a61e42ddb6933

Observation 3baeeb8c-e4ef-4928-8cff-4e32bae54d62 · inbound

MMAT-1M: A Large Reasoning Dataset for Multimodal Agent Tuning cites this paper.

MMAT-1M: A Large Reasoning Dataset for Multimodal Agent Tuning FireAct: Toward Language Agent Fine-tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T12:18:31.887391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:18:31.887391Z digest=sha256:2405fac82f85c0eed8f22094bb03b3516d017defa36c116f3ecae9967e1530b1

Observation 7d149853-178e-4757-a71c-fdf5f94f7ea8 · inbound

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey cites this paper.

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey FireAct: Toward Language Agent Fine-tuning

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-18T19:21:48.752794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T19:19:36.427337Z digest=sha256:1b084f62a6d25e6c686c414823adeb379b25d5c69d42f26b57ad7d71aa2b8146

Observation f45c9255-87da-4228-b342-8b599bef5c8f · inbound

AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning cites this paper.

AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning FireAct: Toward Language Agent Fine-tuning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T16:08:37.751164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:08:37.751164Z digest=sha256:e58048af497fa9194c302d7f7e636d1b8725c8d1ad80dfb0ff4dd649fcbad92d

Observation 233947b3-af51-43fb-aa22-d60d4dd230e7 · inbound

Bridging the Capability Gap: Joint Alignment Tuning for Harmonizing LLM-based Multi-Agent Systems cites this paper.

Bridging the Capability Gap: Joint Alignment Tuning for Harmonizing LLM-based Multi-Agent Systems FireAct: Toward Language Agent Fine-tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T18:50:10.443190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T18:50:10.443190Z digest=sha256:89d6dfac4dbfca284f5ab24206fe4c19612c37ad4565d924cbbbc7b2e7c51e1c

Observation c363136c-8b56-418e-a591-d3f89be6daf9 · inbound

On-Device Fine-Tuning via Backprop-Free Zeroth-Order Optimization cites this paper.

On-Device Fine-Tuning via Backprop-Free Zeroth-Order Optimization FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:05:21.475359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-17T22:03:53.594703Z digest=sha256:307bca94374406d3f33fae1bff70eae9598069b81d94abf06b4ad963d7de5185

Observation a703daeb-00fa-40df-9da0-698878cdb6e4 · inbound

MemVerse: Multimodal Memory for Lifelong Learning Agents cites this paper.

MemVerse: Multimodal Memory for Lifelong Learning Agents FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T18:48:57.657515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:48:57.657515Z digest=sha256:bf81e77f6b0cc9638c39c569cbea70a571862a53c068be9b7d5d6e5df81d4902

Observation 701c787b-a889-4547-acf4-0b55d6b846b8 · inbound

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models cites this paper.

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models FireAct: Toward Language Agent Fine-tuning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T11:44:05.540078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T11:44:05.540078Z digest=sha256:be29a502884c9a36bd3b43bfba6724689d7e4e2132b2e2a9e5bd236a8fff76d6

Observation 215a4d4b-b014-49a7-b769-75b812a26f03 · inbound

SoK: Agentic Skills -- Beyond Tool Use in LLM Agents cites this paper.

SoK: Agentic Skills -- Beyond Tool Use in LLM Agents FireAct: Toward Language Agent Fine-tuning

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:19:31.175043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-14T23:19:31.024268Z digest=sha256:deaf8769889556ed237f5582f9f14f5f7f303d2d16c46d2ff0610c7aeb95f373

Observation a2334525-f988-410c-bc05-b0a5cb16df15 · inbound

ForkKV: Scaling Multi-LoRA Agent Serving via Copy-on-Write Disaggregated KV Cache cites this paper.

ForkKV: Scaling Multi-LoRA Agent Serving via Copy-on-Write Disaggregated KV Cache FireAct: Toward Language Agent Fine-tuning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:10:54.920774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T18:16:49.292491Z digest=sha256:63a1a00fc98f6ba1a234cb1f28b7e69870b7f68573d550f9d4bb77e8cdba5fe2

Observation ebdc028a-c5d4-46cc-9c4e-61c78c258e34 · inbound

PASK: Toward Intent-Aware Proactive Agents with Long-Term Memory cites this paper.

PASK: Toward Intent-Aware Proactive Agents with Long-Term Memory FireAct: Toward Language Agent Fine-tuning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:30:59.754002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T18:05:40.242925Z digest=sha256:74f98aca155e911d60e6e26fd323ebbb869642a56bd99d5e5b045ee2c06557fd

Observation b0613023-41f0-471c-990b-2b4ecff093f3 · inbound

From Coarse to Fine: Self-Adaptive Hierarchical Planning for LLM Agents cites this paper.

From Coarse to Fine: Self-Adaptive Hierarchical Planning for LLM Agents FireAct: Toward Language Agent Fine-tuning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:46:10.623340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-08T08:10:36.579810Z digest=sha256:7532278d81bbe121031f418d42f681aaecc185410fd3a644c1b9b5b6f6154f40

Observation d76db5ba-0679-41f7-aad1-bd04ba976c6a · inbound

Rethinking Agentic Reinforcement Learning In Large Language Models cites this paper.

Rethinking Agentic Reinforcement Learning In Large Language Models FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:21:29.906759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-07T06:30:09.945371Z digest=sha256:443c9900c4603686dd52cba3898e2854db0b925657d6ead4af674f35247d7456

Observation 0cfc38ba-4da7-40ec-b7a5-cd2a40a62e60 · inbound

Rethinking Agentic Reinforcement Learning In Large Language Models cites this paper.

Rethinking Agentic Reinforcement Learning In Large Language Models FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:11:17.132850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-08T03:12:19.414358Z digest=sha256:e72269753ce62e4d3495204bb245bc7d24639d232a09ff4b245ac98c7387dd65

Observation 2816bf74-7e82-4ed4-bfc5-c5539c5ab24a · inbound

Rethinking Agentic Reinforcement Learning In Large Language Models cites this paper.

Rethinking Agentic Reinforcement Learning In Large Language Models FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-19T17:02:40.660252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T16:58:41.558250Z digest=sha256:8c147d25ce8710a09ace706a5d748912f4f22ca618815b6e499f07d805caabe5

Observation 414876f4-f394-4d5f-ae80-fb49c6d560d0 · inbound

SOD: Step-wise On-policy Distillation for Small Language Model Agents cites this paper.

SOD: Step-wise On-policy Distillation for Small Language Model Agents FireAct: Toward Language Agent Fine-tuning

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:40:54.506607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-11T02:25:59.056181Z digest=sha256:413b6dcdf132f2a4b109d7acebe5a0ff84cfa6f5c55c40b865ff251d94169f03

Observation a76083b1-83e4-42ea-9057-87b5a5a12829 · inbound

SOD: Step-wise On-policy Distillation for Small Language Model Agents cites this paper.

SOD: Step-wise On-policy Distillation for Small Language Model Agents FireAct: Toward Language Agent Fine-tuning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T05:20:45.134981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:20:45.134981Z digest=sha256:4c5c3356bb2b65a380a35bef74ecaacd4a751aa9e211ff165d86e027d2ad92de

Observation c713480d-1da4-4114-941f-bc60dfa7ca11 · inbound

Learning CLI Agents with Structured Action Credit under Selective Observation cites this paper.

Learning CLI Agents with Structured Action Credit under Selective Observation FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:00:56.420074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-11T02:59:26.100818Z digest=sha256:d9a165825a126f8e9f06b979b47c31cfd08c02c7162638b927df0b8bd98933cf

Observation e71c9f61-b55d-40a9-ad30-20bcc08b828b · inbound

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation cites this paper.

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation FireAct: Toward Language Agent Fine-tuning

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:11:27.660465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-12T03:36:24.941205Z digest=sha256:41a850f602974a30069f87fdb61c0cc66fcc476fd1bd85e8cee14d883a9e185d

Observation 3701cabe-b463-4441-948b-0692179f4b81 · inbound

Verifiable Process Rewards for Agentic Reasoning cites this paper.

Verifiable Process Rewards for Agentic Reasoning FireAct: Toward Language Agent Fine-tuning

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:46:27.390166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-12T04:59:44.082011Z digest=sha256:2569181c38492abede93140ceeca88a6299c83629fe5facb048ec54b6f20253f

Observation 972a20b1-7617-496f-a8cf-7bf7c87c0396 · inbound

Verifiable Process Rewards for Agentic Reasoning cites this paper.

Verifiable Process Rewards for Agentic Reasoning FireAct: Toward Language Agent Fine-tuning

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-30T22:45:07.256818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T22:43:50.317854Z digest=sha256:22955cd3582e483dbe497ac0ba13a888b20e25f697a3dc538d3728906ebdd42a

Observation 50bf99b1-d234-4603-9715-a2fd78b34f04 · inbound

SkillGen: Verified Inference-Time Agent Skill Synthesis cites this paper.

SkillGen: Verified Inference-Time Agent Skill Synthesis FireAct: Toward Language Agent Fine-tuning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:17:23.043506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T06:14:28.614825Z digest=sha256:11d290f247f621e4fe2873fc0af66d15278ce7bbf57e99e9df6ee92115d8d4f9

Observation d945513a-b9ff-4c3d-8ce0-3cf7ca4ee900 · inbound

ClawForge: Generating Executable Interactive Benchmarks for Command-Line Agents cites this paper.

ClawForge: Generating Executable Interactive Benchmarks for Command-Line Agents FireAct: Toward Language Agent Fine-tuning

Reference 87

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T05:09:45.136629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-15T05:08:15.750558Z digest=sha256:77fb9f015e5f550b1f93fdb1d8e209eed286e8d0e2380bd50a82f53feedc80fd

Observation 03da1012-3696-4a93-90a2-e6fa14d1f4d9 · inbound

ClawForge: Generating Executable Interactive Benchmarks for Command-Line Agents cites this paper.

ClawForge: Generating Executable Interactive Benchmarks for Command-Line Agents FireAct: Toward Language Agent Fine-tuning

Reference 87

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T20:23:43.246128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-20T20:19:21.824216Z digest=sha256:e5728c6b8fa2d0f6aa6417a60b1ce31a5aa6b2ae4c89006478754f741ba1e38a

Observation 7c719705-040b-4d02-826f-e8d9c0595458 · inbound

TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination cites this paper.

TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination FireAct: Toward Language Agent Fine-tuning

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T18:02:42.288078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-19T18:01:06.649723Z digest=sha256:691bf33709b96b737acd1c0f18465ba1a22a36117cae502d899c54040c1bb6d5

Observation 37d965ed-ee2c-4c4d-87e8-d9aac2c260eb · inbound

AMATA: Adaptive Multi-Agent Trajectory Alignment for Knowledge-Intensive Question Answering cites this paper.

AMATA: Adaptive Multi-Agent Trajectory Alignment for Knowledge-Intensive Question Answering FireAct: Toward Language Agent Fine-tuning

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:48:19.927143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-20T13:43:42.097447Z digest=sha256:fef163f04b3242c1c6e95c625826fb52839979d0254b5798c04d5ed2f0661216

Observation bf1dca57-c7ec-4bf1-a977-a4b09a47e521 · inbound

Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost cites this paper.

Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost FireAct: Toward Language Agent Fine-tuning

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T06:31:10.437580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-22T06:26:31.841138Z digest=sha256:56eb88457e6a883e5c90a32e20f665e579bdd0b440d256ee81f239c87c472868

Observation 87fd8886-651e-4c1b-9b21-1a64611edfbf · inbound

Test-Time Deep Thinking to Explore Implicit Rules cites this paper.

Test-Time Deep Thinking to Explore Implicit Rules FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:54:38.220238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T11:52:15.163893Z digest=sha256:0305a9110ecbb991d7f062d75474409dad9bb55e934bbe2349e2142c6ed3f56c

Observation 38045e09-7dbd-448d-89d7-bdbc8e7d98c3 · inbound

Agent Explorative Policy Optimization for Multimodal Agentic Reasoning cites this paper.

Agent Explorative Policy Optimization for Multimodal Agentic Reasoning FireAct: Toward Language Agent Fine-tuning

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:23:23.735371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T12:22:39.655615Z digest=sha256:bb7a1900824095ab93fca4a4e49f8c3569d9937163bc0f5bc1f04dd927272340

Observation 64d64355-8c2a-4094-8ea1-bf6143d4dfe2 · inbound

COMAP: Co-Evolving World Models and Agent Policies for LLM Agents cites this paper.

COMAP: Co-Evolving World Models and Agent Policies for LLM Agents FireAct: Toward Language Agent Fine-tuning

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:16:24.848399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-28T14:27:50.260308Z digest=sha256:3b4005f62a86a17c058884b5e6507f49af0bc880666637dea24dc3f2256e037d

Observation 71ae13f4-84af-4eac-86d7-6e31862d9f7f · inbound

SaliMory: Orchestrating Cognitive Memory for Conversational Agents cites this paper.

SaliMory: Orchestrating Cognitive Memory for Conversational Agents FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:16:31.723007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T10:09:26.995556Z digest=sha256:3cef97fb307cece5130b215a206576278b9bc0e7e52583e7fa42bbab411a07a4

Observation ee620fc9-0804-4521-bc91-bd3ab31fbd65 · inbound

Self-evolving LLM agents with in-distribution Optimization cites this paper.

Self-evolving LLM agents with in-distribution Optimization FireAct: Toward Language Agent Fine-tuning

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:57:09.540367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-27T22:18:27.021136Z digest=sha256:7515d7b36570ac4c6c83ece9667a16b93e0131725e28b663bc60ce52bb8440ee

Observation 93edcb76-e681-48db-94cb-a54fbdee09f9 · inbound

Evoflux: Inference-Time Evolution of Executable Tool Workflows for Compact Agents cites this paper.

Evoflux: Inference-Time Evolution of Executable Tool Workflows for Compact Agents FireAct: Toward Language Agent Fine-tuning

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-06-27T09:50:48.183424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-27T09:46:59.954720Z digest=sha256:9b4bd75e1172ebf92c1a6d4c655d0ce46e21c4693ddce4c7aec627ce8cc1c132

Observation d2f72c89-7104-4a4b-a26a-624139ce3b11 · inbound

Training the Orchestrator: A Supervised Approach to End-to-End PDDL Planning with LLM Agents cites this paper.

Training the Orchestrator: A Supervised Approach to End-to-End PDDL Planning with LLM Agents FireAct: Toward Language Agent Fine-tuning

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:59:38.128661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-26T13:56:51.914966Z digest=sha256:007477cb14a54d9add091ae2ed32041f1859c44b66f403dde0d0e9796b03cee0

Observation 698a2d61-ed51-475a-b800-89780383ac57 · inbound

AgentOdyssey: Open-Ended Long-Horizon Text Game Generation for Test-Time Continual Learning Agents cites this paper.

AgentOdyssey: Open-Ended Long-Horizon Text Game Generation for Test-Time Continual Learning Agents FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:46:11.465290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T21:59:25.449570Z digest=sha256:987d6ffa5dfe760a68708358988789abf4f6abbb1adf15550dd96c67f6a965ed

Observation b5074fd1-6551-4c79-8f50-a25b58c0258c · inbound

Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks cites this paper.

Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks FireAct: Toward Language Agent Fine-tuning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:49:52.879268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T05:51:01.446139Z digest=sha256:257d6198583846537c092f3153a8ef66d7e1aec1ed7e61b737fe260694a0645c

Observation 8096747b-5fb9-4f05-8ddc-2319995977ba · inbound

Agentic-Ideation: Sample Efficient Agentic Trajectories Synthesis for Scientific Ideation Agents cites this paper.

Agentic-Ideation: Sample Efficient Agentic Trajectories Synthesis for Scientific Ideation Agents FireAct: Toward Language Agent Fine-tuning

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:05:41.381986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-01T05:49:52.979376Z digest=sha256:68f7c2c1b9f1b0c0dd4be6924a89efb7b3c47c0397ef3f96d8df2bb00214de79

Observation 2444a382-c0b1-40c2-900f-f5c24b58fa44 · inbound

Atomic Task Graph: A Unified Framework for Agentic Planning and Execution cites this paper.

Atomic Task Graph: A Unified Framework for Agentic Planning and Execution FireAct: Toward Language Agent Fine-tuning

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:08:21.353187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-03T13:59:31.697155Z digest=sha256:02cbd06b1f2361d24e984248cd2ae3499993e013e7418b9439a88f829c990643

Observation 56cecc91-1805-4e49-912f-9797b1ea6101 · inbound

STAPO: Selective Trajectory-Aware Policy Optimization for LLM Agent Training cites this paper.

STAPO: Selective Trajectory-Aware Policy Optimization for LLM Agent Training FireAct: Toward Language Agent Fine-tuning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-11T10:50:54.419477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T10:50:54.419477Z digest=sha256:1a9ab887c7c015c73a670300cabbfc56297c38567608a90f07a7e51ec34e1711

Observation 0d8d235e-bea8-4cda-88cc-f30d8bc5c0a4 · inbound

Eluna: An Agentic LLM System for Automating Warehouse Operations with Reasoning and Task Execution cites this paper.

Eluna: An Agentic LLM System for Automating Warehouse Operations with Reasoning and Task Execution FireAct: Toward Language Agent Fine-tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-13T05:30:10.497783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T05:30:10.497783Z digest=sha256:57cc1bde4a811ee261ad5feae0fa0fb7f55de2f318bb48df411dae0471b3db60

Observation fadb3a24-6dea-449d-8890-a67f0292a037 · inbound

Agentic-DPO: From Imitation to Agentic Policy Optimization on Expert Trajectories cites this paper.

Agentic-DPO: From Imitation to Agentic Policy Optimization on Expert Trajectories FireAct: Toward Language Agent Fine-tuning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T10:33:54.851493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T10:33:54.851493Z digest=sha256:44687cc8aee5de44675bf64224583ef97315f949534af7e77aa757f190a52c05

Observation 0812fa76-c88c-4290-9848-71d20c002f2b · inbound

From Outcomes to Actions: Leveraging Hindsight for Long-Horizon Language Agent Training cites this paper.

From Outcomes to Actions: Leveraging Hindsight for Long-Horizon Language Agent Training FireAct: Toward Language Agent Fine-tuning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T09:45:49.534257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:45:49.534257Z digest=sha256:982e55c96d04e9aef31119dd4ed76e29484b47e59bbaf7031d2f7d690a33a232

Observation 8650e116-c765-4a9c-bdf2-a1b5bc534758 · inbound

Procedural Knowledge Is Not Low-Rank: Why LoRA Fails to Internalize Multi-Step Procedures cites this paper.

Procedural Knowledge Is Not Low-Rank: Why LoRA Fails to Internalize Multi-Step Procedures FireAct: Toward Language Agent Fine-tuning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T13:17:10.198084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T13:17:10.198084Z digest=sha256:57d5c9ffca2241e7f5c202425c17cb5cabcb22a292d8e36d417fa8967d2bd5e6

Observation 7bda8c0e-88b7-4aa2-b1f1-f9856096f671 · inbound

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG cites this paper.

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG FireAct: Toward Language Agent Fine-tuning

Reference 129

Resolution
unresolved
no resolver link, observed 2026-08-02T10:20:56.153236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T10:20:56.153236Z digest=sha256:b5d34984ffe4045697c3ce3c066333403b4fa675df715123447a610c7448e589

Observation 212faca4-28ac-4214-93c4-009490a9726b · inbound

ODYSSE: Episode-wise Policy Optimization for Personalized Agentic Reasoning cites this paper.

ODYSSE: Episode-wise Policy Optimization for Personalized Agentic Reasoning FireAct: Toward Language Agent Fine-tuning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T02:44:29.355941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:44:29.355941Z digest=sha256:eb5a61521c3f674f8a072ba1fcf7b89834993c909bc776240c8dad3e1a0f2157

Observation 5d4c95c5-497b-4e2b-9371-b8e6e72dd8c4 · inbound

AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents cites this paper.

AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents FireAct: Toward Language Agent Fine-tuning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-30T14:30:12.562879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-30T14:30:12.562879Z digest=sha256:270be8704c57f78a6f0654e2df2c9a29bf7bf7657133e1f338ed4efb710311c4

Observation 5191ba9d-6e22-40fb-8c00-8ee41facf7f3 · inbound

AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents cites this paper.

AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents FireAct: Toward Language Agent Fine-tuning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T04:29:02.267081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:29:02.267081Z digest=sha256:7c9cd467ebdfb8b86429df6531c6cfcbb2be4c39d4f1ae247501880e65f4030a

Observation b35807e4-fae5-419d-a086-825e7bb85ae3 · inbound

Leveraging Trajectory Graphs for Pre-Execution Error Diagnosis in Agentic LLM Systems cites this paper.

Leveraging Trajectory Graphs for Pre-Execution Error Diagnosis in Agentic LLM Systems FireAct: Toward Language Agent Fine-tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-31T00:46:08.879772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T00:46:08.879772Z digest=sha256:cf133837f5f49f745733e0766a455854eb67282296f21589970f41dab50ec68b

Observation 6f20d570-efd7-4c93-ab90-7707693ad3a7 · inbound

MemHarness: Memory Is Reconstructed, Not Replayed cites this paper.

MemHarness: Memory Is Reconstructed, Not Replayed FireAct: Toward Language Agent Fine-tuning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-31T12:39:14.155269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T12:39:14.155269Z digest=sha256:874d06644a9d495eccb1a57f687dcaa3a9b3f7104514981b73b400e7532d6136