Pith. sign in

Paper Citation Record · LEDGER

AgentRefine: Enhancing Agent Generalization through Refinement Tuning

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2501.01702.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.01702 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:11:55.463221Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:48:56.423902Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6d14c2c3-a659-4623-831a-f5bc105149cd · inbound

Large Language Models for Planning: A Comprehensive and Systematic Survey cites this paper.

Large Language Models for Planning: A Comprehensive and Systematic Survey AgentRefine: Enhancing Agent Generalization through Refinement Tuning

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T14:11:55.463221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:11:55.463221Z digest=sha256:bdda730abbfe01cd819096f55f6a64d9c22400835e5ea91475268e17eb6d7746

Observation 644eca4f-dee4-4cdd-a7f4-01188eee14f7 · inbound

Agent-Environment Alignment via Automated Interface Generation cites this paper.

Agent-Environment Alignment via Automated Interface Generation AgentRefine: Enhancing Agent Generalization through Refinement Tuning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:31.292672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:45:31.292672Z digest=sha256:0533d3e40be2b0f5f7f676454e38a643d1e673ffec83150d520879430de6aea5

Observation c9a98375-1970-4930-a02b-426f3af5ff02 · inbound

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization cites this paper.

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization AgentRefine: Enhancing Agent Generalization through Refinement Tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.667874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:47:00.667874Z digest=sha256:73c6f26d6e81fa0ca5d2d3f4d9f96aa144358f02c067591c79c3bc91557a4060

Observation 6a5dedfc-061a-41d8-b712-4157991f9fe6 · inbound

RLVMR: Reinforcement Learning with Verifiable Meta-Reasoning Rewards for Robust Long-Horizon Agents cites this paper.

RLVMR: Reinforcement Learning with Verifiable Meta-Reasoning Rewards for Robust Long-Horizon Agents AgentRefine: Enhancing Agent Generalization through Refinement Tuning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T11:21:25.444678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:21:25.444678Z digest=sha256:4428ed326600e9d7d8fa5dbe935cea99bd530c49450e567423f80381a8ef259d

Observation 7c2cb04b-f4bc-4af7-95a3-bae0d1e053a4 · inbound

Source Component Shift Adaptation via Offline Decomposition and Online Mixing Approach cites this paper.

Source Component Shift Adaptation via Offline Decomposition and Online Mixing Approach AgentRefine: Enhancing Agent Generalization through Refinement Tuning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T20:35:42.954633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:35:42.954633Z digest=sha256:90d55ac780c5b4a24df4be94957f871fd4538edfe351f734643789b78962923f

Observation d8d751b2-fc5c-4fbf-8159-3cefec194abb · inbound

Leveraging OS-Level Primitives for Robotic Action Management cites this paper.

Leveraging OS-Level Primitives for Robotic Action Management AgentRefine: Enhancing Agent Generalization through Refinement Tuning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T20:38:39.405045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:38:39.405045Z digest=sha256:e9b03227eb6b6769f11db0c015a25756c8ca6621f3357801a2482557cc804d34

Observation c975513a-09f3-49bd-a2fe-3067633210df · inbound

From History to State: Constant-Context Skill Learning for LLM Agents cites this paper.

From History to State: Constant-Context Skill Learning for LLM Agents AgentRefine: Enhancing Agent Generalization through Refinement Tuning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:01:05.502490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T16:50:43.547830Z digest=sha256:34e3fd989064f1e2c7e6f68b9144e36e215730e64440d10fca5b148a5a5a95be

Observation 072174d9-9e44-4abc-9638-4ea9a8a01b69 · inbound

Test-Time Deep Thinking to Explore Implicit Rules cites this paper.

Test-Time Deep Thinking to Explore Implicit Rules AgentRefine: Enhancing Agent Generalization through Refinement Tuning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:54:38.208867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T11:52:15.163893Z digest=sha256:76c805a3c4a49f4131fa14f4ae25930c9ff4ee4606a5404fa67a29ad3ee468a9

Observation 376eb1ba-609a-495e-b4b3-86f7a7c16cbe · inbound

OPD-Evolver: Cultivating Holistic Agent Evolver via On-Policy Distillation cites this paper.

OPD-Evolver: Cultivating Holistic Agent Evolver via On-Policy Distillation AgentRefine: Enhancing Agent Generalization through Refinement Tuning

Reference 118

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:48:56.425501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T01:07:49.603969Z digest=sha256:c8d17c79b73f6299801bc1b0842df4077a79c11f8292f0c441c2d55fb981f30d

Observation ddddbc5f-1585-4c36-8a98-841c2807c6af · inbound

Agentic-DPO: From Imitation to Agentic Policy Optimization on Expert Trajectories cites this paper.

Agentic-DPO: From Imitation to Agentic Policy Optimization on Expert Trajectories AgentRefine: Enhancing Agent Generalization through Refinement Tuning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-14T10:33:54.851493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T10:33:54.851493Z digest=sha256:63fcd97098771292089f7444d6c0295fe6cace1dd8a85a6ee041cf4103cf1dca

Observation 6af45b6f-4368-4aab-b860-146e11fd5df8 · inbound

Beyond Action Imitation: Learning a Decision-Aware User Simulator for Online Advertising cites this paper.

Beyond Action Imitation: Learning a Decision-Aware User Simulator for Online Advertising AgentRefine: Enhancing Agent Generalization through Refinement Tuning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-30T17:56:01.404123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T17:56:01.404123Z digest=sha256:d41a4c0dd11b3afb04085def40728bcb4233be14cf37828e69e2f72726e86c03