Pith. sign in

Paper Citation Record · LEDGER

Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2308.02151.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.02151 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:24:14.981510Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T21:28:58.626281Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f8659ef3-82dc-41a9-a27b-7aafb6eb37c1 · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 234

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:28:39.056161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:59508e1b72013160b6e1d84ff02cafcea3b682f4e726fad3501186956278c095

Observation c960f693-2689-47f1-85d4-6dd2a9dec05e · inbound

A Survey on the Memory Mechanism of Large Language Model based Agents cites this paper.

A Survey on the Memory Mechanism of Large Language Model based Agents Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:21:39.992190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T07:21:39.440092Z digest=sha256:43e62dffea1ca29c9538a1023dc35f868cd6fa028a8b137c9f5ef23fc0280847

Observation 6682630d-db10-425d-8e0a-61d2a7e13492 · inbound

From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMs cites this paper.

From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMs Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-17T11:05:09.813162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T11:05:09.588491Z digest=sha256:dce4c9b9a8bc82e3a1c7a1686d147fb9e2bc5d578bee75414146a797897b1b85

Observation 498ff026-fe3f-4f37-bd26-0a457ef52ece · inbound

Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team cites this paper.

Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T00:24:14.981510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:24:14.981510Z digest=sha256:c45e57ef45c9e31d799eb274d637add5397bd4d4d2418d0c7f9d4ffa8f2a2688

Observation d97be4d3-3c47-4cd2-b2ba-ae416b0d23cf · inbound

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning cites this paper.

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:06:08.681011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-08T10:23:52.522238Z digest=sha256:22022adefae5076d4fd478f1f774a5e8ac449c51ae4a6deed4a9b034bf0aee94

Observation fd45479c-89ef-47c3-8787-9f632c461736 · inbound

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning cites this paper.

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:00:56.846868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-11T02:00:00.663355Z digest=sha256:bf04411aa7350ed86a0357b59e7dff7afedf751e0d2e1875d6b094371b7228b7

Observation 2257a956-2335-440d-96e2-ff68c3f54f44 · inbound

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning cites this paper.

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:17:28.208672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T07:17:13.708752Z digest=sha256:aa787681c62ecf579f9c64465236a88a92fcf1c2a4f57b17764024611fa70dd9

Observation e9983242-aa4a-4fd2-b415-f426ffd395d4 · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:20:57.093072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T01:47:39.926540Z digest=sha256:df61d423f8333cc891d2c8eda4846fcd1389f31b3bd528e6d7eb05c98be509be

Observation 919c1440-c2f1-490e-b7ce-0b418b1e4dae · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:19:14.904430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T23:15:44.550045Z digest=sha256:3d6927d0afc10213f38f649bb8d87a93b406dcadd881e4032feaec2e51a3d7b0

Observation 20f42e94-cf13-40f8-96f7-4ab6f5e594db · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:25:07.333934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T23:23:42.883286Z digest=sha256:e57b638549e67f350eebb41b73fef1c9bee394dd73fb991d02aca36287bdd47b

Observation b931cb18-5ab5-47e2-b9d4-e2fa4fe981fe · inbound

Training Language Agents to Learn from Experience cites this paper.

Training Language Agents to Learn from Experience Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:14:02.615242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T07:11:09.642275Z digest=sha256:86b7e913b5c09728e385ab5dc3d8ab5638300aa8499a0a4520c8fca2710d4c27

Observation 84dc25dd-01a6-4f6c-a937-6bf2731214a6 · inbound

EvolveNav: Proactive Preflection and Self-Evolving Memory for Zero-Shot Object Goal Navigation cites this paper.

EvolveNav: Proactive Preflection and Self-Evolving Memory for Zero-Shot Object Goal Navigation Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:28:58.628130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:34:47.530464Z digest=sha256:0af5901cd8d237c1e774d7a746d24f4349ad2439404b46cc6866a83e5942907f

Observation 2d5a709f-7554-4495-9096-3cfa72d7bed2 · inbound

S2T-RLHF: Hierarchical Credit Assignment for Stable Preference-Based RLHF cites this paper.

S2T-RLHF: Hierarchical Credit Assignment for Stable Preference-Based RLHF Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T13:58:42.937172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:58:42.937172Z digest=sha256:54240dc1f65731444f1a62d37600fa6b5588111591b839d4f25895da37bf8b8e