Pith. sign in

Paper Citation Record · LEDGER

Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2308.02151.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.02151 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:24:14.981510Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T21:28:58.626281Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f8659ef3-82dc-41a9-a27b-7aafb6eb37c1 · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 234

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:28:39.056161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:578332f7d2b9cdde1cacdc84cb3d82a2c21c31523f55d823a3902c7d1cde3e18

Observation c960f693-2689-47f1-85d4-6dd2a9dec05e · inbound

A Survey on the Memory Mechanism of Large Language Model based Agents cites this paper.

A Survey on the Memory Mechanism of Large Language Model based Agents Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:21:39.992190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T07:21:39.440092Z digest=sha256:7b7e056e023f7a9a51d0921855d79ebb3be54b6675e3b8b908394561e95cc17b

Observation 6682630d-db10-425d-8e0a-61d2a7e13492 · inbound

From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMs cites this paper.

From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMs Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-17T11:05:09.813162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T11:05:09.588491Z digest=sha256:db0e23a8545c44f632e699d8c0de8350c1d5bf9207df244610b7c40723b08ddf

Observation 498ff026-fe3f-4f37-bd26-0a457ef52ece · inbound

Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team cites this paper.

Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T00:24:14.981510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:24:14.981510Z digest=sha256:c45e57ef45c9e31d799eb274d637add5397bd4d4d2418d0c7f9d4ffa8f2a2688

Observation d97be4d3-3c47-4cd2-b2ba-ae416b0d23cf · inbound

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning cites this paper.

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:06:08.681011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-08T10:23:52.522238Z digest=sha256:8f13ea62b1993baf649fce1a704745345227f02287be6051b976f2cc8ed942ae

Observation fd45479c-89ef-47c3-8787-9f632c461736 · inbound

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning cites this paper.

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:00:56.846868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-11T02:00:00.663355Z digest=sha256:325c08ec9079f89050f8f773b73a308835eb3830dbbf776489a90b3cbf4dec6d

Observation 2257a956-2335-440d-96e2-ff68c3f54f44 · inbound

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning cites this paper.

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:17:28.208672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T07:17:13.708752Z digest=sha256:f6ce3b16f6191401f807df5155a4e96289d84e8216eb4383096eb88996bcdb14

Observation e9983242-aa4a-4fd2-b415-f426ffd395d4 · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:20:57.093072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T01:47:39.926540Z digest=sha256:608f02b258ec2f01398ca8e7a5fd2ddd34bcc9ad631394404bb1aaa8b48293c2

Observation 919c1440-c2f1-490e-b7ce-0b418b1e4dae · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:19:14.904430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T23:15:44.550045Z digest=sha256:962c892155b3ac28425a106fa856136b700bfbc0b8e41a052d535d4328d322bf

Observation 20f42e94-cf13-40f8-96f7-4ab6f5e594db · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:25:07.333934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T23:23:42.883286Z digest=sha256:2cfbbfa606842896840cbe615a2535a7ef11c4ccd4b73f6a9a41960ad19b4b4e

Observation b931cb18-5ab5-47e2-b9d4-e2fa4fe981fe · inbound

Training Language Agents to Learn from Experience cites this paper.

Training Language Agents to Learn from Experience Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:14:02.615242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T07:11:09.642275Z digest=sha256:5a973182f3b0c45413820c532775e274cfd61fd9639ef1de97d32f41aad9fd76

Observation 84dc25dd-01a6-4f6c-a937-6bf2731214a6 · inbound

EvolveNav: Proactive Preflection and Self-Evolving Memory for Zero-Shot Object Goal Navigation cites this paper.

EvolveNav: Proactive Preflection and Self-Evolving Memory for Zero-Shot Object Goal Navigation Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:28:58.628130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T00:34:47.530464Z digest=sha256:a8efee8442a066555d2b0c021471e769e0bb5f337c4436a22d2be68d9e2351ae

Observation 2d5a709f-7554-4495-9096-3cfa72d7bed2 · inbound

S2T-RLHF: Hierarchical Credit Assignment for Stable Preference-Based RLHF cites this paper.

S2T-RLHF: Hierarchical Credit Assignment for Stable Preference-Based RLHF Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T13:58:42.937172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:58:42.937172Z digest=sha256:54240dc1f65731444f1a62d37600fa6b5588111591b839d4f25895da37bf8b8e