Pith. sign in

Paper Citation Record · LEDGER

ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2312.10003.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.10003 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T19:34:28.337466Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T06:15:00.866473Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0d5d98b6-93e2-4658-8d54-4003accc3b97 · inbound

From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review cites this paper.

From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:57:38.252446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T02:57:37.873567Z digest=sha256:8a5c268b67664c3b686ab2c31843eb662e0e9775ae6cff911dbc4cb019a8431f

Observation 7bf92f9b-25fa-45a2-949f-e52db5e4775e · inbound

AI Reasoning for Wireless Communications and Networking: A Survey and Perspectives cites this paper.

AI Reasoning for Wireless Communications and Networking: A Survey and Perspectives ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-04T19:34:28.337466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:34:28.337466Z digest=sha256:244401804907fe4995c5ca53fe719d71ef09398cef2960488e06ae1c03e004dc

Observation af7de01d-ebad-4268-b54c-87434c6f3bcd · inbound

Estimating the Empowerment of Language Model Agents cites this paper.

Estimating the Empowerment of Language Model Agents ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T14:53:14.599607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:53:14.599607Z digest=sha256:04970fd00f60c57a4a9b483ba8711b71860aa1cd91e9a45e2b45b46caa4729f3

Observation 70ba945b-1d86-4212-81f1-419d5feffc55 · inbound

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory cites this paper.

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 234

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:15.679804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-14T23:13:15.016486Z digest=sha256:5842e4f1c15253f69b57a731def8cc334b6b0eb3519156f735153797b12d900f

Observation cb05160e-8cee-452a-8503-e66f5b5c1c96 · inbound

Agentic Reasoning for Large Language Models cites this paper.

Agentic Reasoning for Large Language Models ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:14:25.988543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T15:14:25.558878Z digest=sha256:6903fdab32bac3f2d029f789ec4c19a172bf5691d27def271ad8371b32c3f662

Observation c3be1648-17fc-4441-ab3d-2bf62885c8f5 · inbound

Toward Efficient Agents: Memory, Tool learning, and Planning cites this paper.

Toward Efficient Agents: Memory, Tool learning, and Planning ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T09:21:32.684964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:21:32.684964Z digest=sha256:4d411497865ae5d6df639862569a50c0a82bea02182dcc04d609827317ce224d

Observation 037ae162-fc22-492f-bde3-b91cfe59372d · inbound

Quantifying Trust: Financial Risk Management for Trustworthy AI Agents cites this paper.

Quantifying Trust: Financial Risk Management for Trustworthy AI Agents ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:18:01.300645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T17:16:17.464937Z digest=sha256:2d2ba327b3d94aa1507c6ff97cba7f862c6801ea8d9cf765bf1abb45e90c3b43

Observation 7258b9e7-2cdf-4f52-9335-e577337d4d82 · inbound

DuIVRS-2: An LLM-based Interactive Voice Response System for Large-scale POI Attribute Acquisition cites this paper.

DuIVRS-2: An LLM-based Interactive Voice Response System for Large-scale POI Attribute Acquisition ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:33:12.922872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-20T10:30:48.957568Z digest=sha256:56c8b5f931b518c99a0df5dac0935c7667e93fb94542e6683488040c285441dd

Observation 3352d076-4318-4d6b-b64a-b506e727b548 · inbound

EvoGens: A Population-Based Heuristic Search Framework for Scientific Idea Generation cites this paper.

EvoGens: A Population-Based Heuristic Search Framework for Scientific Idea Generation ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:32:44.146779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T22:29:44.078911Z digest=sha256:40269c22d07508f7d6c92f1a480e763fff10ef78affa8718110168b55fd28dbc

Observation 4c09d938-fcb1-48b2-865a-423e2a18bd28 · inbound

SlimSearcher: Training Efficiency-Aware Web Agents via Adaptive Reward Gating cites this paper.

SlimSearcher: Training Efficiency-Aware Web Agents via Adaptive Reward Gating ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:17:08.710917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T22:54:44.329613Z digest=sha256:ba4bd75ac7e342dbe6ca7033a73e23f347dd59b846a6bfe266e1bab8d93b3253

Observation d379452a-44fc-4c5a-bf92-c4ce6d53b990 · inbound

Graph2Idea:Retrieval-Augmented Scientific Idea Generation with Graph-Structured Contexts cites this paper.

Graph2Idea:Retrieval-Augmented Scientific Idea Generation with Graph-Structured Contexts ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:17:31.037482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T16:39:36.945808Z digest=sha256:93a30ad50a1ea72e133b62b5558cd0bf3a5e9c0201dafb6752df460c596aa5ac

Observation 079675ac-308f-40a4-a1e8-8918767a667b · inbound

Self-Improvement Can Self-Regress: The Rise-and-Collapse Failure Mode of LLM Self-Training cites this paper.

Self-Improvement Can Self-Regress: The Rise-and-Collapse Failure Mode of LLM Self-Training ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:29:16.726094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T21:07:28.660119Z digest=sha256:8380f4893ef2a40ce2b88d0c6c67eb8a7c78c98fc6ab7451d96e6d96a69e98a7

Observation b5f02c49-29fc-4e89-8f0d-210781c0d1e7 · inbound

Stop Hand-Holding Your Coding Agent: Engineering the Loops that Replace Step-by-Step Prompting cites this paper.

Stop Hand-Holding Your Coding Agent: Engineering the Loops that Replace Step-by-Step Prompting ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:47:22.277629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-02T20:40:11.059493Z digest=sha256:acab081492f1a479fcfa820a84f2a903ce8fe6e710cc8284581b4198af4bc5ad

Observation 8ceacf65-bebc-4793-b2c4-7fc4ebf54b23 · inbound

LLMs for Agentic Home Energy Management cites this paper.

LLMs for Agentic Home Energy Management ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-11T17:08:27.234032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T17:08:27.234032Z digest=sha256:647d9bea50089317e11462844005ec234e6f07fbcf4898fd6497f1db88b20f1c

Observation ee669dc4-56fa-4ab4-9f3a-51b5978ba771 · inbound

LLMs for Agentic Home Energy Management cites this paper.

LLMs for Agentic Home Energy Management ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T08:40:33.501969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:40:33.501969Z digest=sha256:65327c1ab42e6960185cd04a9dcadff3f92013218131baa4506ea94515603f85

Observation 7ffba785-912c-490c-9786-776a4b314105 · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 78

Resolution
unresolved
no resolver link, observed 2026-07-11T13:53:36.775836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T13:53:36.775836Z digest=sha256:8f74a9f268fdebe08ec8713159f11efa223bacb69e835fc17d31663cf4adac53

Observation 6f322826-12d6-45ff-ba02-ecde46fa59be · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-02T08:40:40.326974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:40:40.326974Z digest=sha256:1ba6741ded7fd1852efc1a7d4c387dd1a35b019a795ac7325100173e33da5aab

Observation a194edc7-af73-4ddd-8eb2-ca258a572d92 · inbound

Engineering Trustworthy Agentic AI for Critical Systems cites this paper.

Engineering Trustworthy Agentic AI for Critical Systems ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:28.624570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:28.624570Z digest=sha256:af8482ad11794133bb3bfd2b35b5777a1d3cb6cec693567e76937265050eaf8f