Pith. sign in

Paper Citation Record · LEDGER

Learning Planning-based Reasoning by Trajectories Collection and Process Reward Synthesizing

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2402.00658.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.00658 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:22.042716Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T15:42:26.911183Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7749b063-e0fb-4649-b3c8-43bf035ed78e · inbound

DiagnosisArena: Benchmarking Diagnostic Reasoning for Large Language Models cites this paper.

DiagnosisArena: Benchmarking Diagnostic Reasoning for Large Language Models Learning Planning-based Reasoning by Trajectories Collection and Process Reward Synthesizing

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:42:27.076360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:42:22.042716Z digest=sha256:71f37c29d5345a0635f64bdb37f94e5ad1097858cfd2b7ff7d05bfd1758f9e67

Observation 0c4735c7-4597-42d8-a28f-021c161a5a18 · inbound

Large Language Models for Planning: A Comprehensive and Systematic Survey cites this paper.

Large Language Models for Planning: A Comprehensive and Systematic Survey Learning Planning-based Reasoning by Trajectories Collection and Process Reward Synthesizing

Reference 111

Resolution
unresolved
no resolver link, observed 2026-08-07T14:11:57.753578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:11:57.753578Z digest=sha256:1d36025e4c4b217561a9e253717bf59f643d31da6cd6eee5f2f43c3a92565c73

Observation 990160fb-6307-4daa-9858-00d3fdbbe47c · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay Learning Planning-based Reasoning by Trajectories Collection and Process Reward Synthesizing

Reference 92

Resolution
unresolved
no resolver link, observed 2026-07-11T13:53:36.775836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T13:53:36.775836Z digest=sha256:ca4baf21f0fb02ce1244515ce82ed0bed8e1b808717f1174540301bf672fbefb

Observation 594a8fd3-7461-4573-91c8-b2fdff369f9e · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay Learning Planning-based Reasoning by Trajectories Collection and Process Reward Synthesizing

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-02T08:40:41.861888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:40:41.861888Z digest=sha256:f9342caf6ae0f506633df46dca1244634b46de16db5b956fc0a319c6b3d377bd