Pith. sign in

Paper Citation Record · LEDGER

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning

As of 8 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2505.17829.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17829 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:45:20.428126Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

53 of 53 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved52
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cbb5158a-e8f9-4d14-893a-453804d038da · outbound

This paper cites L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:15.860789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:15.860789Z digest=sha256:03b43a378126cd026811e187fd67019e744977a6032d64cef3063466d0311987

Observation d84d6c05-a592-4a30-9678-7e5de1b18f6f · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:15.928935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:15.928935Z digest=sha256:8a448fbfa0036424fb7ed45844db0207d8ad09ef1bb512762e3caa371e6059e3

Observation 8198a83c-174c-47bb-9380-71fc6c8467c9 · outbound

This paper cites Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:15.995858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:15.995858Z digest=sha256:f0135448664e4048877bf20495db663de17a271ea762beed8fce89f1c9c8931d

Observation bf5dd1f9-059f-4b2f-83db-a84b73358037 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.052650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.052650Z digest=sha256:1bf39d5a481006e1c6a2daf0e910923e3964f79e391f181d7b78a7f8462fd0db

Observation 5408ac83-e892-4b72-8be3-502414775fe8 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.157711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.157711Z digest=sha256:bb462d107f611072517e7b8000790a08aee08f2f5d92c89349a3334e2c45bfe6

Observation 622a08fe-2874-4516-9be6-0009b06f3410 · outbound

This paper cites An Empirical Study on Eliciting and Improving R1-like Reasoning Models.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning An Empirical Study on Eliciting and Improving R1-like Reasoning Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.258476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.258476Z digest=sha256:3b3f9cfdd3a80a0f633c915fec9e9062285120e75546a8bfe111be1db01d6cbc

Observation f1958060-a26f-4663-849c-d564bdc09acb · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Training Verifiers to Solve Math Word Problems

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.355876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.355876Z digest=sha256:34f4144cc163826484d47be2a8250e1b0b37f24256c47e7fd27c153bcb124780

Observation cc788a5d-fee3-4fcf-a0f5-5b710919b473 · outbound

This paper cites Rethinking External Slow-Thinking: From Snowball Errors to Probability of Correct Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Rethinking External Slow-Thinking: From Snowball Errors to Probability of Correct Reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.471846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.471846Z digest=sha256:ac33dc6108788a79faca0fafe50b8d372517598c658d789b0200bc057a7be372

Observation 94ccc782-0e46-4550-b340-3a0efbd2a360 · outbound

This paper cites The Llama 3 Herd of Models.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.567297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.567297Z digest=sha256:593b886bac338d7a8f09d952f2d31a3c07a114437be9d0168e2981cd8dfb94ae

Observation 547bb6f3-c113-4139-be22-5ddb1d03a8c4 · outbound

This paper cites Search, Verify and Feedback: Towards Next Generation Post-training Paradigm of Foundation Models via Verifier Engineering.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Search, Verify and Feedback: Towards Next Generation Post-training Paradigm of Foundation Models via Verifier Engineering

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.667323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.667323Z digest=sha256:2f7ec38da559db98ab87e36f98a949cbcd23d0dee09a1874e24704b41c09aaa8

Observation 38b3018e-0145-4e90-a9ac-f5ecea67b504 · outbound

This paper cites rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.786796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.786796Z digest=sha256:0fe5ac29edfaa844a651e63ba49c24ac5ff3d6aa5bde7b62e39ac5276387770e

Observation 91b7c566-b336-4bef-aefc-352dc56b0f58 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.870979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.870979Z digest=sha256:3803b18200559e9f9bef0ea9340f6adfdc11116491f0da35e99f39ec7f549a58

Observation a45e6103-0b37-4c7a-ace9-db5f77a7673e · outbound

This paper cites Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.967173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.967173Z digest=sha256:7cf50d49344fe364799a56eeb5b5076167576df5227fa8f3047d68a06a2f682e

Observation 69c03bf8-35d6-4f45-a1d2-c4ae7fd881da · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.034269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.034269Z digest=sha256:b7069600df657fd22fe09a4b07a78dba8941d79df2d75aebfb8855b85919bddd

Observation 4c3263cb-f17e-41de-bd6a-ca488c2b2211 · outbound

This paper cites ETS: Efficient Tree Search for Inference-Time Scaling.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning ETS: Efficient Tree Search for Inference-Time Scaling

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.115091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.115091Z digest=sha256:40764ed76ae34f9a49601a3b8999361d8e4b2c637b44ee09056b68c192d4d804

Observation 895124a2-d20a-4fc3-8433-9edc9e6e7bd9 · outbound

This paper cites Efficient Test-Time Scaling via Self-Calibration.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Efficient Test-Time Scaling via Self-Calibration

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.220177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.220177Z digest=sha256:98cf05ef0026efd15d2a03e098024c4d6df5f9a89b542524403e2a603de7f87a

Observation b0e05e90-6c61-41af-995c-45701c748fc2 · outbound

This paper cites A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.297639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.297639Z digest=sha256:732b2689a6a418a7cbb69239af7fdf759699db0f7a0a157ae95bf0e17f7de036

Observation 9f9ffe20-ea0b-4f37-9e95-257e32a45eae · outbound

This paper cites Enhancing LLM Reasoning with Reward-guided Tree Search.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.406362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.406362Z digest=sha256:c31864e897fed90460de5818e5ceb2c390585205a3a171e352c3b8ce44ad1652

Observation 1a881508-053e-43ab-953e-65baf416e635 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.487613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.487613Z digest=sha256:90fcab5e38c244b7b1427f7eacb04997e7d4b1b2406845e654747b1d5457e66e

Observation b25cc179-2cb9-4684-876d-15fefb5f7993 · outbound

This paper cites Escape Sky-high Cost: Early-stopping Self-Consistency for Multi-step Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Escape Sky-high Cost: Early-stopping Self-Consistency for Multi-step Reasoning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.591164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.591164Z digest=sha256:5461ae30ab2ab0c0eefbe57fd93e1efd4bdc9f5a2217597036fcf7ad20c2c0af

Observation 21c2f839-5f79-4d08-a540-8d2ff69b315b · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.671386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.671386Z digest=sha256:c8b1e7439d3b915ca1685d227880445dbab3696c0fa7f126a589a64c416ac93f

Observation cec42c1f-e556-47dc-b89a-fa22ea3aa623 · outbound

This paper cites Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.775098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.775098Z digest=sha256:350fa1cbd07656023b90b89741f57dfb6348e64faf0c568d4beab6a6a8e79e56

Observation 1249619d-e278-4318-80bc-9206c8487c21 · outbound

This paper cites Improve Mathematical Reasoning in Language Models by Automated Process Supervision.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Improve Mathematical Reasoning in Language Models by Automated Process Supervision

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.861908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.861908Z digest=sha256:e5d354934e96e0b073106aeb839f11712ec938a04400a057becb7126c495a13d

Observation 91847d96-bfdd-4c5d-82ce-097a4fdeb616 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:21.696466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:45:17.962098Z digest=sha256:14c184f53da9048796266977e229f3b625c7c969f63114e7ad2f3644740ef8db

Observation f3feb280-534b-4aff-be6f-08be76e2c893 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.067650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.067650Z digest=sha256:110a610d4e3b081ed6712298caaef1eea57e931a67b147a2928f354b800f5624

Observation 8205ddc8-fe73-4af5-90ba-658f20f51cca · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.171836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.171836Z digest=sha256:de08a32780e139957d1c0978d47a475f201ddfc08d574c2899c6e7d29b702ac0

Observation eb92f12a-5360-4e3d-840b-43ff23e88e91 · outbound

This paper cites Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.293852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.293852Z digest=sha256:45b462b4bd85a284e3f669db54ddf133a5211f2fcab90c83a196dccff40f39c8

Observation 8e2d8aad-fc61-4f3d-927e-56495a8c0ea2 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.388703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.388703Z digest=sha256:4158d60b26937486ca1fd09f454bea37df8c0d6146e4ae089e81a9957054b5af

Observation d1c1ddc5-266a-468e-8699-bb896de78f2f · outbound

This paper cites Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.474040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.474040Z digest=sha256:6bb44d08c12abca24b104ced81568c1ed6d5b53490cde6f177d30487179b007f

Observation 77f4f98b-17b4-4dde-9671-f1a15288d3a4 · outbound

This paper cites Solving math word problems with process- and outcome-based feedback.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Solving math word problems with process- and outcome-based feedback

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.564018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.564018Z digest=sha256:466a879db0bf7d44f83b2d000a20dd545f1449eec8647c270184ecb3a44f3049

Observation a15fb25e-acc2-4aea-84cd-154f347a5f9e · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:21.514394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:45:18.664750Z digest=sha256:b9cc9b89d93f488b20995fb690b11cb69326c15b4140b55d988ea023a16745eb

Observation 3aa87878-559e-499f-a8ea-fbc5ec7b7e47 · outbound

This paper cites OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.769402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.769402Z digest=sha256:7f99593bf3a54c28f48d96e3a792ab673b41db533ded0b0c3b0d1c70cade5bd8

Observation db8b369f-b0ab-4444-9dac-49822359ca93 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.860070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.860070Z digest=sha256:778f541f76b27248dad41dcb3347e0f621bfbb119a6062f643b34d2edbcb953f

Observation a085871c-a7ba-4de6-806d-5b693720461d · outbound

This paper cites Towards Self-Improvement of LLMs via MCTS: Leveraging Stepwise Knowledge with Curriculum Preference Learning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Towards Self-Improvement of LLMs via MCTS: Leveraging Stepwise Knowledge with Curriculum Preference Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.952147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.952147Z digest=sha256:7033c748ca362d5335ca59c76ca74402f69db8603222a5c27b3a2f264334fc45

Observation 86aacc1c-ec73-49e0-b39a-bd72eb478738 · outbound

This paper cites Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.024357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.024357Z digest=sha256:7e1444493c27695986a2fe0b1daf7ab37c27fe0f4340f5e5dccd5e638fcf313d

Observation 38e75dee-09bc-464f-9e5b-dbbb4e5cc9e8 · outbound

This paper cites Chain-of-Probe: Examining the Necessity and Accuracy of CoT Step-by-Step.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Chain-of-Probe: Examining the Necessity and Accuracy of CoT Step-by-Step

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:45:20.790255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:45:19.073344Z digest=sha256:46649b976a31783cfe8a162319f5472dda877cda28841b860182398dc7577fb0

Observation 834c6eb3-761f-43ec-b151-317162dcc95a · outbound

This paper cites Chi, Quoc V.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Chi, Quoc V

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.142078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.142078Z digest=sha256:6613a0aad346beb864f191e07e9a1be91c21940dfee2ccc341c4cac7c1429e2c

Observation e76e0fb6-1f3d-40d6-b397-352717e997df · outbound

This paper cites Beyond Examples: High-level Automated Reasoning Paradigm in In-Context Learning via MCTS.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Beyond Examples: High-level Automated Reasoning Paradigm in In-Context Learning via MCTS

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.217464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.217464Z digest=sha256:965f8ff0dbe4e47a8c57d1274cc22d93a6805ef8637a2c8cf9c8a6087abba3ac

Observation 984c6ccb-4398-4e84-a8da-5fdbe806bc49 · outbound

This paper cites A Comparative Study on Reasoning Patterns of OpenAI's o1 Model.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning A Comparative Study on Reasoning Patterns of OpenAI's o1 Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.282901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.282901Z digest=sha256:07b7e054d354d363712ac7f2251360db40213ac1b1943c17b659a3c9050e988a

Observation 10c14ed8-4a82-4f5e-923e-154a71b61934 · outbound

This paper cites When More is Less: Understanding Chain-of-Thought Length in LLMs.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning When More is Less: Understanding Chain-of-Thought Length in LLMs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.333659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.333659Z digest=sha256:baac53f3c84767e496a26e38a2cf2f199255e537fb53708ccf1c7c25fed42882

Observation e23b3e40-b19e-44cb-9332-01a0262f1503 · outbound

This paper cites Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.417118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.417118Z digest=sha256:1708790b2e07b6ffef3fd44d42e1a0b467dd66b70e0132b19934e1a26e78cc3d

Observation 13e43f31-d822-4b72-a7f6-63725eaf9c21 · outbound

This paper cites Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.472575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.472575Z digest=sha256:8f48043a9c79931082ff19731f4d6479792926ea8bd5c63ebfee1d542975050b

Observation 85cb7881-06b8-4663-b4dc-993a664390ef · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.536322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.536322Z digest=sha256:8295ca476dba57b3be0c777ec68193598b9cf4444ba13a96897ef75856a50771

Observation 0aee6a39-b94b-4531-8ad2-82b8a53dfab7 · outbound

This paper cites Qwen3 Technical Report.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Qwen3 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.631357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.631357Z digest=sha256:f9d4a0102d80c6ede53a87cfbedab7f21e83056eba972f9370d57e7a6b90888b

Observation 62608014-6a3b-4aaa-9240-9b01f50e2ffb · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:21.351151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:45:19.734663Z digest=sha256:63bf67d7af4c65b8fca5e19a1bdb32d44d9c57d2530dc739736001bf12b2cc8c

Observation 74de4af0-8adf-4a44-9744-56519d94e536 · outbound

This paper cites B-STaR: Monitoring and Balancing Exploration and Exploitation in Self-Taught Reasoners.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning B-STaR: Monitoring and Balancing Exploration and Exploitation in Self-Taught Reasoners

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.802170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.802170Z digest=sha256:6eb1c8ab0b1151a2809be9b2d8e8c05f11b218c18936fcc7ee5fc51323c54f1b

Observation fb9a8d92-a6fb-4231-9b1a-804b31ec998f · outbound

This paper cites Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.883202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.883202Z digest=sha256:06c3e33af289f91ee07375ed6015cbb1ef3b2dcef75c14868f610617e2e41f3f

Observation f03cbe38-d485-4e5e-8883-c58344b01d95 · outbound

This paper cites Generative Verifiers: Reward Modeling as Next-Token Prediction.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Generative Verifiers: Reward Modeling as Next-Token Prediction

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.980989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.980989Z digest=sha256:62d65eaa544bafea2f49f80c01fcf610db49f23a7224ee82c83302de14b3a3a3

Observation 9ac2fd92-d22d-4ebe-a714-758e178760ec · outbound

This paper cites A Survey on Test-Time Scaling in Large Language Models: What, How, Where, and How Well?.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning A Survey on Test-Time Scaling in Large Language Models: What, How, Where, and How Well?

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:20.078930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:20.078930Z digest=sha256:d1f425cadac01c413cda1a4c008fb924e505d4c89d4cb4be3a91c3252bc5c744

Observation b5839cb0-47d4-4ac5-98dd-348c6f1c80cd · outbound

This paper cites Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:20.163992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:20.163992Z digest=sha256:773fa8651913b3f44d79bfc367903c3ce065d71254999ff9585121fd4100e8e0

Observation 2e2d339d-9926-4eb1-a780-9707b0305fe9 · outbound

This paper cites ProcessBench: Identifying Process Errors in Mathematical Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning ProcessBench: Identifying Process Errors in Mathematical Reasoning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:20.276917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:20.276917Z digest=sha256:3df98642eaa1a1cf75ae2a8d9d2e055e7010f4ea2b53ae32332d003c68594ab9

Observation 9cc3677d-eeab-4864-a7b2-214235b9b7de · outbound

This paper cites online" 'onlinestring :=.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning online" 'onlinestring :=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:20.356452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:20.356452Z digest=sha256:9fefab176d9e16b5df0e4610ade2b67a55aac1d0f65c3ecfe505d21dde87500a

Observation 938ded77-216d-48cf-82da-478b2e82f1c6 · outbound

This paper cites write newline.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning write newline

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:20.428126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:20.428126Z digest=sha256:35633909ac578ca5d555d19077068f53b944ce3385c7198d97bf0bcfa5510606

Pith citing papers

No inbound Pith citation observations are available.