Pith. sign in

Paper Citation Record · LEDGER

Revisiting the Test-Time Scaling of o1-like Models: Do they Truly Possess Test-Time Scaling Capabilities?

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2502.12215.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.12215 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-18T22:42:53.663354Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T22:46:53.881797Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 130422e1-a575-4ee2-a484-2476cbc21231 · inbound

Pruning Long Chain-of-Thought of Large Reasoning Models via Small-Scale Preference Optimization cites this paper.

Pruning Long Chain-of-Thought of Large Reasoning Models via Small-Scale Preference Optimization Revisiting the Test-Time Scaling of o1-like Models: Do they Truly Possess Test-Time Scaling Capabilities?

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:31:53.235073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T22:26:52.748349Z digest=sha256:8344c2d382dbe55d1bbac6a38c0fa85aec9ef347168b9cc7be570781eb4e9b05

Observation dce502be-b0b0-4c45-8894-f8c8ae49c6d3 · inbound

Input-Time Scaling: Adding Noise and Irrelevance into Less-Is-More Drastically Improves Reasoning Performance and Efficiency cites this paper.

Input-Time Scaling: Adding Noise and Irrelevance into Less-Is-More Drastically Improves Reasoning Performance and Efficiency Revisiting the Test-Time Scaling of o1-like Models: Do they Truly Possess Test-Time Scaling Capabilities?

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:46:53.885666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T22:42:53.663354Z digest=sha256:7fe2de3bb8663bb8e114129d093039d0da9a7fceea027339f8777b5991f055c8

Observation d973d329-cd95-4725-9222-397e2154d88c · inbound

AdverMCTS: Combating Pseudo-Correctness in Code Generation via Adversarial Monte Carlo Tree Search cites this paper.

AdverMCTS: Combating Pseudo-Correctness in Code Generation via Adversarial Monte Carlo Tree Search Revisiting the Test-Time Scaling of o1-like Models: Do they Truly Possess Test-Time Scaling Capabilities?

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:40:57.348127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T16:35:16.056397Z digest=sha256:0a7a790dcc18d7282de7ad0234c1f694bdd174e847898d9d7137d36a36255ded

Observation 12d053bc-7e64-45dc-b332-134912e9daeb · inbound

VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model cites this paper.

VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model Revisiting the Test-Time Scaling of o1-like Models: Do they Truly Possess Test-Time Scaling Capabilities?

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:46:06.353279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T15:10:16.533927Z digest=sha256:1531836749ed5a922ef0890fae53ba58b583541d7896f6ccfdfc8c2533fd84ec

Observation 83a8a0f6-5e6c-406e-b65d-e84870f734eb · inbound

V-ABS: Action-Observer Driven Beam Search for Dynamic Visual Reasoning cites this paper.

V-ABS: Action-Observer Driven Beam Search for Dynamic Visual Reasoning Revisiting the Test-Time Scaling of o1-like Models: Do they Truly Possess Test-Time Scaling Capabilities?

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:41:42.117003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T04:03:01.612516Z digest=sha256:48c21eab7756f4ffa59a6cbcc1482705c0cf3d9356d0743861d42b292106abfd