Pith. sign in

Paper Citation Record · LEDGER

OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2507.09075.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.09075 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T15:28:10.958789Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T20:00:08.055225Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a76b05f4-2c2b-4b34-b51d-36be98887fbe · inbound

Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation cites this paper.

Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:06:13.680163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T10:04:39.223895Z digest=sha256:52e1cabed4d4d6cc8fb9815bd122a4b03d66a3e3ba22d2f2df56e9cb9cbca06c

Observation 4ff35ad4-42f7-4fd2-8ac6-1d4a993ac00c · inbound

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment cites this paper.

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:15:56.628516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T18:36:44.401045Z digest=sha256:eca5191a6bc18c62902d07e01265612e9778c4bcb71e308cb9ca2df2d32477e9

Observation 5703933e-02f5-43cc-a133-40f58b8ef3f1 · inbound

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling cites this paper.

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:31:26.105108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-07T10:42:27.644514Z digest=sha256:86c2e481d5603e7edf48c29c97d9924862b7d165c51bbd623ad2d08d42e5630c

Observation 8293d704-dd6e-4237-bebd-db601ca6c682 · inbound

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling cites this paper.

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T15:28:10.958789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T15:28:10.958789Z digest=sha256:6f4415a14e37fb69244bf8fc2a2a3bed28a29087cda1d86493cffe006f854366

Observation efc7b009-d233-4ea9-a580-6ed7e062ac14 · inbound

Primal Generation, Dual Judgment: Self-Training from Test-Time Scaling cites this paper.

Primal Generation, Dual Judgment: Self-Training from Test-Time Scaling OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:27:07.026561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T02:25:58.830629Z digest=sha256:a98c3bd3eb513f97208df6007a635ea8e6e4dfa5a79a34824e2c1290bb460c0d

Observation 7aea010d-59fc-41f0-9a3e-c6cf9fc372cb · inbound

Primal Generation, Dual Judgment: Self-Training from Test-Time Scaling cites this paper.

Primal Generation, Dual Judgment: Self-Training from Test-Time Scaling OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:47:58.169224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-14T20:46:38.558668Z digest=sha256:bf3118f50a379085236ff2e29d9e3e769ac0100ecbbfde4ccf3162597ca8b09a

Observation 7fa51c6d-8ed9-4aa5-918b-1d3c764bc5f0 · inbound

Extrapolative Weight Averaging Reveals Correctness-Efficiency Frontiers in Code RL cites this paper.

Extrapolative Weight Averaging Reveals Correctness-Efficiency Frontiers in Code RL OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-29T14:43:30.569364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-29T14:41:13.191919Z digest=sha256:5f401c1c138b28d293f7ce0e2c363d5480bbfaa59557dc2af36d34978203e683

Observation c66e47ac-f5df-46a0-8f44-9a6b6024c46a · inbound

The Generalization Spectrum: A Chromatographic Approach to Evaluating Learning Algorithms cites this paper.

The Generalization Spectrum: A Chromatographic Approach to Evaluating Learning Algorithms OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-04T20:00:08.057207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-25T20:55:15.784610Z digest=sha256:6478b87ca5719bfe461be529d3134fde6b97e26e2632854c99fa4bbf7520ce6a

Observation 0f4aea5e-41cf-4374-91a4-ed9f05dc9d47 · inbound

The Generalization Spectrum: A Chromatographic Approach to Evaluating Learning Algorithms cites this paper.

The Generalization Spectrum: A Chromatographic Approach to Evaluating Learning Algorithms OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:09:50.551077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T05:29:21.598397Z digest=sha256:c581b10a625fe790eef8e4b1ee5ce8a42dcfbaeb32ac07c5c02b16f6e48f588e

Observation 011d3c8b-f5df-496a-856a-475beecca7bd · inbound

Don't Let Gains FADE: Breaking Down Policy Gradient Weights in RL cites this paper.

Don't Let Gains FADE: Breaking Down Policy Gradient Weights in RL OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:57.716470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-07-03T20:59:57.539909Z digest=sha256:0866ba5ff641ebd65d1e46859bdc475b650925545b1564f5f3c18df1cfed552b