Pith. sign in

Paper Citation Record · LEDGER

Logical Reasoning with Outcome Reward Models for Test-Time Scaling

As of 7 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2508.19903.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.19903 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:26:52.816300Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

23 of 23 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b7087924-0729-49ef-a3bf-8212c0530578 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.728328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.728328Z digest=sha256:9b2f68ba4b4fe71ea9be7e129951bc2ca6fba26005ec78992d00287093fbe987

Observation c3056f3b-8e04-49fa-acf3-cd2deee7a1bb · outbound

This paper cites JustLogic: A Comprehensive Benchmark for Evaluating Deductive Reasoning in Large Language Models.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling JustLogic: A Comprehensive Benchmark for Evaluating Deductive Reasoning in Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.732798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.732798Z digest=sha256:cc1ba0cd4e295a9539b10ceef9667d5779f0ea68c70acc1a66d6f4b31e2ee01d

Observation 40436993-5ee3-41b0-853b-afd8acce21bd · outbound

This paper cites The Llama 3 Herd of Models.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling The Llama 3 Herd of Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.736892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.736892Z digest=sha256:fa195baa685683f3ec4a93bef40589fbbd205dc4131888cf78be1a72d77b4195

Observation 053b84d0-a2eb-4912-94f8-93da3eb77f65 · outbound

This paper cites Fabbri, Wojciech Kryscinski, Semih Yavuz, Ye Liu, Xi Victoria Lin, Shafiq Joty, Yingbo Zhou, Caiming Xiong, Rex Ying, Arman Cohan, and Dragomir Radev.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Fabbri, Wojciech Kryscinski, Semih Yavuz, Ye Liu, Xi Victoria Lin, Shafiq Joty, Yingbo Zhou, Caiming Xiong, Rex Ying, Arman Cohan, and Dragomir Radev

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:26:53.093107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T15:26:52.741046Z digest=sha256:eaf1d957523f54b7f818fad7177f3c6f125341cb397d0667895f5107cf5b8db1

Observation 1bacd92f-c50c-450c-9dc0-4e92fdbacd45 · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.744805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.744805Z digest=sha256:5f0c7d63559d6d70d8dae4dd8f300ee420e589bc368f7ff6ce70bea2cc74bcb2

Observation 67afe112-f1a2-483e-a564-6376c518eee4 · outbound

This paper cites Qwen2.5-Coder Technical Report.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Qwen2.5-Coder Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.748712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.748712Z digest=sha256:d66d39b63e26f8e276bc284c3162b093d659d7201f508fecfe51299e86f504cf

Observation 39104c15-7d71-463c-9d02-66be191b2844 · outbound

This paper cites Gemma 3 Technical Report.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Gemma 3 Technical Report

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.752965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.752965Z digest=sha256:d22f7c527075f76c603249d22eb623b6db315666c19376f9600747035444b742

Observation 37c1c938-6b4f-4244-8901-a6f83db59867 · outbound

This paper cites A Closer Look at Logical Reasoning with LLMs: The Choice of Tool Matters.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling A Closer Look at Logical Reasoning with LLMs: The Choice of Tool Matters

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.757698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.757698Z digest=sha256:4d5f1963c81b6e94e4b31e55c59ff5d415d486f125dd3a562f8d7b0a39b2d312

Observation b0b8ccb1-fad0-4035-9e9e-14ad90696a83 · outbound

This paper cites START: Self-taught Reasoner with Tools.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling START: Self-taught Reasoner with Tools

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.761394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.761394Z digest=sha256:5893406cb4aa6ce794d3196a42082220f9ac2aff602b11fb9856b71b19b0996e

Observation f402c349-5e83-409c-ac41-8e1fd96a9448 · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.765584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.765584Z digest=sha256:8df27f5ebe20819825a78cb827c65ea10b15cb3ddfd0dbc592f5e68b5ed03339

Observation cdac56e0-1ff4-4ba7-8015-06b60a9d04a6 · outbound

This paper cites Zhang, Armando Solar - Lezama, Joshua B.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Zhang, Armando Solar - Lezama, Joshua B

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.768976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.768976Z digest=sha256:db46b25e955e37177dec06069c8e42b9a600ccb79404a847e4d8d9a03caf439b

Observation e003ffb1-1682-48dd-8b3c-14cc9c6be5f2 · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.773005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.773005Z digest=sha256:3515311c829bc44d88ddd5c2b708177a9ff8257347bc40b0770bfd74cb5e3cae

Observation ce3e289a-d08b-4d3d-8d3f-f37884a6f23d · outbound

This paper cites Large Language Models Meet Symbolic Provers for Logical Reasoning Evaluation.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Large Language Models Meet Symbolic Provers for Logical Reasoning Evaluation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.776686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.776686Z digest=sha256:8a934c0016482fd6d064856e303ab8ad8ce78b734ee6a2dc89fe8a7955111d6f

Observation 1031d570-310b-4458-acf8-86cdf3ff0bc1 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.780303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.780303Z digest=sha256:77a3de3f053c0355ef7d320ba4c0d49a817c1c334b1c92bf7913711ab768b3e6

Observation 364cffa8-0ada-4631-844e-ad15dfe307d4 · outbound

This paper cites Strategies for Improving NL-to-FOL Translation with LLMs: Data Generation, Incremental Fine-Tuning, and Verification.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Strategies for Improving NL-to-FOL Translation with LLMs: Data Generation, Incremental Fine-Tuning, and Verification

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.784006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.784006Z digest=sha256:671f21b11aece4120393a18f9ee0968e31ee68ce8c02a867fce3b2eaedde5485

Observation b98684d4-f6b6-4466-a858-9772f8c2be91 · outbound

This paper cites Solving math word problems with process- and outcome-based feedback.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Solving math word problems with process- and outcome-based feedback

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.787792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.787792Z digest=sha256:53078279990f662840ce789027f20ff572237d5a2398682a28ba8fc6e9dd02f5

Observation 353a5442-2b9c-4ae6-824e-3574bcc1efc0 · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.791295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.791295Z digest=sha256:27d4ac28286804eb23e9f58aa5f99c711e41ac7598a0ec68f754438a2b900ce6

Observation 3920aab4-2a9c-41be-80ec-c2477fbf92bd · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.794880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.794880Z digest=sha256:7e1fe3a932e1d2a20f8529cae8282142deed3dd099113ac9a99f3da1894f5028

Observation 972dbe22-2d59-4f61-8e6b-b249ffb6da92 · outbound

This paper cites Qwen3 Technical Report.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Qwen3 Technical Report

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.798870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.798870Z digest=sha256:a88910204e3dc5f22b1f1be680686b09cd42bd5a119d47e389520780eca0a608

Observation ea7e3f4b-556a-409a-88cd-7619fc670b71 · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:26:53.054868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T15:26:52.802869Z digest=sha256:f14a28fd97507841049e4f2d6eaf9c9a61b7f29abe68db329a64a11abfe40f16

Observation eb39a6cd-e778-4ec0-b42a-f65491dbd8f1 · outbound

This paper cites The Lessons of Developing Process Reward Models in Mathematical Reasoning.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling The Lessons of Developing Process Reward Models in Mathematical Reasoning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.807424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.807424Z digest=sha256:ed3b809bf15b8ff9ade941740fc2ab45203629b385b35b0a12f2d10fedb898cf

Observation 6bc34e11-3105-44b6-a9cb-79f84048d237 · outbound

This paper cites online" 'onlinestring :=.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling online" 'onlinestring :=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.811818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.811818Z digest=sha256:9297753de2809fd04337ef9900673e2324a96bf4a3f86d83f64420760b509e8f

Observation ba7e1584-de20-4324-abec-2241b515dea5 · outbound

This paper cites write newline.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling write newline

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.816300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.816300Z digest=sha256:b300069d4478c441dfd339dec109a7e91959707d6200caaef5cf1bdfbbb94f62

Pith citing papers

No inbound Pith citation observations are available.