Pith. sign in

Paper Citation Record · LEDGER

It's Not That Simple. An Analysis of Simple Test-Time Scaling

As of 7 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 0 inbound Pith citation observations for arXiv:2507.14419.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.14419 v1

Coverage vector

measured 27 of 27 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:09:06.934747Z

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

27 of 27 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1616d673-4dd0-485b-985a-3085911dca6d · outbound

This paper cites KCTS: Knowledge-Constrained Tree Search Decoding with Token-Level Hallucination Detection.

It's Not That Simple. An Analysis of Simple Test-Time Scaling KCTS: Knowledge-Constrained Tree Search Decoding with Token-Level Hallucination Detection

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:04.122606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:04.122606Z digest=sha256:ce72768b2fe0e5052709cddb0db7e7e0bca7d4b356bbb7e16d75c1cec120f2f6

Observation 5231d89a-6cc3-4559-b7ce-2cdb49bf5719 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

It's Not That Simple. An Analysis of Simple Test-Time Scaling DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:04.187424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:04.187424Z digest=sha256:9c30b44f243a2cc80e8e705eb7cc706cc39d2179db5d2c133be1cdbf79c65214

Observation 2a39a040-6db1-4621-9b31-ed1e1b145f3a · outbound

This paper cites Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:04.281051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:04.281051Z digest=sha256:3a37947fa1c6ba2852db2615d0c069e1ed7f3ffdf0f1c1bff89dd61c456ebf5a

Observation 5bd1b6a5-fe47-4c10-9c4d-37b73e4d41e1 · outbound

This paper cites Stream of Search (SoS): Learning to Search in Language.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Stream of Search (SoS): Learning to Search in Language

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:04.402455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:04.402455Z digest=sha256:4df4e2385a9f393b8cec400de80bfc0244f5ee597de782e4dfd99abf3ae01418

Observation 3a7b2f06-1c87-4437-b501-b5ad243d8677 · outbound

This paper cites Accessed: 2025-03-26.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Accessed: 2025-03-26

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:09:08.002681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:09:04.473476Z digest=sha256:48ab6c8ad4cc5afa08ba1310ab4b1250f7238d45811905f9066577eff905d99a

Observation 2b2eba29-5b23-454f-9ea9-301c316c0e1f · outbound

This paper cites Rewarding Chatbots for Real-World Engagement with Millions of Users.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Rewarding Chatbots for Real-World Engagement with Millions of Users

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:04.740135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:04.740135Z digest=sha256:8c7ef1b164e62c60c0e5e5b50d37fba7b252a564cb25c3f7a0954359306b0de1

Observation 8d5495d8-dcf5-49f1-a9b5-ad9236aa4942 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:04.830836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:04.830836Z digest=sha256:d35732496b8e5ac411d9ce52c5222c89db7c29699499ce7db86f97a902f4a33e

Observation dba945db-1dd1-4fc4-8e32-1c4c6db77b99 · outbound

This paper cites Don't throw away your value model! Generating more preferable text with Value-Guided Monte-Carlo Tree Search decoding.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Don't throw away your value model! Generating more preferable text with Value-Guided Monte-Carlo Tree Search decoding

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:04.911668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:04.911668Z digest=sha256:d2114e3f0d4ae51f73c3d62a80456930d249b829b8aaba9f26f76da666799a47

Observation e642f830-5024-48fb-a6d2-eeb6e5876820 · outbound

This paper cites s1: Simple test-time scaling.

It's Not That Simple. An Analysis of Simple Test-Time Scaling s1: Simple test-time scaling

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:05.062891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:05.062891Z digest=sha256:1a382ec1075648c7e8d61feb2f8139c3b38bea8d66c6b7f6e4b8991c35ced7ef

Observation 8d851a03-0dea-4b3c-b7f1-b53a3a1975fb · outbound

This paper cites Show Your Work: Scratchpads for Intermediate Computation with Language Models.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Show Your Work: Scratchpads for Intermediate Computation with Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:05.173663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:05.173663Z digest=sha256:42a5d8fbaa20f7a61a3fb9d2e6b0a56ec00f2f36bb3b65c865365b490865e4b0

Observation 425dbdef-37d9-4bce-8336-a06c7efbd413 · outbound

This paper cites Accessed: 2025-03-26.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Accessed: 2025-03-26

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:09:07.825754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:09:05.313138Z digest=sha256:7d7d205a27765b81b040b293912abf2fe66d6454144f69f1c331d74332b73a65

Observation 9846710a-07a5-49e6-91e8-eee876a6c015 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:05.450870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:05.450870Z digest=sha256:8a7ad6fb0c3c34d30ff885541c9dfc3dd8de0ff304fd9be3456222989f68c6d9

Observation e490b109-c17f-4cd5-a8c2-c1fc2c6ca889 · outbound

This paper cites GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems.

It's Not That Simple. An Analysis of Simple Test-Time Scaling GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:05.621221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:05.621221Z digest=sha256:ec5a09da2c98bbcebb82c7806314d5489cc3a11473a7c0299fe022e795d7d4ec

Observation 13dce90c-d178-4a3d-af4b-f92c5057619c · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

It's Not That Simple. An Analysis of Simple Test-Time Scaling LLaMA: Open and Efficient Foundation Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:05.727718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:05.727718Z digest=sha256:5adafc41fd7f1680c5371e508b7ac2e4d9453a8431d49a24b7c274262c032e95

Observation 2f697e7b-534d-47fd-8537-d7a287e2882e · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:05.883911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:05.883911Z digest=sha256:8ad7cfed3a2b704357acd216340ec3154961bc721a5b451be77a2d859caff25e

Observation 862812bc-f356-4f09-840b-3bab5c256cfe · outbound

This paper cites Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:06.009237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:06.009237Z digest=sha256:3399a3a667751b26493b402e416bdf28def965f96dd54fcea28aac44247fe10a

Observation 0d24e271-2bb9-4483-b848-69cc9b7c9ff8 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:06.171068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:06.171068Z digest=sha256:a7e017a0d05926c60e6ccff80b1647b09b9603cc94da16059da4a6c8e33832d8

Observation 0585e424-feaf-4a11-a316-e6e23ca7574f · outbound

This paper cites Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:06.343306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:06.343306Z digest=sha256:ad8b1677109bb92ddda750d794becb4054ce98b867b33dd1ba7ef0869e0ce95b

Observation 9c665c9f-b7e4-4dc8-8d47-e0efc2e95bb7 · outbound

This paper cites DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search.

It's Not That Simple. An Analysis of Simple Test-Time Scaling DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:06.494941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:06.494941Z digest=sha256:7de873c05812c73feed7743e37368e7cf1803c02abf1aee62153b8b8b5b83e52

Observation 49d7a066-9c35-4052-b442-a4c847089f4f · outbound

This paper cites LIMO: Less is More for Reasoning.

It's Not That Simple. An Analysis of Simple Test-Time Scaling LIMO: Less is More for Reasoning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:06.614864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:06.614864Z digest=sha256:a056acf848d92ab4375c396fdd5b09476397ba169fd17a26040aa8ccafc460d9

Observation 2c04983f-9390-4f99-90d9-f94622b3c594 · outbound

This paper cites Demystifying Long Chain-of-Thought Reasoning in LLMs.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Demystifying Long Chain-of-Thought Reasoning in LLMs

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:06.765738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:06.765738Z digest=sha256:b5d0441a23884734d20dff5bd9ef681aba02556954e9809d358f93abbb71989d

Observation bbdf5ee1-a7e9-4875-926d-c0dd4acb0659 · outbound

This paper cites Planning with Large Language Models for Code Generation.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Planning with Large Language Models for Code Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:06.934747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:06.934747Z digest=sha256:1050fd2e3fb3775745c5ac9c5cb263339976e493428f110c533e27f4d71db241

Observation c9565a23-3c7d-4e6c-ad50-fe44218488f9 · outbound

This paper cites Xingyu Chen, Jiahao Xu, Tian Liang, Zhiwei He, Jianhui Pang, Dian Yu, Linfeng Song, Qiuzhi Liu, Mengfei Zhou, Zhuosheng Zhang, Rui Wang, Zhaopeng Tu, Haitao Mi, and Dong Yu.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Xingyu Chen, Jiahao Xu, Tian Liang, Zhiwei He, Jianhui Pang, Dian Yu, Linfeng Song, Qiuzhi Liu, Mengfei Zhou, Zhuosheng Zhang, Rui Wang, Zhaopeng Tu, Haitao Mi, and Dong Yu

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:09:08.245962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:09:04.026907Z digest=sha256:3316f6195e95779b66f3cd2832ef0a97233d65820d28bebef66764aa0ff6399c

Observation b2ab9dd1-5f22-4f82-a79b-99d6b2e5e7bf · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Measuring Mathematical Problem Solving With the MATH Dataset

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:04.605352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:04.605352Z digest=sha256:da12a9f9a2f8a448df4c21d2456a9957946aaa1046ccea68825916803203e30d

Observation 1edf09c2-9f43-4316-8882-eeab12e7a859 · outbound

This paper cites Qwen Technical Report.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Qwen Technical Report

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:03.858313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:03.858313Z digest=sha256:511af6f078ff7401630512aab2b3fd3901ec2ebe58895913638e1f5afd4d26b1

Observation 66639b2f-fa3d-478d-9316-11e70abed175 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:03.927810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:03.927810Z digest=sha256:e389d38726e3385ee1e5dbeb7af61e347e59c2dd491c29b399a6efbedf1f9c2f

Observation 57540151-6150-408d-8677-2867d09389e4 · outbound

This paper cites Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs.

It's Not That Simple. An Analysis of Simple Test-Time Scaling Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:04.090730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:04.090730Z digest=sha256:6a44f2040e48b2bed4fc3fe46299b75cec24df5e3e851f7b15f6c4a3e3573b18

Pith citing papers

No inbound Pith citation observations are available.