Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:37:24.974376Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2507.15512.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:37:24.974376Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
54 of 54 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3fb0fe32-d866-4d42-a53a-1af59abac6f5 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models online" 'onlinestring :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e472068-85f0-41fa-85f2-924d6c0ea33c · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06ed4b1f-d4ed-4cc3-b07b-41f36c918a98 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00d2d16e-e232-4f97-b42b-9ff633586a45 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0888cbf8-91aa-4517-a4cf-70e55c111605 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Efficient Prompting Methods for Large Language Models: A Survey
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fec87c8d-ba97-4723-9c28-e1dca317d7b0 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 387a8b27-0780-4018-8b0b-d7d4ae9d039c · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf468863-344c-42fa-8c3d-fbdcf602386d · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf547cc7-471a-431c-a0fa-dd6c01112d2b · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Dynamic Parallel Tree Search for Efficient LLM Reasoning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b919a394-ae99-48ee-a1d0-fda14ca763de · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 79fffff3-dc27-4602-9fb6-b8e1030a7516 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8794fd29-cfe3-42de-8b6e-43d1c798a70c · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models The Llama 3 Herd of Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd5cbf11-bbb8-4adf-9c00-05a41dab1f82 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Measuring Mathematical Problem Solving With the MATH Dataset
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 270f58dd-1586-44d9-b353-99dc74c54396 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 32281b08-23e6-4e8c-8a35-9f3f0d922fae · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f219167-10ac-4d6a-82cf-37dfdece25e0 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models MindStar: Enhancing Math Reasoning in Pre-trained LLMs at Inference Time
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 122bc742-36ce-4876-a508-ca2ceaf498f3 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Scaling Laws for Neural Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd30e8d6-b71f-4d93-9b37-cce2ecf9e573 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Gonzalez, Hao Zhang, and Ion Stoica
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25fa7974-b8ca-4780-b4e5-9fc7d7d2fc39 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models From System 1 to System 2: A Survey of Reasoning Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93694884-1d37-48aa-9d6c-fe9a1fb6f47f · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94a41992-67de-4659-88cf-986d501d1dac · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7d89afc0-0d57-4e87-8861-e1e16611743d · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0dc4419a-3029-4923-8ceb-d323da4ce8fe · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8f289958-4208-4114-9432-419e4611a362 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation af9f2b31-e30b-4a7a-8262-590c58eb3025 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models s1: Simple test-time scaling
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de22a193-3e29-4776-8967-80c4c52146bb · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 54efe568-fdcf-4759-b2bb-ba8298b473d7 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3d9c8bdf-a42b-4db3-b90f-a95684e72fd2 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models MC-NEST: Enhancing Mathematical Reasoning in Large Language Models leveraging a Monte Carlo Self-Refine Tree
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ded48cee-0d04-43ff-beca-400b55807270 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fc70004-692f-4544-9468-1df91cef1faf · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 56eda410-0fa7-473f-936f-6e655336c53b · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 255745b6-cfa4-4de7-8ab4-231be8d982ba · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf571c63-c71a-4caa-946a-eea22154689b · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models PRMBench: A Fine-grained and Challenging Benchmark for Process-Level Reward Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31c72ada-d685-449e-8af8-0f8c0db39c97 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Gemma 3 Technical Report
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 942872f9-e760-4cfb-8a5f-5c709f276273 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Solving math word problems with process- and outcome-based feedback
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe7517a8-eff5-442b-81c5-822196e81a48 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ea2de8a-6282-446c-b3c8-dcb133f67084 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1d1a203-5b25-4b88-83ee-0776460c306d · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15a8f3ad-d36f-491d-b35c-69e2c695b8dd · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Le, Ed H
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 01f4e93f-2cbf-4427-8840-5990a60af7c4 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Chi, Quoc V
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 079a00e7-b7f0-40a1-8453-95ada4905eb1 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 888b3b45-63a5-4bf8-a25e-c57afb5d0ea5 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 47daccca-fc97-4a3f-a989-a7a8b31a624a · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Foundations of Large Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 663df73f-68df-4176-9922-391a9287f81a · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Qwen2.5 Technical Report
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a27e4034-1814-4d89-afa5-11b468fa5a14 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 40a8dd28-cd25-4b90-86d0-9a9856adc807 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1ecf1ba-7fac-4019-a561-0bda22fb1dce · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98a9c10f-eba1-46ef-b98e-3ff54639620e · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1bfd7f65-1aeb-45a8-a77f-4127ea42aed0 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 109efe03-ab17-4f2a-b832-038edf186e46 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models A Survey on Test-Time Scaling in Large Language Models: What, How, Where, and How Well?
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f21145a-2da8-4827-b183-28c5bb86f8d7 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Small Language Models Need Strong Verifiers to Self-Correct Reasoning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d07800da-c014-48f5-a681-d6d9f3143dd2 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models The Lessons of Developing Process Reward Models in Mathematical Reasoning
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afad08a3-61b1-4782-8c08-1af9a7c37026 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb2dd002-b98d-4d15-a457-1b9382094597 · outbound
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models Learning to Reason via Mixture-of-Thought for Logical Reasoning
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.