Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:11:46.192441Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2506.00396.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:11:46.192441Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f6377368-b056-412a-96ae-e14ccad40c31 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively URL: " 'urlintro :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33a3c908-e5a4-4da1-9430-d564e72bd382 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4499d80-d173-495d-aa62-b5e11faf4df3 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Graph of Thoughts: Solving Elaborate Problems with Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a774c6f8-9cbb-4f70-a7eb-df45b99c2156 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Accelerating Large Language Model Decoding with Speculative Sampling
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc94f70a-b3fa-48af-abc8-c7a354fcb0f0 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively FinQA: A Dataset of Numerical Reasoning over Financial Data
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a086ed3-0756-41da-896f-065a4a73b3f6 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Training Verifiers to Solve Math Word Problems
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95746e22-c505-4434-a73e-f0f025859af1 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4bb8b3b9-fc0f-4f72-a8fd-e7e244cb4162 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Everything of Thoughts: Defying the Law of Penrose Triangle for Thought Generation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e2df33f-7c43-4e80-bbd8-fd38cc1f67d5 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively The Llama 3 Herd of Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 188d4c1d-6f1a-4d79-a559-54be10710330 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Reasoning with Language Model is Planning with World Model
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75a67c94-5ce5-4074-bbb4-7174c7500b37 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Large Language Models Cannot Self-Correct Reasoning Yet
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb996b5b-de16-4817-a101-d0dafa3b1dfe · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cd972443-e59a-40ce-b06d-3303f40eea88 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Reward Design with Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac6ecb87-4432-4707-b893-a0045ec5f7c2 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 62239cb0-9347-4121-8d7d-05f951a4c1c9 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 73869741-03f6-4365-8c9b-9dc415079e0e · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively GPT-4 Technical Report
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4055b49f-ec1b-4352-84a0-f9ce9dd6a23b · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 49912d8b-c8cf-4072-bdc1-e0727f766815 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65b26389-f98b-4547-aab2-6754cc213b0b · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4790b78-16e9-45ec-bb8b-b429c0fc16df · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b9aaa710-1c47-41ed-9dd1-89d68d6435ff · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdbd1863-aabd-4658-bf01-a14a595989a2 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b52a0d0b-9606-428e-8a7f-0f4c3a020837 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 25ecfa1a-a068-49e0-beea-9488145e4ca5 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54b69683-c3c1-46c1-919f-080e91e0a93d · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20b8ee25-6ca8-437f-a48c-90b04c7aca29 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22ae8d7f-691c-417f-975f-b2e50f5f1b5c · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively PRMBench: A Fine-grained and Challenging Benchmark for Process-Level Reward Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06ebea2b-a996-450d-90de-d4a4cf82c051 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dce3f3e-4cef-4567-9e38-923e7a5aaa63 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 58fb2e04-6789-4528-ab20-838d42e59af2 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00d41d8c-0668-40df-9829-13e3d9fee315 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22c6d246-76e5-403c-be61-4cca2e589cfc · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Chain-of-Table: Evolving Tables in the Reasoning Chain for Table Understanding
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c97218ba-3f3b-4c83-9e10-22a1c2722c86 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02aadac5-3a10-41c4-ab28-8da1516f555c · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26fa552f-2b33-4684-9e9a-661ef3c1d765 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7e877f1d-7658-4cd9-8983-0f96efd48657 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77ab271b-cb31-40fb-a81d-147fbdc2ca4a · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d01445e-2841-406d-808a-45f48acc6633 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Tree of Thoughts: Deliberate Problem Solving with Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb7fb42d-6cd0-4eca-a422-d9b0e69eddbf · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23584c56-45e1-40e6-a1e6-fb8c3a2c20d5 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively ToolChain*: Efficient Action Space Navigation in Large Language Models with A* Search
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb77b58f-2502-4520-af81-f354e4335d57 · outbound
Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.