Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:13:55.402176Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 1 inbound Pith citation observation for arXiv:2506.03295.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:13:55.402176Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-10T05:23:08.478393Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-10T09:23:37.521970Z
34 of 34 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f61a3d36-b194-4a19-80e4-5fe8042ad77e · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem online" 'onlinestring :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bddc807-a88e-4a84-8e28-f448c411a530 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 032176ae-7252-468b-9ff9-fb94657eb39b · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Phi-4-reasoning Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64dc34c3-5951-4705-aec2-9f1d5c9b3859 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem GPT-4 Technical Report
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ceb413f-2a7b-4c3b-b255-b662ea55ec38 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f5a9ef3e-62e5-40ab-b996-0d4c84c0436d · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8d4d64b-3532-482b-a0a8-9cd360d41574 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem SuperGPQA: Scaling LLM Evaluation across 285 Graduate Disciplines
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c3c6318-4650-477a-826d-0a6ef184648c · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ad00419f-80f7-4192-9c27-d8b928d75df3 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e35f8357-73d7-4ba2-a547-0b006bb08109 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f36f3d9b-9820-40ef-b053-f8c2ce170c51 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Measuring Mathematical Problem Solving With the MATH Dataset
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b47a0b5-9e12-4b7c-92b3-386e12b15de2 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef3b0a0a-d011-4cc8-8d46-32c24501447a · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Transient Non-Stationarity and Generalisation in Deep Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95815a1c-d9b2-4c25-ac55-c5ec47921bb3 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem OpenAI o1 System Card
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dab77a8c-5779-48ce-9821-740a9524f26c · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem BIG-Bench Extra Hard
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18c74c37-490f-495c-8640-0c7dd3b64025 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0be86bb8-dca2-4b9f-9e6b-4ee186768982 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c458e86f-bf83-4911-8d45-be01a33a7d17 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem General-Reasoner: Advancing LLM Reasoning Across All Domains
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8850b35-548f-4380-a5f3-600086c3753c · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem s1: Simple test-time scaling
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2e15b19-515e-4dd6-ad16-7b64c773c0b9 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation be07fda1-f257-458c-b527-5313acd2bbc2 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fe1384b7-7f29-4423-85e6-839528e25287 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cb61302-0316-49ef-8f2a-b4ef49a9bdfe · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5ff1a335-898a-4926-b228-1eeacf38255a · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d2fe63a-e74e-4ea6-b315-b09ac42ba3c4 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb53f15c-9bda-41fe-b338-77fba178d538 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Reinforcement Learning for Reasoning in Large Language Models with One Training Example
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c47b0fc3-0508-435c-aa00-a904a00345d2 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a427ab1-20d6-42e5-875b-457b07433dc2 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Critique Fine-Tuning: Learning to Critique is More Effective than Learning to Imitate
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b679c15-fa0b-45ea-9eb4-314ab6cb5bf7 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem MiMo: Unlocking the Reasoning Potential of Language Model -- From Pretraining to Posttraining
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72345245-f1cc-42a6-84b1-174caee69774 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Qwen3 Technical Report
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a341be0-3ea0-4c17-9ce4-a97b22bc3daa · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a5f9fc7-85ed-475f-8b00-cb6c7d8884c3 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem LIMO: Less is More for Reasoning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfc1964e-dd6b-4193-892f-3d6fbb2eff31 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 475ea730-30f7-4685-a1f9-f66e4a87bee5 · outbound
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 444f8f0e-d11a-4ce5-a8bc-e6279d55bb8a · inbound
HEALing Entropy Collapse: Enhancing Exploration in Few-Shot RLVR via Hybrid-Domain Entropy Dynamics Alignment Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.