Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 100 inbound Pith citation observations for arXiv:2501.04519.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:30:38.008772Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
8
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 3b9c28fe-e1d3-42d1-966e-c279540ba545 · inbound
Stepwise Reasoning Error Disruption Attack of LLMs rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae3cd6c2-b95c-4168-b2ff-fe81fead2e87 · inbound
Reasoning Language Models: A Blueprint rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 034e27fd-824d-48e9-bf03-ea75755be632 · inbound
Scaling Inference-Efficient Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 586d5cd4-62e1-4a27-bd81-499fb27f53dd · inbound
Efficient Multi-Agent System Training with Data Influence-Oriented Tree Search rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1d2a5277-40df-4e3c-a160-c552a7fc7537 · inbound
LongDPO: Unlock Better Long-form Generation Abilities for LLMs via Critique-augmented Stepwise Information rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bf0e9b3-3ecd-4465-98cd-4db3051e59e0 · inbound
Brief analysis of DeepSeek R1 and its implications for Generative AI rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05159511-cc5c-4ca3-aad5-3350ce6dfb06 · inbound
Safety Reasoning with Guidelines rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 996155ce-2629-4548-a377-bcbfe0093546 · inbound
CodeSteer: Symbolic-Augmented Language Models via Code/Text Guidance rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f97b4a9-2d70-4e09-8a11-3d0d9e19b7cc · inbound
Improving Language Models with Intentional Analysis rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d9fb27c5-3ba9-4b3c-a5b9-5fc998f8172d · inbound
Iterative Deepening Sampling as Efficient Test-Time Scaling rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb67e40f-4cf7-42ae-8179-a6522be88ea6 · inbound
PIPA: Preference Alignment as Prior-Informed Statistical Estimation rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e1abdd1-50fb-48bf-8a14-e2f02454c213 · inbound
Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7e1636c-8b03-4263-91b1-ae5239e0bc6a · inbound
On the Emergence of Thinking in LLMs I: Searching for the Right Intuition rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1593a911-7d7b-4aa5-ba85-16fd0f66797b · inbound
Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0db4016a-38d6-4cf6-ae2a-a510da52520a · inbound
The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 654b6651-bd04-4583-a1d9-2df3143df84d · inbound
Typhoon T1: An Open Thai Reasoning Model rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a43164e7-7775-49b3-b0df-02a5a357d75e · inbound
CoT-Valve: Length-Compressible Chain-of-Thought Tuning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67a98d22-7542-4dc7-9e50-1c9897cc54c2 · inbound
From System 1 to System 2: A Survey of Reasoning Large Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 143
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 14267516-3810-4312-bc97-390ab80baef5 · inbound
Meta-Reasoner: Dynamic Guidance for Optimized Inference-time Reasoning in Large Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d9affb95-ebff-4890-bb90-3ffdcfd77928 · inbound
Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 226
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 14c552b0-c3f1-4f16-aa6b-e44d33d2395b · inbound
Phi-4-reasoning Technical Report rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0ead29b4-f399-46dc-a0f7-0837fd53f05c · inbound
TokUR: Token-Level Uncertainty Estimation for Large Language Model Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 50cd3b33-c295-4122-838b-a37faeff1e13 · inbound
DiagnosisArena: Benchmarking Diagnostic Reasoning for Large Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcd871aa-cf77-468b-981c-26d38e7d523a · inbound
TemplateRL: Structured Template-Guided Reinforcement Learning for LLM Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0b1ccf3e-7388-41e7-bc96-c6e6f01be5f8 · inbound
Learning to Reason via Mixture-of-Thought for Logical Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8694828c-2bb2-4f8b-acde-9d3ae93b6fb4 · inbound
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38b3018e-0145-4e90-a9ac-f5ecea67b504 · inbound
Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bedae641-e22c-4a95-97c6-a0b014872ec4 · inbound
VeriThinker: Learning to Verify Makes Reasoning Model Efficient rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ac5c1b8-bbad-4caf-b8b5-86409613798d · inbound
RaDeR: Reasoning-aware Dense Retrieval Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55471775-4c12-4441-a108-254b5ab1853e · inbound
MMATH: A Multilingual Benchmark for Mathematical Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf166085-f870-4f4d-bd71-c6c1dacb4a81 · inbound
Faster and Better LLMs via Latency-Aware Test-Time Scaling rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e200c4b-2622-4226-bfdb-4d622dc54603 · inbound
Concise Reasoning, Big Gains: Pruning Long Reasoning Trace with Difficulty-Aware Prompting rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1216ef42-4154-4d8a-b68e-e40f0f23babc · inbound
Can Past Experience Accelerate LLM Reasoning? rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6db72fc4-c218-4d06-a80d-91872720b3d0 · inbound
rStar-Coder: Scaling Competitive Code Reasoning with a Large-Scale Verified Dataset rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f779bc17-e800-4b2d-bdc9-34e5539ba275 · inbound
UI-Genie: A Self-Improving Approach for Iteratively Boosting MLLM-based Mobile GUI Agents rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d721ffe-0a6c-461f-9e63-ff41c627315e · inbound
RAG-Zeval: Towards Robust and Interpretable Evaluation on RAG Responses through End-to-End Rule-Guided Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f7f1c3c-cd7a-45be-ad00-ebd75ad3a8b4 · inbound
How Much Backtracking is Enough? Exploring the Interplay of SFT and RL in Enhancing LLM Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 317aff50-0850-44ce-bba9-846ccbf27ed4 · inbound
Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation abc0b61a-31c9-4586-8bf7-70f503b7e8fe · inbound
One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94f56415-7df1-433c-b465-91a8350e4ce4 · inbound
AdaDecode: Accelerating LLM Decoding with Adaptive Layer Parallelism rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82381dd6-b3cb-4230-a172-84bd0c8bfe11 · inbound
Structured Pruning for Diverse Best-of-N Reasoning Optimization rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e199341-7a67-43c8-a4fa-e3400918a3eb · inbound
DynamicMind: A Tri-Mode Thinking System for Large Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffc07461-ffdc-44a3-8845-da654ca7584f · inbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfdd0ab2-dde6-46e3-a06c-933d250e130b · inbound
A Survey on Large Language Models for Mathematical Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37401fbc-e87f-498d-9c1b-9519b094ab57 · inbound
Socratic-MCTS: Test-Time Visual Reasoning by Asking the Right Questions rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06e02db3-49d4-4303-b2bb-2dae77510a5a · inbound
SwS: Self-aware Weakness-driven Problem Synthesis in Reinforcement Learning for LLM Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 301449ee-3612-4f6f-a783-45d2679d11f0 · inbound
TTT-Bench: A Benchmark for Evaluating Reasoning Ability with Simple and Novel Tic-Tac-Toe-style Games rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59d93812-8bd4-4b1c-8421-7c8c5550685c · inbound
MCTS-Refined CoT: High-Quality Fine-Tuning Data for LLM-Based Repository Issue Resolution rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05224726-4f85-4934-aa07-8d6c13110327 · inbound
Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2656d7cb-e0c5-402f-bb80-5de2452c58f7 · inbound
DuaShepherd: Integrating Stepwise Correctness and Potential Rewards for Mathematical Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e85fecc2-2d41-4bdd-ad63-eef3dcabfd24 · inbound
Distilling Tool Knowledge into Language Models via Back-Translated Traces rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d26efe52-e749-4d62-9209-0bb1d6082afe · inbound
Towards Understanding the Cognitive Habits of Large Reasoning Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8226f95e-f046-4cbf-919f-72c19f342603 · inbound
Test-Time Scaling with Reflective Generative Model rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c2624e4-c766-48db-adb9-7fdc2332769a · inbound
Enhancing Test-Time Scaling of Large Language Models with Hierarchical Retrieval-Augmented MCTS rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3113f3ab-8d7c-4976-bfb2-2ace325cef0b · inbound
From Language to Logic: A Bi-Level Framework for Structured Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 906
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 474e7e1f-ebec-4769-90c1-086e16e92013 · inbound
EduFlow: Advancing MLLMs' Problem-Solving Proficiency through Multi-Stage, Multi-Perspective Critique rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dfdf8e0-b155-410e-8274-540a77f81581 · inbound
AI-Powered Math Tutoring: Platform for Personalized and Adaptive Education rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d753e3c7-eb92-43de-a765-31c634f38b70 · inbound
Dynamic and Generalizable Process Reward Modeling rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dfeb8e6-c345-49ff-9bc1-20b33a167609 · inbound
AQuilt: Weaving Logic and Self-Inspection into Low-Cost, High-Relevance Data Synthesis for Specialist LLMs rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdb1f98b-6f30-4fc9-aa4c-de8503fd383a · inbound
A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 234
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9f30538b-8588-4cba-add5-7cc3256747de · inbound
StructVRM: Aligning Multimodal Reasoning with Structured and Verifiable Reward Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9fac228-abbe-4aac-8de2-91b4c37dae0f · inbound
An Explainable Machine Learning Framework for Railway Predictive Maintenance using Data Streams from the Metro Operator of Portugal rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62909ce8-c75d-4330-8c4d-8d2d6a3e3c6e · inbound
Sample-efficient LLM Optimization with Reset Replay rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1ab3c225-c272-492f-8113-19d1233e2401 · inbound
InteChar: A Unified Oracle Bone Character List for Ancient Chinese Language Modeling rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e6a9ace-d75a-451e-b5c1-cc2d3f741e21 · inbound
rStar2-Agent: Agentic Reasoning Technical Report rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e887d306-6f75-4c59-9230-544deff51a20 · inbound
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 202
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 255eb9be-145d-4f37-a01e-831da3972f87 · inbound
ParaThinker: Native Parallel Thinking as a New Paradigm to Scale LLM Test-time Compute rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c04d4468-5067-4343-8006-c47994ef6c65 · inbound
DeepSearch: Overcome the Bottleneck of Reinforcement Learning with Verifiable Rewards via Monte Carlo Tree Search rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2b12d6ce-7fe1-40c5-baf4-25f423b5945b · inbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0c1c197b-2910-421a-b4f6-af84b593fac8 · inbound
MASPRM: Multi-Agent System Process Reward Model rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7b835e5-3650-4486-af52-465b278ec30d · inbound
Sharpness-Guided Group Relative Policy Optimization via Probability Shaping rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4e046aa7-c60f-4465-9b24-55412e400030 · inbound
TreeCoder: Systematic Exploration and Optimisation of Decoding and Constraints for LLM Code Generation rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 152560ad-ddcc-4670-a512-74c44d9220e2 · inbound
PyFi: Toward Pyramid-like Financial Image Understanding for VLMs via Adversarial Agents rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4ca6de60-51a9-4c36-ab8b-9c0b293768c0 · inbound
On the Overscaling Curse of Parallel Thinking: System Efficacy Contradicts Sample Efficiency rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2a65ce3f-459f-4877-9f61-32c6531bc8d9 · inbound
A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45829dcc-6a81-4821-b00a-25494ece48f0 · inbound
Can I Have Your Order? Monte-Carlo Tree Search for Slot Filling Ordering in Diffusion Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d82cbb9a-454d-484b-972c-986e1b46981a · inbound
SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bda38c06-9ec6-46ea-8562-c5027c81e588 · inbound
SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76a04c12-74a7-4501-a022-b6575dd8aab0 · inbound
CoTEvol: Self-Evolving Chain-of-Thoughts for Data Synthesis in Mathematical Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bf86c04b-90ac-4460-a3bf-f6cb036765e5 · inbound
Fine-Tuning Small Reasoning Models for Quantum Field Theory rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 84324a20-d793-4820-9f8a-78fecad56a1f · inbound
A Limit Theory of Foundation Models: A Mathematical Approach to Understanding Emergent Intelligence and Scaling Laws rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 48368654-04ff-49e0-84d2-81553fe4a869 · inbound
A Limit Theory of Foundation Models: A Mathematical Approach to Understanding Emergent Intelligence and Scaling Laws rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 431bed34-8c12-4e7b-b2f4-f4f571024482 · inbound
IRIS: Interleaved Reinforcement with Incremental Staged Curriculum for Cross-Lingual Mathematical Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 980b8362-b7da-4dd0-943b-77bf37fa5c77 · inbound
Confidence-Aware Alignment Makes Reasoning LLMs More Reliable rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f820cb88-4dff-4174-a2d3-7ab13abeaacc · inbound
Mathematical Reasoning via Intervention-Based Time-Series Causal Discovery Using LLMs as Concept Mastery Simulators rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 68428fb5-62ff-4d18-b2d9-6d86d8def2e4 · inbound
PriorZero: Bridging Language Priors and World Models for Decision Making rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 90b203d4-4837-4398-ba2a-202f27da154a · inbound
Many-Shot CoT-ICL: Making In-Context Learning Truly Learn rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a8d5e88f-8379-4512-ab09-629fe4024d65 · inbound
Graphs of Research: Citation Evolution Graphs as Supervision for Research Idea Generation rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 335db1e3-ef65-49d7-a9e9-c6de0432b6f6 · inbound
PopuLoRA: Co-Evolving LLM Populations for Reasoning Self-Play rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d890c629-6899-4748-9604-6f67267dcb76 · inbound
STRIDE: Learnable Stepwise Language Feedback for LLM Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 89f1e9a3-e938-492c-bd54-04a02c7bbea4 · inbound
Mathematical Reasoning in Large Language Models: Benchmarks, Architectures, Evaluation, and Open Challenges rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdd2fe36-2bab-4888-9b2f-9ea87f2b85d7 · inbound
When Self-Belief Misleads: Active Label Acquisition for Reinforcement Learning with Verifiable Rewards rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0eaae3cb-75a9-4dd7-8adf-e5e48f807740 · inbound
Learning to Adapt SFT Data for Better Reasoning Generalization rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1ea2f4e8-f7a8-4aec-aab2-c024f249b556 · inbound
Efficient Test-time Inference for Generative Planning Models with OCL Search rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 59d40125-828c-4873-971f-b962ae58cf41 · inbound
From Answers to States: Verifiable Process-Level Evaluation of Chemical Reasoning in Large Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 87c65115-8ac5-42d5-b79b-9fea48415870 · inbound
Step-by-Step Optimization-like Reasoning in LLMs over Expanding Search Spaces rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9e8d504b-0622-434f-a03e-94d45c695e4a · inbound
Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 90d15e96-d409-4d4a-9b50-c4a9c5f5f5ca · inbound
Improving Multimodal Reasoning via Worst Dimension Optimization rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d65ea05a-f974-467a-a5e5-e3cd492beba8 · inbound
CATPO: Critique-Augmented Tree Policy Optimization rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ed7f0660-4b5f-487c-bb4d-736115358585 · inbound
The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.