Pith. sign in

Paper Citation Record · LEDGER

rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 100 inbound Pith citation observations for arXiv:2501.04519.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.04519 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 100 of 107 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:30:38.008772Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

8
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3b9c28fe-e1d3-42d1-966e-c279540ba545 · inbound

Stepwise Reasoning Error Disruption Attack of LLMs cites this paper.

Stepwise Reasoning Error Disruption Attack of LLMs rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-11T14:30:38.008772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:30:38.008772Z digest=sha256:c1649bb9b789be0460c0d2159bb8abd2786878782524eb8c6405be71869de7f6

Observation ae3cd6c2-b95c-4168-b2ff-fe81fead2e87 · inbound

Reasoning Language Models: A Blueprint cites this paper.

Reasoning Language Models: A Blueprint rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T18:36:54.504256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:36:54.504256Z digest=sha256:de64d745b186ee02790c5c2c71484fabd44cee30e94286e1e15e0be4e53da52b

Observation 034e27fd-824d-48e9-bf03-ea75755be632 · inbound

Scaling Inference-Efficient Language Models cites this paper.

Scaling Inference-Efficient Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T00:43:28.905055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:43:28.905055Z digest=sha256:f8ea067252e66b201d248167833bf687a5cb90dbead8acdc3c602c1d99c80c2c

Observation 586d5cd4-62e1-4a27-bd81-499fb27f53dd · inbound

Efficient Multi-Agent System Training with Data Influence-Oriented Tree Search cites this paper.

Efficient Multi-Agent System Training with Data Influence-Oriented Tree Search rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-23T04:07:30.497780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-23T04:06:23.521344Z digest=sha256:3e4bd9169408be595e4d7e1dc70c715a8ad645de8950335e3ff08629bbbc0482

Observation 1d2a5277-40df-4e3c-a160-c552a7fc7537 · inbound

LongDPO: Unlock Better Long-form Generation Abilities for LLMs via Critique-augmented Stepwise Information cites this paper.

LongDPO: Unlock Better Long-form Generation Abilities for LLMs via Critique-augmented Stepwise Information rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T13:25:52.006995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T13:25:52.006995Z digest=sha256:35bf425922402a56fb155885fd5cadbbddbb031376173e2dca785a25d8000a98

Observation 4bf0e9b3-3ecd-4465-98cd-4db3051e59e0 · inbound

Brief analysis of DeepSeek R1 and its implications for Generative AI cites this paper.

Brief analysis of DeepSeek R1 and its implications for Generative AI rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T11:53:29.385817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:53:29.385817Z digest=sha256:52c79e92c6ed99c05a8791693f7036fb444cfa9936129b340ccaec68e6b1f731

Observation 05159511-cc5c-4ca3-aad5-3350ce6dfb06 · inbound

Safety Reasoning with Guidelines cites this paper.

Safety Reasoning with Guidelines rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T23:50:35.741225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T23:50:35.741225Z digest=sha256:14bdc7d35eb3815cef4b94a0dc54da1d35ea7ff5bae85a2664048da86224493a

Observation 996155ce-2629-4548-a377-bcbfe0093546 · inbound

CodeSteer: Symbolic-Augmented Language Models via Code/Text Guidance cites this paper.

CodeSteer: Symbolic-Augmented Language Models via Code/Text Guidance rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T12:16:27.380681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T12:16:27.380681Z digest=sha256:054d4d6e34f83407455ff26fcfb9b09760842c7ec66b22ecc3cf2e6ce67de3bc

Observation 2f97b4a9-2d70-4e09-8a11-3d0d9e19b7cc · inbound

Improving Language Models with Intentional Analysis cites this paper.

Improving Language Models with Intentional Analysis rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-05-23T03:45:21.598416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-23T03:44:33.625252Z digest=sha256:9f34d3423c79a73a537ec221664df3cb37985676c4229db58e02881954e6f25f

Observation d9fb27c5-3ba9-4b3c-a5b9-5fc998f8172d · inbound

Iterative Deepening Sampling as Efficient Test-Time Scaling cites this paper.

Iterative Deepening Sampling as Efficient Test-Time Scaling rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T19:23:26.305225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:23:26.305225Z digest=sha256:aadd9179e212cfeb916c28bd7ac7493290befd674bb231483a04dd36d42f0160

Observation bb67e40f-4cf7-42ae-8179-a6522be88ea6 · inbound

PIPA: Preference Alignment as Prior-Informed Statistical Estimation cites this paper.

PIPA: Preference Alignment as Prior-Informed Statistical Estimation rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T18:10:53.373963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:10:53.373963Z digest=sha256:544dc905e3a9b05792570a1c1dd0775501ca2c97df0a3f9989be25273aaeb277

Observation 7e1abdd1-50fb-48bf-8a14-e2f02454c213 · inbound

Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling cites this paper.

Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T14:40:35.941372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:40:35.941372Z digest=sha256:44980a5b061fcc032c986e021aac1b2b2749de2575991e0cb29d023b41eeed67

Observation c7e1636c-8b03-4263-91b1-ae5239e0bc6a · inbound

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition cites this paper.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.340082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.340082Z digest=sha256:e30018c5d53c0c116c3eb830e5f766595700c47665e4c9a3b8d509cab0525cfa

Observation 1593a911-7d7b-4aa5-ba85-16fd0f66797b · inbound

Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning cites this paper.

Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T14:27:51.099046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:27:51.099046Z digest=sha256:22de3efde360b76da046b5fbb97afea0198c8b35ab8044108a044790dc8a6e38

Observation 0db4016a-38d6-4cf6-ae2a-a510da52520a · inbound

The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks cites this paper.

The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T06:00:32.234438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T06:00:32.234438Z digest=sha256:cb966bf3d9af2272740e5ff65e9ed589e0a63f0d7dded969bfdd424de734c632

Observation 654b6651-bd04-4583-a1d9-2df3143df84d · inbound

Typhoon T1: An Open Thai Reasoning Model cites this paper.

Typhoon T1: An Open Thai Reasoning Model rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T22:52:54.318244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T22:52:54.318244Z digest=sha256:c5a0a6b4955ad813b0beae5f2b0808f6c71a9b00b286b985c3f2b5e03db166d6

Observation a43164e7-7775-49b3-b0df-02a5a357d75e · inbound

CoT-Valve: Length-Compressible Chain-of-Thought Tuning cites this paper.

CoT-Valve: Length-Compressible Chain-of-Thought Tuning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T20:57:42.034406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T20:57:42.034406Z digest=sha256:24901a674921f844ad3cd13c884512c0af9f4131626f58059072256f666884ed

Observation 67a98d22-7542-4dc7-9e50-1c9897cc54c2 · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 143

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:14a14d5c0acb2409b1cdcc6d38e89ba03e0fd2eaf4a8823396c729b0454700eb

Observation 14267516-3810-4312-bc97-390ab80baef5 · inbound

Meta-Reasoner: Dynamic Guidance for Optimized Inference-time Reasoning in Large Language Models cites this paper.

Meta-Reasoner: Dynamic Guidance for Optimized Inference-time Reasoning in Large Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-23T02:47:26.385462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-23T02:46:20.578680Z digest=sha256:db9f284b3be98f81ef27c66f629012e52107e1c727d969af4a77707b5407ab54

Observation d9affb95-ebff-4890-bb90-3ffdcfd77928 · inbound

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models cites this paper.

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 226

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-12T08:40:40.910461Z digest=sha256:27f6ced6da0f1eccff5d4359db55fd954dea9c786227c8e70f86fb22056553db

Observation 14c552b0-c3f1-4f16-aa6b-e44d33d2395b · inbound

Phi-4-reasoning Technical Report cites this paper.

Phi-4-reasoning Technical Report rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-17T03:40:25.706499Z digest=sha256:1354e74b798bea9757d73b2c00deeb6af97195d55e719b6e7e878359f3976fb9

Observation 0ead29b4-f399-46dc-a0f7-0837fd53f05c · inbound

TokUR: Token-Level Uncertainty Estimation for Large Language Model Reasoning cites this paper.

TokUR: Token-Level Uncertainty Estimation for Large Language Model Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-22T14:01:38.382405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T13:58:07.913104Z digest=sha256:3d8890bbc6aab9b6d48ebdbcbc503c26590ca305a9185b2d9b023cf7a418da1d

Observation 50cd3b33-c295-4122-838b-a37faeff1e13 · inbound

DiagnosisArena: Benchmarking Diagnostic Reasoning for Large Language Models cites this paper.

DiagnosisArena: Benchmarking Diagnostic Reasoning for Large Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:21.412901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:21.412901Z digest=sha256:9f48c8f36b2bcc4438879c7d253c92ca47df6b8b5e79201f05275413322b33f4

Observation bcd871aa-cf77-468b-981c-26d38e7d523a · inbound

TemplateRL: Structured Template-Guided Reinforcement Learning for LLM Reasoning cites this paper.

TemplateRL: Structured Template-Guided Reinforcement Learning for LLM Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-22T13:41:36.183624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T13:39:17.205077Z digest=sha256:82b8064d5b96a33cd7de38e40ff7f01e5f7b150767bfbfdc5a50aa9827fba036

Observation 0b1ccf3e-7388-41e7-bc96-c6e6f01be5f8 · inbound

Learning to Reason via Mixture-of-Thought for Logical Reasoning cites this paper.

Learning to Reason via Mixture-of-Thought for Logical Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.134427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.134427Z digest=sha256:773a0e10189450a5efed05329f0c3b18de48c7cb51c7caac84b518fa8c905e38

Observation 8694828c-2bb2-4f8b-acde-9d3ae93b6fb4 · inbound

Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models cites this paper.

Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:17.111371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:17.111371Z digest=sha256:9962d4b410ea171809dbf153465b7d0bd7b3ee1cf1ef3f81f27a3373a0d53923

Observation 38b3018e-0145-4e90-a9ac-f5ecea67b504 · inbound

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning cites this paper.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.786796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.786796Z digest=sha256:e43e6bc4afdfeff1f6ac710163020341549f452531ae3e8ded06cfe6b62d348b

Observation bedae641-e22c-4a95-97c6-a0b014872ec4 · inbound

VeriThinker: Learning to Verify Makes Reasoning Model Efficient cites this paper.

VeriThinker: Learning to Verify Makes Reasoning Model Efficient rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:42:11.309063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:42:11.309063Z digest=sha256:fdfbdbe83b8fd33c17221bc5720a1da881bbcdbace1b510258bdf53d7f94e5c0

Observation 6ac5c1b8-bbad-4caf-b8b5-86409613798d · inbound

RaDeR: Reasoning-aware Dense Retrieval Models cites this paper.

RaDeR: Reasoning-aware Dense Retrieval Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:35:19.843349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:35:19.843349Z digest=sha256:c64c30f350d3793c9cfed70017cb654fe1fbcfbf3060d255d64394d2512aa782

Observation 55471775-4c12-4441-a108-254b5ab1853e · inbound

MMATH: A Multilingual Benchmark for Mathematical Reasoning cites this paper.

MMATH: A Multilingual Benchmark for Mathematical Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:23:07.990118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:23:07.990118Z digest=sha256:340eaebe8a8aea06a5c0d1d3f3acb7bcc12575a458c03a59ca9a58c112927487

Observation bf166085-f870-4f4d-bd71-c6c1dacb4a81 · inbound

Faster and Better LLMs via Latency-Aware Test-Time Scaling cites this paper.

Faster and Better LLMs via Latency-Aware Test-Time Scaling rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:14:21.569416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:14:21.569416Z digest=sha256:0489b8272834429dfdaadc63c711f8fe2df7cb10777be1ad02f83bc89c8f5261

Observation 3e200c4b-2622-4226-bfdb-4d622dc54603 · inbound

Concise Reasoning, Big Gains: Pruning Long Reasoning Trace with Difficulty-Aware Prompting cites this paper.

Concise Reasoning, Big Gains: Pruning Long Reasoning Trace with Difficulty-Aware Prompting rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:20.744113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:20.744113Z digest=sha256:8b26bad9c43bb435f8c3b54629f4e84d6166f9f3be65e9747fef9114086efea2

Observation 1216ef42-4154-4d8a-b68e-e40f0f23babc · inbound

Can Past Experience Accelerate LLM Reasoning? cites this paper.

Can Past Experience Accelerate LLM Reasoning? rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:56.145963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:56.145963Z digest=sha256:0d8060ed471e0e61d9068468e789add3a0f0149dc70b01ddeaca01fab1e31f54

Observation 6db72fc4-c218-4d06-a80d-91872720b3d0 · inbound

rStar-Coder: Scaling Competitive Code Reasoning with a Large-Scale Verified Dataset cites this paper.

rStar-Coder: Scaling Competitive Code Reasoning with a Large-Scale Verified Dataset rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:40:16.872558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:40:16.872558Z digest=sha256:79f33b9e24646d4e89d5be337e1858e66424a27cb686dd350767603fb52c5763

Observation f779bc17-e800-4b2d-bdc9-34e5539ba275 · inbound

UI-Genie: A Self-Improving Approach for Iteratively Boosting MLLM-based Mobile GUI Agents cites this paper.

UI-Genie: A Self-Improving Approach for Iteratively Boosting MLLM-based Mobile GUI Agents rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:34:12.734899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:34:12.734899Z digest=sha256:a055e120e270f4198a666f50aa0b2e78d9b233c2f8ad0d8b0fa485895ad5dbf9

Observation 2d721ffe-0a6c-461f-9e63-ff41c627315e · inbound

RAG-Zeval: Towards Robust and Interpretable Evaluation on RAG Responses through End-to-End Rule-Guided Reasoning cites this paper.

RAG-Zeval: Towards Robust and Interpretable Evaluation on RAG Responses through End-to-End Rule-Guided Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:14:29.991659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:14:29.991659Z digest=sha256:1c1559e74ed930faf7634b31c09c5aded4cd40d705680ad931f7e4b470b0d188

Observation 5f7f1c3c-cd7a-45be-ad00-ebd75ad3a8b4 · inbound

How Much Backtracking is Enough? Exploring the Interplay of SFT and RL in Enhancing LLM Reasoning cites this paper.

How Much Backtracking is Enough? Exploring the Interplay of SFT and RL in Enhancing LLM Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:32:21.738872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:32:21.738872Z digest=sha256:28b27583eeb73a0993bea9c61e2d4e322e73ded11b0d620cfa19f7c34feb597f

Observation 317aff50-0850-44ce-bba9-846ccbf27ed4 · inbound

Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning cites this paper.

Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-12T12:12:08.724844Z digest=sha256:f2892b815fedacb3fdfa8a6c316789c9f434c9c45fe214f07211e89d4e220d76

Observation abc0b61a-31c9-4586-8bf7-70f503b7e8fe · inbound

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL cites this paper.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.139932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.139932Z digest=sha256:577376f562cf7a415bd5317c0bbe88d87b08f962ace77c24e9d99c0930b74ffa

Observation 94f56415-7df1-433c-b465-91a8350e4ce4 · inbound

AdaDecode: Accelerating LLM Decoding with Adaptive Layer Parallelism cites this paper.

AdaDecode: Accelerating LLM Decoding with Adaptive Layer Parallelism rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:02:39.226807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:02:39.226807Z digest=sha256:b3df986cbf6c65bafd8f984dfa7621d9e51d263b16c1085fffe15061e02d4bea

Observation 82381dd6-b3cb-4230-a172-84bd0c8bfe11 · inbound

Structured Pruning for Diverse Best-of-N Reasoning Optimization cites this paper.

Structured Pruning for Diverse Best-of-N Reasoning Optimization rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:56:23.571793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:56:23.571793Z digest=sha256:9becb489e9bf0904134e62f7db2b8153994d544e78c97de975f99f100185ea87

Observation 8e199341-7a67-43c8-a4fa-e3400918a3eb · inbound

DynamicMind: A Tri-Mode Thinking System for Large Language Models cites this paper.

DynamicMind: A Tri-Mode Thinking System for Large Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T10:21:27.136837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:21:27.136837Z digest=sha256:3d14a7e17bca86617f16f75fdd4f7ec0c72c6695b19cc00c90277f26b512d3b7

Observation ffc07461-ffdc-44a3-8845-da654ca7584f · inbound

Unlocking Recursive Thinking of LLMs: Alignment via Refinement cites this paper.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.005898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.005898Z digest=sha256:d77984ed9d26f65fdc173482841bc72d6b937fc2b56b641210398876201f4437

Observation bfdd0ab2-dde6-46e3-a06c-933d250e130b · inbound

A Survey on Large Language Models for Mathematical Reasoning cites this paper.

A Survey on Large Language Models for Mathematical Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:45.965584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:14:45.965584Z digest=sha256:878c8638312c61154dd52a2467f3a2a13711fde69adcf46395b0ee98f59da5bf

Observation 37401fbc-e87f-498d-9c1b-9519b094ab57 · inbound

Socratic-MCTS: Test-Time Visual Reasoning by Asking the Right Questions cites this paper.

Socratic-MCTS: Test-Time Visual Reasoning by Asking the Right Questions rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:18.720845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:01:18.720845Z digest=sha256:f654a72045412cf1c0e221130df5cc8b25b099876fb8492ca1d42c8cd9472999

Observation 06e02db3-49d4-4303-b2bb-2dae77510a5a · inbound

SwS: Self-aware Weakness-driven Problem Synthesis in Reinforcement Learning for LLM Reasoning cites this paper.

SwS: Self-aware Weakness-driven Problem Synthesis in Reinforcement Learning for LLM Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:24.112886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:24.112886Z digest=sha256:ef1241b8786bc25c83e73fec02bb428b8c673f802226b7e1e7d3eb0cbf0aabd2

Observation 301449ee-3612-4f6f-a783-45d2679d11f0 · inbound

TTT-Bench: A Benchmark for Evaluating Reasoning Ability with Simple and Novel Tic-Tac-Toe-style Games cites this paper.

TTT-Bench: A Benchmark for Evaluating Reasoning Ability with Simple and Novel Tic-Tac-Toe-style Games rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T04:38:41.146393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:38:41.146393Z digest=sha256:63a6c295af182b0a33599a342f3f60a75c00c45b528aa3be2b732a386bc5f2ba

Observation 59d93812-8bd4-4b1c-8421-7c8c5550685c · inbound

MCTS-Refined CoT: High-Quality Fine-Tuning Data for LLM-Based Repository Issue Resolution cites this paper.

MCTS-Refined CoT: High-Quality Fine-Tuning Data for LLM-Based Repository Issue Resolution rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:49:34.039982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:49:34.039982Z digest=sha256:4927d6430769e21599fcb8b9331f37cb085cf1de368c9fce3b5c7ed7ccce1b83

Observation 05224726-4f85-4934-aa07-8d6c13110327 · inbound

Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team cites this paper.

Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:24:14.735972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:24:14.735972Z digest=sha256:c620f7b2db0d230d43a908bdf855f047739a7dbda5cbb86d49e43dec49c14106

Observation 2656d7cb-e0c5-402f-bb80-5de2452c58f7 · inbound

DuaShepherd: Integrating Stepwise Correctness and Potential Rewards for Mathematical Reasoning cites this paper.

DuaShepherd: Integrating Stepwise Correctness and Potential Rewards for Mathematical Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:35.759763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:35:35.759763Z digest=sha256:29838ea8044c171777c97a088f54e939f5548346bc22aa6da0309465bd8881b9

Observation e85fecc2-2d41-4bdd-ad63-eef3dcabfd24 · inbound

Distilling Tool Knowledge into Language Models via Back-Translated Traces cites this paper.

Distilling Tool Knowledge into Language Models via Back-Translated Traces rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:11:58.354030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:11:58.354030Z digest=sha256:a686e4fe1f3c4cca59fa2836ecc55e18eace9162a596a120420f52daa0426d89

Observation d26efe52-e749-4d62-9209-0bb1d6082afe · inbound

Towards Understanding the Cognitive Habits of Large Reasoning Models cites this paper.

Towards Understanding the Cognitive Habits of Large Reasoning Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:07:56.268149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:07:56.268149Z digest=sha256:e858c2925d2dba07346aaae7fb0cce10316c1728e2fbebdd3b41fce27e23377d

Observation 8226f95e-f046-4cbf-919f-72c19f342603 · inbound

Test-Time Scaling with Reflective Generative Model cites this paper.

Test-Time Scaling with Reflective Generative Model rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:46:44.759235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:46:44.759235Z digest=sha256:4363ea15e63a5963d8397d90e8a5eb05acf11746de00f1eb6ed996dfa5dd0522

Observation 6c2624e4-c766-48db-adb9-7fdc2332769a · inbound

Enhancing Test-Time Scaling of Large Language Models with Hierarchical Retrieval-Augmented MCTS cites this paper.

Enhancing Test-Time Scaling of Large Language Models with Hierarchical Retrieval-Augmented MCTS rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T19:28:12.313777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:28:12.313777Z digest=sha256:417d354dba07924f9d708615dee1f14bd2f3bc37cc1e6d14adaea74cd072a771

Observation 3113f3ab-8d7c-4976-bfb2-2ace325cef0b · inbound

From Language to Logic: A Bi-Level Framework for Structured Reasoning cites this paper.

From Language to Logic: A Bi-Level Framework for Structured Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 906

Resolution
unresolved
no resolver link, observed 2026-08-06T18:24:50.462671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:24:50.462671Z digest=sha256:36f4768e1b6ef7580a10fcdbaec4485d0bb758d5d59b11862e6c0d8274a636cf

Observation 474e7e1f-ebec-4769-90c1-086e16e92013 · inbound

EduFlow: Advancing MLLMs' Problem-Solving Proficiency through Multi-Stage, Multi-Perspective Critique cites this paper.

EduFlow: Advancing MLLMs' Problem-Solving Proficiency through Multi-Stage, Multi-Perspective Critique rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:04:53.340378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:04:53.340378Z digest=sha256:04c5956c58a5c19a6bff90459f96e29c3dfff955b5ddb83d329051be1fdb6d58

Observation 6dfdf8e0-b155-410e-8274-540a77f81581 · inbound

AI-Powered Math Tutoring: Platform for Personalized and Adaptive Education cites this paper.

AI-Powered Math Tutoring: Platform for Personalized and Adaptive Education rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:29:26.331533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:29:26.331533Z digest=sha256:b95965bc073764b8a636c1c4f6dc56ecd8f68cd1fd112c2d953332e832f340e7

Observation d753e3c7-eb92-43de-a765-31c634f38b70 · inbound

Dynamic and Generalizable Process Reward Modeling cites this paper.

Dynamic and Generalizable Process Reward Modeling rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T14:44:35.823791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:44:35.823791Z digest=sha256:064a29ec5640fa96945ded842a95e5d206be98934dc0307e0c58ae9f9e944355

Observation 6dfeb8e6-c345-49ff-9bc1-20b33a167609 · inbound

AQuilt: Weaving Logic and Self-Inspection into Low-Cost, High-Relevance Data Synthesis for Specialist LLMs cites this paper.

AQuilt: Weaving Logic and Self-Inspection into Low-Cost, High-Relevance Data Synthesis for Specialist LLMs rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T14:37:28.746721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:37:28.746721Z digest=sha256:ffe2493a5173ca9f107e5bc38bcda9f96da2e0dc438d2560f3c4588da8f278c8

Observation fdb1f98b-6f30-4fc9-aa4c-de8503fd383a · inbound

A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence cites this paper.

A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 234

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-14T22:23:14.621091Z digest=sha256:748b5e62e74c094b99d965eda641ded79e758efed8f2b84e80d3276a7e14cf27

Observation 9f30538b-8588-4cba-add5-7cc3256747de · inbound

StructVRM: Aligning Multimodal Reasoning with Structured and Verifiable Reward Models cites this paper.

StructVRM: Aligning Multimodal Reasoning with Structured and Verifiable Reward Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T23:29:17.082217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:29:17.082217Z digest=sha256:e348ff3cd9632f26c877bd2891c6ac3df4b0e8e6778497b6b8f0a73fa5fba25a

Observation d9fac228-abbe-4aac-8de2-91b4c37dae0f · inbound

An Explainable Machine Learning Framework for Railway Predictive Maintenance using Data Streams from the Metro Operator of Portugal cites this paper.

An Explainable Machine Learning Framework for Railway Predictive Maintenance using Data Streams from the Metro Operator of Portugal rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T23:25:50.433469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:25:50.433469Z digest=sha256:43149273a2bbedab8d6184bad39daa7e44e622ed4ae4c2b5f67b0e0e0b6177ec

Observation 62909ce8-c75d-4330-8c4d-8d2d6a3e3c6e · inbound

Sample-efficient LLM Optimization with Reset Replay cites this paper.

Sample-efficient LLM Optimization with Reset Replay rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-18T23:41:54.694793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T23:38:08.044450Z digest=sha256:9ca42ddcdd343e3531dd6879605fbd2f8c31a5888d8dae486e5f86fa4ecad800

Observation 1ab3c225-c272-492f-8113-19d1233e2401 · inbound

InteChar: A Unified Oracle Bone Character List for Ancient Chinese Language Modeling cites this paper.

InteChar: A Unified Oracle Bone Character List for Ancient Chinese Language Modeling rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T21:23:23.591867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:23:23.591867Z digest=sha256:26a6087ae16afb6b767d90670fc4a34e80c72b80ab89fc839b28e30104206541

Observation 1e6a9ace-d75a-451e-b5c1-cc2d3f741e21 · inbound

rStar2-Agent: Agentic Reasoning Technical Report cites this paper.

rStar2-Agent: Agentic Reasoning Technical Report rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T14:57:39.405467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:57:39.405467Z digest=sha256:8a9229aceb008958c4566a047aabd3e943075fe9393b7bf76e60764a57dc9d32

Observation e887d306-6f75-4c59-9230-544deff51a20 · inbound

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey cites this paper.

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 202

Resolution
verified exact
local_arxiv, observed 2026-05-18T19:21:48.008345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T19:19:36.427337Z digest=sha256:2ced7b89b326feddb9e199482b1666e0e5c06aed5d6b3f9efe160d986e5a4444

Observation 255eb9be-145d-4f37-a01e-831da3972f87 · inbound

ParaThinker: Native Parallel Thinking as a New Paradigm to Scale LLM Test-time Compute cites this paper.

ParaThinker: Native Parallel Thinking as a New Paradigm to Scale LLM Test-time Compute rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T13:52:04.430116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:52:04.430116Z digest=sha256:36940940705e5aad9c7fa572372d62a47f9d45bf4b5c5a92bca586611cadf9f2

Observation c04d4468-5067-4343-8006-c47994ef6c65 · inbound

DeepSearch: Overcome the Bottleneck of Reinforcement Learning with Verifiable Rewards via Monte Carlo Tree Search cites this paper.

DeepSearch: Overcome the Bottleneck of Reinforcement Learning with Verifiable Rewards via Monte Carlo Tree Search rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-18T12:12:35.851659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T12:12:25.437344Z digest=sha256:b8abf16646e0a2e1b916d5a9bddcb4dc7c62db031b29775769115dcb7ae36c48

Observation 2b12d6ce-7fe1-40c5-baf4-25f423b5945b · inbound

Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs cites this paper.

Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-18T05:30:55.039488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T05:30:11.389756Z digest=sha256:2319f2a7560b82a085dc49f1e8b10b3d781cc416681c0701d8e78bae7d6b2a82

Observation 0c1c197b-2910-421a-b4f6-af84b593fac8 · inbound

MASPRM: Multi-Agent System Process Reward Model cites this paper.

MASPRM: Multi-Agent System Process Reward Model rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T07:56:46.376327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:56:46.376327Z digest=sha256:8a30cae1644ab7d302547d5825fe99808f41ccc9716644e7ea13d215086868f6

Observation a7b835e5-3650-4486-af52-465b278ec30d · inbound

Sharpness-Guided Group Relative Policy Optimization via Probability Shaping cites this paper.

Sharpness-Guided Group Relative Policy Optimization via Probability Shaping rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T03:12:22.378053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T03:10:57.839146Z digest=sha256:eb8fbffc7acd4090cbfd7db6a7d09622bfd9a37e8426fe66024c49405be153de

Observation 4e046aa7-c60f-4465-9b24-55412e400030 · inbound

TreeCoder: Systematic Exploration and Optimisation of Decoding and Constraints for LLM Code Generation cites this paper.

TreeCoder: Systematic Exploration and Optimisation of Decoding and Constraints for LLM Code Generation rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-17T04:19:18.805697Z digest=sha256:1deb3cad709ed48b1ed55afd0f428820a01bc1c0e78cd5d5f87558ea8bccd76e

Observation 152560ad-ddcc-4670-a512-74c44d9220e2 · inbound

PyFi: Toward Pyramid-like Financial Image Understanding for VLMs via Adversarial Agents cites this paper.

PyFi: Toward Pyramid-like Financial Image Understanding for VLMs via Adversarial Agents rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-16T23:38:28.197932Z digest=sha256:f5bc51ec32856deef41bbee825da9b973ce9d47270080e510adaeb2a5b2afe40

Observation 4ca6de60-51a9-4c36-ab8b-9c0b293768c0 · inbound

On the Overscaling Curse of Parallel Thinking: System Efficacy Contradicts Sample Efficiency cites this paper.

On the Overscaling Curse of Parallel Thinking: System Efficacy Contradicts Sample Efficiency rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-16T10:38:09.875786Z digest=sha256:63ec3ca9078dc0299cbfc6ab62cdf2c9b1efad73707c8af0f430c57a69616444

Observation 2a65ce3f-459f-4877-9f61-32c6531bc8d9 · inbound

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula cites this paper.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:34.738415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:34.738415Z digest=sha256:9eb1f6b20e29784d611f8dad98a46c8b81b5f0ddecb58a4c0a26b8abaeea5d41

Observation 45829dcc-6a81-4821-b00a-25494ece48f0 · inbound

Can I Have Your Order? Monte-Carlo Tree Search for Slot Filling Ordering in Diffusion Language Models cites this paper.

Can I Have Your Order? Monte-Carlo Tree Search for Slot Filling Ordering in Diffusion Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T23:50:34.097649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T23:50:34.097649Z digest=sha256:48e5c89606f3f735a0076a727c2f722c145a09c57f0a0307db1e7b681152f9d3

Observation d82cbb9a-454d-484b-972c-986e1b46981a · inbound

SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning cites this paper.

SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T15:59:04.802241Z digest=sha256:9242e58de109698edb77d877b238a91ac8c8e448674b5c295295da6291e5cfb7

Observation bda38c06-9ec6-46ea-8562-c5027c81e588 · inbound

SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning cites this paper.

SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T22:47:33.917420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T22:47:33.917420Z digest=sha256:de0c489bb681632c91dfe67c7fe62a68fda49cca3c56907e1c0c954d4c9011f5

Observation 76a04c12-74a7-4501-a022-b6575dd8aab0 · inbound

CoTEvol: Self-Evolving Chain-of-Thoughts for Data Synthesis in Mathematical Reasoning cites this paper.

CoTEvol: Self-Evolving Chain-of-Thoughts for Data Synthesis in Mathematical Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T10:58:27.736448Z digest=sha256:04a75d79262a50348111f881951ec3fff7007a180799719d6e5146443189e91e

Observation bf86c04b-90ac-4460-a3bf-f6cb036765e5 · inbound

Fine-Tuning Small Reasoning Models for Quantum Field Theory cites this paper.

Fine-Tuning Small Reasoning Models for Quantum Field Theory rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T03:23:18.770963Z digest=sha256:2c3461c3276e94855b3252b71d587f3149e438c441c83199f9b9125e09a53dd1

Observation 84324a20-d793-4820-9f8a-78fecad56a1f · inbound

A Limit Theory of Foundation Models: A Mathematical Approach to Understanding Emergent Intelligence and Scaling Laws cites this paper.

A Limit Theory of Foundation Models: A Mathematical Approach to Understanding Emergent Intelligence and Scaling Laws rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-13T07:27:21.118156Z digest=sha256:adaed50a0fa5b87e619fc9de536294ad42d9acdc88517bdc63f743ca822ad179

Observation 48368654-04ff-49e0-84d2-81553fe4a869 · inbound

A Limit Theory of Foundation Models: A Mathematical Approach to Understanding Emergent Intelligence and Scaling Laws cites this paper.

A Limit Theory of Foundation Models: A Mathematical Approach to Understanding Emergent Intelligence and Scaling Laws rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 76

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T09:05:36.328782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-01T09:03:59.522516Z digest=sha256:91810e648c970fc3010181a3090958f8331205c2c701dfdd1a6e445391cd2396

Observation 431bed34-8c12-4e7b-b2f4-f4f571024482 · inbound

IRIS: Interleaved Reinforcement with Incremental Staged Curriculum for Cross-Lingual Mathematical Reasoning cites this paper.

IRIS: Interleaved Reinforcement with Incremental Staged Curriculum for Cross-Lingual Mathematical Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T03:44:56.098128Z digest=sha256:18b00544fa0ae7954830af53fc8cd4e090bf03ba5e9e4018dd1c0df931cf42ca

Observation 980b8362-b7da-4dd0-943b-77bf37fa5c77 · inbound

Confidence-Aware Alignment Makes Reasoning LLMs More Reliable cites this paper.

Confidence-Aware Alignment Makes Reasoning LLMs More Reliable rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-11T02:10:40.020460Z digest=sha256:e0bd2dbaa1fd37c7b3f148f73923a7fc4b1aa92bad20e7948a73b454cbc60e17

Observation f820cb88-4dff-4174-a2d3-7ab13abeaacc · inbound

Mathematical Reasoning via Intervention-Based Time-Series Causal Discovery Using LLMs as Concept Mastery Simulators cites this paper.

Mathematical Reasoning via Intervention-Based Time-Series Causal Discovery Using LLMs as Concept Mastery Simulators rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-11T02:05:48.483640Z digest=sha256:ac11ab5a78a52062248c3f945d9512e65babd9c42bcb05b0c9b7fd1c741ded4a

Observation 68428fb5-62ff-4d18-b2d9-6d86d8def2e4 · inbound

PriorZero: Bridging Language Priors and World Models for Decision Making cites this paper.

PriorZero: Bridging Language Priors and World Models for Decision Making rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-13T05:25:05.907123Z digest=sha256:af2c0fcee41d4d2ae1c1c15bd23cf2660d2fa8abb9c672db1d683dbbcb762516

Observation 90b203d4-4837-4398-ba2a-202f27da154a · inbound

Many-Shot CoT-ICL: Making In-Context Learning Truly Learn cites this paper.

Many-Shot CoT-ICL: Making In-Context Learning Truly Learn rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T04:43:48.343981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-14T19:15:44.379686Z digest=sha256:01e18279c1def8b6112c5c0c18eca5ddd0ca813ec00afe69011406f95d4ff298

Observation a8d5e88f-8379-4512-ab09-629fe4024d65 · inbound

Graphs of Research: Citation Evolution Graphs as Supervision for Research Idea Generation cites this paper.

Graphs of Research: Citation Evolution Graphs as Supervision for Research Idea Generation rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T20:55:04.014261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-30T20:52:32.567107Z digest=sha256:29a14214c6f36f8b666c863f32822ca74095704d4d8f7c27a6a38d0611230af8

Observation 335db1e3-ef65-49d7-a9e9-c6de0432b6f6 · inbound

PopuLoRA: Co-Evolving LLM Populations for Reasoning Self-Play cites this paper.

PopuLoRA: Co-Evolving LLM Populations for Reasoning Self-Play rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 64

Resolution
metadata mismatch
local_arxiv, observed 2026-05-19T21:42:48.129101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-19T21:37:56.570173Z digest=sha256:b7445528792d69ba651e31056c6a61f333ee0171c2a2201adeb2a60eb2b192a8

Observation d890c629-6899-4748-9604-6f67267dcb76 · inbound

STRIDE: Learnable Stepwise Language Feedback for LLM Reasoning cites this paper.

STRIDE: Learnable Stepwise Language Feedback for LLM Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-05-20T20:49:00.759777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T20:47:16.236629Z digest=sha256:90590e88b9dbce4a8e65b033b3b1f5bbb7b7ce9bb562bd345bdcd7b2ffe7b922

Observation 89f1e9a3-e938-492c-bd54-04a02c7bbea4 · inbound

Mathematical Reasoning in Large Language Models: Benchmarks, Architectures, Evaluation, and Open Challenges cites this paper.

Mathematical Reasoning in Large Language Models: Benchmarks, Architectures, Evaluation, and Open Challenges rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-12T16:29:56.962393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T16:29:56.962393Z digest=sha256:302686c72ef521edb6bc3b805e56648ae638d69f71ace0069e11e8506cf662d7

Observation cdd2fe36-2bab-4888-9b2f-9ea87f2b85d7 · inbound

When Self-Belief Misleads: Active Label Acquisition for Reinforcement Learning with Verifiable Rewards cites this paper.

When Self-Belief Misleads: Active Label Acquisition for Reinforcement Learning with Verifiable Rewards rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-06-29T23:04:01.277826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T22:58:46.313028Z digest=sha256:473ce4958af2298dc62d071ad46eb03f709c83febbaf7df53d2c89a328ed220f

Observation 0eaae3cb-75a9-4dd7-8adf-e5e48f807740 · inbound

Learning to Adapt SFT Data for Better Reasoning Generalization cites this paper.

Learning to Adapt SFT Data for Better Reasoning Generalization rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T18:23:51.093488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T18:14:22.896656Z digest=sha256:6e19f107e26266ffc611f448aeed823be294421b46ccce3f67ca849bae566a80

Observation 1ea2f4e8-f7a8-4aec-aab2-c024f249b556 · inbound

Efficient Test-time Inference for Generative Planning Models with OCL Search cites this paper.

Efficient Test-time Inference for Generative Planning Models with OCL Search rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T19:52:35.142296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T18:53:08.784890Z digest=sha256:0b392da81b450db7a0409738a1914814edab4acbebb7d5e5d9296c8162152da0

Observation 59d40125-828c-4873-971f-b962ae58cf41 · inbound

From Answers to States: Verifiable Process-Level Evaluation of Chemical Reasoning in Large Language Models cites this paper.

From Answers to States: Verifiable Process-Level Evaluation of Chemical Reasoning in Large Language Models rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:46:32.853097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T09:42:02.298729Z digest=sha256:7a1410a7e622f5d0a0292de2a7c9aa09e8cd52082d817b085584852e59449305

Observation 87c65115-8ac5-42d5-b79b-9fea48415870 · inbound

Step-by-Step Optimization-like Reasoning in LLMs over Expanding Search Spaces cites this paper.

Step-by-Step Optimization-like Reasoning in LLMs over Expanding Search Spaces rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-07-02T08:46:48.975501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-28T05:46:26.938277Z digest=sha256:5ca385df5ceb2c8a91e29cac3c80e50c9272021289f8ba52ea344b2db4bbde78

Observation 9e8d504b-0622-434f-a03e-94d45c695e4a · inbound

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces cites this paper.

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 48

Resolution
metadata mismatch
local_arxiv, observed 2026-06-27T22:31:21.447916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-27T22:22:52.690010Z digest=sha256:65cdcac08518401016961b802b414d61d2f858680014a66935d97b3c16ea5e9d

Observation 90d15e96-d409-4d4a-9b50-c4a9c5f5f5ca · inbound

Improving Multimodal Reasoning via Worst Dimension Optimization cites this paper.

Improving Multimodal Reasoning via Worst Dimension Optimization rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-02T17:37:14.796624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-27T21:55:57.691185Z digest=sha256:ecfff5b6ab96211606d7f0c432446cecc18ebaa1e4de30ad1d5ee5313eccd43e

Observation d65ea05a-f974-467a-a5e5-e3cd492beba8 · inbound

CATPO: Critique-Augmented Tree Policy Optimization cites this paper.

CATPO: Critique-Augmented Tree Policy Optimization rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-06-27T19:31:09.966790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:99ddb597f7cf44020cee4bcdabf560c81064c8295c109bf34e8fe222b5e40e4a

Observation ed7f0660-4b5f-487c-bb4d-736115358585 · inbound

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes cites this paper.

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 82

Resolution
verified exact
local_arxiv, observed 2026-06-27T13:00:56.139993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-27T12:59:51.091008Z digest=sha256:52ea836636ee65809d5010589621bfa0a49ebfad8e6e71f51dc616007ddd0df1