Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 39 inbound Pith citation observations for arXiv:2502.17387.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T10:39:31.109183Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T16:39:58.415508Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 26485edc-3072-4b24-925c-5d34a4ec5208 · inbound
DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ee83f16d-1103-4596-9b5b-2018b82c3374 · inbound
DeepDistill: Enhancing LLM Reasoning Capabilities via Large-Scale Difficulty-Graded Data Training Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33a3da34-7ee6-46b7-a0f3-3e436693dc6f · inbound
100 Days After DeepSeek-R1: A Survey on Replication Studies and More Directions for Reasoning Language Models Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfc6ffb7-03ba-4eba-9fd6-71af80d16ba8 · inbound
Exploring the Potential of Offline RL for Reasoning in LLMs: A Preliminary Study Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8caaf516-6c3a-4be0-be69-13f6632377a8 · inbound
Nature's Insight: A Novel Framework and Comprehensive Analysis of Agentic Reasoning Through the Lens of Neuroscience Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 205
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef46c4eb-d007-43dd-bcac-af4fe3fbe2f0 · inbound
Reward Reasoning Model Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b167c31-980d-4731-958d-ea1dbf8b670d · inbound
SituatedThinker: Grounding LLM Reasoning with Real-World through Situated Thinking Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 587d4f97-b00c-4cbe-a252-20115473630a · inbound
Improving Multilingual Math Reasoning for African Languages Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f97a8a18-1b56-469a-95f7-1f54877170e9 · inbound
Enigmata: Scaling Logical Reasoning in Large Language Models with Synthetic Verifiable Puzzles Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e21e5946-6e2c-4958-aa70-05157d3d6c78 · inbound
SwS: Self-aware Weakness-driven Problem Synthesis in Reinforcement Learning for LLM Reasoning Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 791f41b8-8b5a-4841-9f32-c105a5483f26 · inbound
Ring-lite: Scalable Reasoning via C3PO-Stabilized Reinforcement Learning for LLMs Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f41dd2f5-c890-404e-8a61-91943603b5a6 · inbound
MiroMind-M1: An Open-Source Advancement in Mathematical Reasoning via Context-Aware Multi-Stage Policy Optimization Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92b27c15-f29f-4181-a11c-98ad96e447ed · inbound
Seed-Prover: Deep and Broad Reasoning for Automated Theorem Proving Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00f51b38-fa8d-4a5c-8994-ee8640e98248 · inbound
A Survey of Reinforcement Learning for Large Reasoning Models Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 871fff2f-8de7-47c9-8004-2127fc9f1b20 · inbound
EngiBench: A Benchmark for Evaluating Large Language Models on Engineering Problem Solving Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d02e1bcf-dcbf-4c8c-a7af-9087641b776b · inbound
Breaking the Self-Confirming Loop: Diagnosing and Mitigating Systemic Reward Bias in Self-Rewarding RL Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e18af844-6888-4e1a-b75e-da20f4075c36 · inbound
SPHINX: A Synthetic Environment for Visual Perception and Reasoning Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f42e5f60-9a6c-4cfc-a001-1d14b14a6a21 · inbound
GradAlign: Gradient-Aligned Data Selection for LLM Reinforcement Learning Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 255142bb-7241-47ae-aaf3-15fc3d45213d · inbound
On the Role of Reasoning Patterns in the Generalization Discrepancy of Long Chain-of-Thought Supervised Fine-Tuning Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 45719082-7f4e-4f94-98cd-19f9444c2ce3 · inbound
Detecting and Suppressing Reward Hacking with Gradient Fingerprints Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 69b63e41-8a87-47a3-8f37-4ce773ee594c · inbound
MathNet: a Global Multimodal Benchmark for Mathematical Reasoning and Retrieval Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4116eac9-a3c3-40b1-8972-e9d431b9f50c · inbound
Evaluating Answer Leakage Robustness of LLM Tutors against Adversarial Student Attacks Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c422f2e8-8fcc-4627-85cc-0db327271bfd · inbound
$R^2$-dLLM: Accelerating Diffusion Large Language Models via Spatio-Temporal Redundancy Reduction Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 65af2582-dfcd-4132-a8b4-7b49bf4b24cd · inbound
Forge: Quality-Aware Reinforcement Learning for NP-Hard Optimization in LLMs Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 64990604-f0ab-4334-9870-fad57353e7d4 · inbound
Scalable Token-Level Hallucination Detection in Large Language Models Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2107cf75-4187-41e9-931b-7d103dc5108d · inbound
Stress-Testing the Reasoning Competence of LLMs With Proofs Under Minimal Formalism Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ae0853b1-1f32-48ed-a850-59f1bd637ca3 · inbound
Boosting Reinforcement Learning with Verifiable Rewards via Randomly Selected Few-Shot Guidance Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1dbdf28d-a2cf-4b39-84ba-34d5fe2097de · inbound
MONA: Muon Optimizer with Nesterov Acceleration for Scalable Language Model Training Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9c301df9-2be6-498b-a8b6-79e6b47f0dc2 · inbound
RLVR Datasets and Where to Find Them: Tracing Data Lineage for Better Training Data Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1bf67cc3-4465-4359-93aa-6d8cc8eba11f · inbound
LLMs Are Already Good Tutors: Training-Free Prompt Optimization for Pedagogical Math Tutoring Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ae0ae862-e140-4be1-996d-8f5b4db0e93b · inbound
Where Rollouts Begin: Low-Load, High-Leverage First-Token Diversification for RLVR Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7d7824bb-7a69-4480-a017-63393c37a913 · inbound
The Tutoring Effectiveness Index: Predicting LLM Math Tutor Quality from Four Conversation Signals Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4229a3b1-0471-4e0b-b1b2-0def05f79229 · inbound
EvoTrainer: Co-Evolving LLM Policies and Training Harnesses for Autonomous Agentic Reinforcement Learning Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 553c98fc-cd8b-438a-a4f5-94e298da6728 · inbound
CALIBER: Calibrating Confidence Before and After Reasoning in Language Models Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a8d52ec6-e740-45ba-8bee-e16094dcb35e · inbound
GRPO, Dr. GRPO, and DAPO Are Three Operations on One Number: The Group-Standard-Deviation Identity Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6a4d4a1d-dfc8-46ab-93ae-5bb9458538c6 · inbound
Aligning Language Models with Selective Prediction Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b88bc4ac-058b-458a-800b-2418edf48972 · inbound
On the effectiveness of reward functions in reinforcement learning for confidence calibration of large language models Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b37ddf84-76f4-437b-9745-7ec94e741c6a · inbound
Cost of Reasoning in non-English Languages: A Case Study on Japanese Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f6994d2-75e0-4909-a4a6-0a4901aee4ef · inbound
AngelSpec: Towards Real-World High Performance Inference with Speculative Decoding Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.