{"record_type":"pith_number_record","schema_url":"https://pith.science/schemas/pith-number/v1.json","pith_number":"pith:2024:QGSTJI5JABMU5WZIVWMQOFEHBG","short_pith_number":"pith:QGSTJI5J","schema_version":"1.0","canonical_sha256":"81a534a3a900594edb28ad9907148709ad52ae974bad64003fad9bc99ced66e8","source":{"kind":"arxiv","id":"2410.01560","version":2},"attestation_state":"computed","paper":{"title":"OpenMathInstruct-2: Accelerating AI for Math with Massive Open-Source Instruction Data","license":"http://creativecommons.org/licenses/by-nc-nd/4.0/","headline":"","cross_cats":["cs.AI","cs.LG"],"primary_cat":"cs.CL","authors_text":"Alexan Ayrapetyan, Branislav Kisacanin, Igor Gitman, Ivan Moshkov, Shubham Toshniwal, Wei Du","submitted_at":"2024-10-02T14:00:09Z","abstract_excerpt":"Mathematical reasoning continues to be a critical challenge in large language model (LLM) development with significant interest. However, most of the cutting-edge progress in mathematical reasoning with LLMs has become \\emph{closed-source} due to lack of access to training data. This lack of data access limits researchers from understanding the impact of different choices for synthesizing and utilizing the data. With the goal of creating a high-quality finetuning (SFT) dataset for math reasoning, we conduct careful ablation experiments on data synthesis using the recently released \\texttt{Llam"},"verification_status":{"content_addressed":true,"pith_receipt":true,"author_attested":false,"weak_author_claims":0,"strong_author_claims":0,"externally_anchored":false,"storage_verified":false,"citation_signatures":0,"replication_records":0,"graph_snapshot":true,"references_resolved":false,"formal_links_present":false},"canonical_record":{"source":{"id":"2410.01560","kind":"arxiv","version":2},"metadata":{"license":"http://creativecommons.org/licenses/by-nc-nd/4.0/","primary_cat":"cs.CL","submitted_at":"2024-10-02T14:00:09Z","cross_cats_sorted":["cs.AI","cs.LG"],"title_canon_sha256":"83e64df4cda700f04ce50761af1a573f6c008aeefc69b705f8fcae511023ed63","abstract_canon_sha256":"a48429e6bd2845ecf5e108fe27d78d1102f27c1e6d88e1df8a5e27a5c5923883"},"schema_version":"1.0"},"receipt":{"kind":"pith_receipt","key_id":"pith-v1-2026-05","algorithm":"ed25519","signed_at":"2026-07-05T09:16:22.728302Z","signature_b64":"X/mGtx27PiyadtW4WOdKyl9U49dgQjEeEG8vsX5l6CfrtsyOlta945/M0gtFlSeNrgA49nJK0LkCUTQ/ONizCg==","signed_message":"canonical_sha256_bytes","builder_version":"pith-number-builder-2026-05-17-v1","receipt_version":"0.3","canonical_sha256":"81a534a3a900594edb28ad9907148709ad52ae974bad64003fad9bc99ced66e8","last_reissued_at":"2026-07-05T09:16:22.727807Z","signature_status":"signed_v1","first_computed_at":"2026-07-05T09:16:22.727807Z","public_key_fingerprint":"8d4b5ee74e4693bcd1df2446408b0d54"},"graph_snapshot":{"paper":{"title":"OpenMathInstruct-2: Accelerating AI for Math with Massive Open-Source Instruction Data","license":"http://creativecommons.org/licenses/by-nc-nd/4.0/","headline":"","cross_cats":["cs.AI","cs.LG"],"primary_cat":"cs.CL","authors_text":"Alexan Ayrapetyan, Branislav Kisacanin, Igor Gitman, Ivan Moshkov, Shubham Toshniwal, Wei Du","submitted_at":"2024-10-02T14:00:09Z","abstract_excerpt":"Mathematical reasoning continues to be a critical challenge in large language model (LLM) development with significant interest. However, most of the cutting-edge progress in mathematical reasoning with LLMs has become \\emph{closed-source} due to lack of access to training data. This lack of data access limits researchers from understanding the impact of different choices for synthesizing and utilizing the data. With the goal of creating a high-quality finetuning (SFT) dataset for math reasoning, we conduct careful ablation experiments on data synthesis using the recently released \\texttt{Llam"},"claims":{"count":0,"items":[],"snapshot_sha256":"258153158e38e3291e3d48162225fcdb2d5a3ed65a07baac614ab91432fd4f57"},"source":{"id":"2410.01560","kind":"arxiv","version":2},"verdict":{"id":null,"model_set":{},"created_at":null,"strongest_claim":"","one_line_summary":"","pipeline_version":null,"weakest_assumption":"","pith_extraction_headline":""},"integrity":{"clean":true,"summary":{"advisory":0,"critical":0,"by_detector":{},"informational":0},"endpoint":"/pith/2410.01560/integrity.json","findings":[],"available":true,"detectors_run":[],"snapshot_sha256":"c28c3603d3b5d939e8dc4c7e95fa8dfce3d595e45f758748cecf8e644a296938"},"references":{"count":0,"sample":[],"resolved_work":0,"snapshot_sha256":"258153158e38e3291e3d48162225fcdb2d5a3ed65a07baac614ab91432fd4f57","internal_anchors":0},"formal_canon":{"evidence_count":0,"snapshot_sha256":"258153158e38e3291e3d48162225fcdb2d5a3ed65a07baac614ab91432fd4f57"},"author_claims":{"count":0,"strong_count":0,"snapshot_sha256":"258153158e38e3291e3d48162225fcdb2d5a3ed65a07baac614ab91432fd4f57"},"builder_version":"pith-number-builder-2026-05-17-v1"},"aliases":[{"alias_kind":"arxiv","alias_value":"2410.01560","created_at":"2026-07-05T09:16:22.727867+00:00"},{"alias_kind":"arxiv_version","alias_value":"2410.01560v2","created_at":"2026-07-05T09:16:22.727867+00:00"},{"alias_kind":"doi","alias_value":"10.48550/arxiv.2410.01560","created_at":"2026-07-05T09:16:22.727867+00:00"},{"alias_kind":"pith_short_12","alias_value":"QGSTJI5JABMU","created_at":"2026-07-05T09:16:22.727867+00:00"},{"alias_kind":"pith_short_16","alias_value":"QGSTJI5JABMU5WZI","created_at":"2026-07-05T09:16:22.727867+00:00"},{"alias_kind":"pith_short_8","alias_value":"QGSTJI5J","created_at":"2026-07-05T09:16:22.727867+00:00"}],"events":[],"event_summary":{},"paper_claims":[],"inbound_citations":{"count":21,"internal_anchor_count":0,"sample":[{"citing_arxiv_id":"2606.06418","citing_title":"Double Preconditioning (DoPr): Optimization for Test-Time Performance, not Validation Loss","ref_index":111,"is_internal_anchor":false},{"citing_arxiv_id":"2606.03391","citing_title":"When Model Merging Breaks Routing: Training-Free Calibration for MoE","ref_index":11,"is_internal_anchor":false},{"citing_arxiv_id":"2606.28560","citing_title":"Depth-Staggered Fibonacci Spacing for Sparse Attention: Static Schedules Beat Learned Dilation and Extrapolate Where Dense Attention Fails","ref_index":12,"is_internal_anchor":false},{"citing_arxiv_id":"2605.25263","citing_title":"Mimir: Large-scale Multilingual Concept Modeling","ref_index":17,"is_internal_anchor":false},{"citing_arxiv_id":"2605.28247","citing_title":"IRDS: Interpretable RLVR Data Selection via Verifier-Coupled Sparse Autoencoder Coverage","ref_index":6,"is_internal_anchor":false},{"citing_arxiv_id":"2606.10184","citing_title":"Dropout-GRPO: Variational Stochasticity for Continuous Latent Reasoning","ref_index":30,"is_internal_anchor":false},{"citing_arxiv_id":"2605.18851","citing_title":"STRIDE: Learnable Stepwise Language Feedback for LLM Reasoning","ref_index":58,"is_internal_anchor":false},{"citing_arxiv_id":"2605.16462","citing_title":"Asking Back: Interaction-Layer Antidistillation Watermarks","ref_index":37,"is_internal_anchor":false},{"citing_arxiv_id":"2504.11456","citing_title":"DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning","ref_index":20,"is_internal_anchor":false},{"citing_arxiv_id":"2603.15956","citing_title":"ExpertGen: Scalable Sim-to-Real Expert Policy Learning from Imperfect Behavior Priors","ref_index":2,"is_internal_anchor":false},{"citing_arxiv_id":"2502.05171","citing_title":"Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach","ref_index":159,"is_internal_anchor":false},{"citing_arxiv_id":"2604.15529","citing_title":"LACE: Lattice Attention for Cross-thread Exploration","ref_index":34,"is_internal_anchor":false},{"citing_arxiv_id":"2605.03344","citing_title":"RAG over Thinking Traces Can Improve Reasoning Tasks","ref_index":34,"is_internal_anchor":false},{"citing_arxiv_id":"2502.01456","citing_title":"Process Reinforcement through Implicit Rewards","ref_index":50,"is_internal_anchor":false},{"citing_arxiv_id":"2412.10302","citing_title":"DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding","ref_index":85,"is_internal_anchor":false},{"citing_arxiv_id":"2604.12229","citing_title":"HintMR: Eliciting Stronger Mathematical Reasoning in Small Language Models","ref_index":3,"is_internal_anchor":false},{"citing_arxiv_id":"2605.07284","citing_title":"Instruction Tuning Changes How Upstream State Conditions Late Readout: A Cross-Patching Diagnostic","ref_index":24,"is_internal_anchor":false},{"citing_arxiv_id":"2604.15529","citing_title":"LACE: Lattice Attention for Cross-thread Exploration","ref_index":34,"is_internal_anchor":false},{"citing_arxiv_id":"2604.14768","citing_title":"CoTEvol: Self-Evolving Chain-of-Thoughts for Data Synthesis in Mathematical Reasoning","ref_index":3,"is_internal_anchor":false},{"citing_arxiv_id":"2604.15529","citing_title":"LACE: Lattice Attention for Cross-thread Exploration","ref_index":34,"is_internal_anchor":false},{"citing_arxiv_id":"2604.18473","citing_title":"Train Separately, Merge Together: Modular Post-Training with Mixture-of-Experts","ref_index":42,"is_internal_anchor":false}]},"formal_canon":{"evidence_count":0,"sample":[],"anchors":[]},"links":{"html":"https://pith.science/pith/QGSTJI5JABMU5WZIVWMQOFEHBG","json":"https://pith.science/pith/QGSTJI5JABMU5WZIVWMQOFEHBG.json","graph_json":"https://pith.science/api/pith-number/QGSTJI5JABMU5WZIVWMQOFEHBG/graph.json","events_json":"https://pith.science/api/pith-number/QGSTJI5JABMU5WZIVWMQOFEHBG/events.json","paper":"https://pith.science/paper/QGSTJI5J"},"agent_actions":{"view_html":"https://pith.science/pith/QGSTJI5JABMU5WZIVWMQOFEHBG","download_json":"https://pith.science/pith/QGSTJI5JABMU5WZIVWMQOFEHBG.json","view_paper":"https://pith.science/paper/QGSTJI5J","resolve_alias":"https://pith.science/api/pith-number/resolve?arxiv=2410.01560&json=true","fetch_graph":"https://pith.science/api/pith-number/QGSTJI5JABMU5WZIVWMQOFEHBG/graph.json","fetch_events":"https://pith.science/api/pith-number/QGSTJI5JABMU5WZIVWMQOFEHBG/events.json","actions":{"anchor_timestamp":"https://pith.science/pith/QGSTJI5JABMU5WZIVWMQOFEHBG/action/timestamp_anchor","attest_storage":"https://pith.science/pith/QGSTJI5JABMU5WZIVWMQOFEHBG/action/storage_attestation","attest_author":"https://pith.science/pith/QGSTJI5JABMU5WZIVWMQOFEHBG/action/author_attestation","sign_citation":"https://pith.science/pith/QGSTJI5JABMU5WZIVWMQOFEHBG/action/citation_signature","submit_replication":"https://pith.science/pith/QGSTJI5JABMU5WZIVWMQOFEHBG/action/replication_record"}},"created_at":"2026-07-05T09:16:22.727867+00:00","updated_at":"2026-07-05T09:16:22.727867+00:00"}