Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:29:35.826782Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2506.15706.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:29:35.826782Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c61dd3a6-0f87-4fba-bed9-78da6f6c6430 · outbound
MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning Qwen Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9ad4ebe-4bfa-4155-a22e-7471c5277dd7 · outbound
MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning ORPO: Monolithic Preference Optimization without Reference Model
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d4f9a7d-1972-40c9-b23c-eae567766c12 · outbound
MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning MathGenie: Generating Synthetic Data with Question Back-translation for Enhancing Mathematical Reasoning of LLMs
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 946a8099-ad50-43ce-ada0-929c11b49985 · outbound
MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning Orca-Math: Unlocking the potential of SLMs in Grade School Math
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d94a3e4-526c-4e4e-a214-63c1098c235c · outbound
MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning LLaMA: Open and Efficient Foundation Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b4612a2-3425-4484-9dae-1f571c030c6a · outbound
MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning MathPile: A Billion-Token-Scale Pretraining Corpus for Math
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6560daf-019a-4a7a-85c5-15fd3c253c84 · outbound
MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning Automatic Chain of Thought Prompting in Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a397e16-b92e-468e-8660-2dbfc3d1c1e6 · outbound
MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning Training Verifiers to Solve Math Word Problems
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ada7517d-270b-424d-bb77-6823d4c674a6 · outbound
MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning MathScale: Scaling Instruction Tuning for Mathematical Reasoning
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afa863ba-4dbf-4806-9fc1-8135714b49ef · outbound
MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning KTO: Model Alignment as Prospect Theoretic Optimization
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b40b769e-4c06-48e5-9d9d-e10fc34fc93c · outbound
MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83c80812-fb4f-4f46-8b67-046aef42f866 · outbound
MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning Measuring Mathematical Problem Solving With the MATH Dataset
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 930ec7a3-8608-4099-8aa6-f78e94c68117 · outbound
MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning Improve Mathematical Reasoning in Language Models by Automated Process Supervision
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.