Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 61 inbound Pith citation observations for arXiv:2602.04942.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T18:07:21.573005Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T00:26:39.270366Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation ffe965c4-710a-4edc-ade9-001c6692847f · inbound
Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation Privileged Information Distillation for Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0ed377c2-72a4-4294-8b83-1c2f0f9d703d · inbound
Embarrassingly Simple Self-Distillation Improves Code Generation Privileged Information Distillation for Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e5dbd66-ba12-4559-bfd5-93997e57e823 · inbound
Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents Privileged Information Distillation for Language Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ec78b8c4-8b51-4007-a19b-0137648ac39d · inbound
Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation Privileged Information Distillation for Language Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5ebe4292-7ac0-45c8-937f-d7f52b6f0daf · inbound
Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation Privileged Information Distillation for Language Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 612a8927-5dae-4f46-8951-4ce08df33aaa · inbound
$\pi$-Play: Multi-Agent Self-Play via Privileged Self-Distillation without External Data Privileged Information Distillation for Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation bb77c09b-b739-4431-a4c2-b02badc0ce0b · inbound
Reward Weighted Classifier-Free Guidance as Policy Improvement in Autoregressive Models Privileged Information Distillation for Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 78479ed0-6f56-4576-a841-cf6bf83ae808 · inbound
TCOD: Exploring Temporal Curriculum in On-Policy Distillation for Multi-turn Autonomous Agents Privileged Information Distillation for Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation baaebeef-1d64-40e4-b31f-bffd6e6052f3 · inbound
MAD-OPD: Breaking the Ceiling in On-Policy Distillation via Multi-Agent Debate Privileged Information Distillation for Language Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a159b673-3049-435f-b698-ce87c363e5d2 · inbound
D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models Privileged Information Distillation for Language Models
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d388fb8a-6378-4287-a5fe-d4519f9ea38c · inbound
D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models Privileged Information Distillation for Language Models
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2316f8ab-056a-4058-9cdc-b4901d134cb5 · inbound
UniSD: Towards a Unified Self-Distillation Framework for Large Language Models Privileged Information Distillation for Language Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d2b5a07e-82c1-4f98-9349-b4b9f16a1ed1 · inbound
Signal Reshaping for GRPO in Weak-Feedback Agentic Code Repair Privileged Information Distillation for Language Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b1e518f7-5015-4e73-8957-0c918aa3f206 · inbound
SOD: Step-wise On-policy Distillation for Small Language Model Agents Privileged Information Distillation for Language Models
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 4cd43221-c830-45d1-8484-a6635327a95d · inbound
SOD: Step-wise On-policy Distillation for Small Language Model Agents Privileged Information Distillation for Language Models
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f6de619-4937-4576-bfe3-4636ff447e27 · inbound
TRACE: Distilling Where It Matters via Token-Routed Self On-Policy Alignment Privileged Information Distillation for Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e8a69d5c-9c9b-4028-b70f-ae8d3e965e64 · inbound
Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why Privileged Information Distillation for Language Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 94e94df9-0b87-46f0-b2f4-a234f4f35026 · inbound
From Generic Correlation to Input-Specific Credit in On-Policy Self Distillation Privileged Information Distillation for Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2e05cd63-f5bd-4c92-b777-47e963b74fb2 · inbound
Multi-Rollout On-Policy Distillation via Peer Successes and Failures Privileged Information Distillation for Language Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d708a2ee-5b78-40ba-8e42-5ebe73006382 · inbound
Learning with Rare Success but Rich Feedback via Reflection-Enhanced Self-Distillation Privileged Information Distillation for Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 30691877-32e6-43e2-9e93-e04e7e774518 · inbound
Learning from Language Feedback via Variational Policy Distillation Privileged Information Distillation for Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c0ff8b8a-d77a-4176-9b12-ad6241303b2b · inbound
A Brief Overview: On-Policy Self-Distillation In Large Language Models Privileged Information Distillation for Language Models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 88a5c801-7be1-43af-8153-da58b2e61155 · inbound
A Brief Overview: On-Policy Self-Distillation In Large Language Models Privileged Information Distillation for Language Models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 46a7e490-75b2-42af-83ab-17634ea31430 · inbound
Next-Acceleration-Scale Prediction for Autoregressive MRI Reconstruction Privileged Information Distillation for Language Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2d2c7633-b550-4205-baa4-9fab09c1892c · inbound
Next-Acceleration-Scale Prediction for Autoregressive MRI Reconstruction Privileged Information Distillation for Language Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e0793516-e4ec-483a-b45b-b9f87305e1cf · inbound
CEPO: RLVR Self-Distillation using Contrastive Evidence Policy Optimization Privileged Information Distillation for Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 971f7dc3-1261-4dd5-b6c8-24ac4fb9a243 · inbound
Survive or Collapse: The Asymmetric Roles of Data Gating and Reward Grounding in Self-Play RL Privileged Information Distillation for Language Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 87490850-665c-46d0-beb9-236f9311db60 · inbound
Tailoring Teaching to Aptitude: Direction-Adaptive Self-Distillation for LLM Reasoning Privileged Information Distillation for Language Models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c175ca74-6d8a-460c-925f-0d49e92dc74b · inbound
Self-Policy Distillation via Capability-Selective Subspace Projection Privileged Information Distillation for Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6874a8a3-0c4a-4256-9ad5-2d54906575b0 · inbound
StepOPSD: Step-Aware Online Preference Distillation for Agent Reinforcement Learning Privileged Information Distillation for Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation cefce76c-8e83-41a4-97f6-cd994bc81952 · inbound
Self-Distilled Policy Gradient Privileged Information Distillation for Language Models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b6790bb3-d48b-418a-adc8-bfc77713c643 · inbound
Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions Privileged Information Distillation for Language Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 65f75f0f-f3c6-47b0-8184-50193743d983 · inbound
Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions Privileged Information Distillation for Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff700a35-f8c9-424c-8087-7e019b34f4dc · inbound
Self-Distillation Policy Optimization via Visual Feedback: Bridging Code and Visual Artifacts Privileged Information Distillation for Language Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6d35a4cc-40f9-419a-8642-65d44bb1c2e1 · inbound
Beyond Absolute Imitation: Anchored Residual Guidance for Privileged On-Policy Distillation Privileged Information Distillation for Language Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c3d31d0d-f718-4252-ade1-f7950cf93721 · inbound
HERO: Hindsight-Enhanced Reflection from Environment Observations for Agentic Self-Distillation Privileged Information Distillation for Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ec01ee55-5ce1-4f1f-a21d-dfd4df889de9 · inbound
ROAD-VLA: Robust Online Adaptation via Self-Distillation for Vision-Language-Action Models Privileged Information Distillation for Language Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2cbdfe62-ef0a-4716-a818-828af8904b68 · inbound
On-Policy Self-Distillation with Sampled Demonstrations Reduces Output Diversity Privileged Information Distillation for Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a22a6ef2-263b-4dd7-943d-b8651e98514e · inbound
Regime-Aware Peer Specialization for Robust RAG under Heterogeneous Knowledge Conflicts Privileged Information Distillation for Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0e8a175d-d3b7-4f5b-98c5-8336b591c47d · inbound
DOPD: Dual On-policy Distillation Privileged Information Distillation for Language Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f0e7e388-ff58-4425-a764-6195cda5f8ce · inbound
TREK: Distill to Explore, Reinforce to Refine Privileged Information Distillation for Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 243a904d-3dc2-4e64-a5cf-e60a24035d5a · inbound
Geometric Self-Distillation for Reasoning Generalization Privileged Information Distillation for Language Models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5b3b8c12-6253-4733-8524-2de3eadc56a3 · inbound
Contrastive On-Policy Distillation Privileged Information Distillation for Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a9e510c-3637-4742-b0c4-f8c16dcd3e72 · inbound
Sample-Efficient Learning from Agent Experience Privileged Information Distillation for Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2925f4a0-efba-41b0-9bf5-0385742deb9b · inbound
The Physics of Multi-Turn Long-Horizon Planning: From Pre-training to Post-training via Single- and Multi-Teacher On-Policy Agentic Distillation Privileged Information Distillation for Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d81f8ef2-0add-49a1-8a84-d1e750300254 · inbound
Pass the Baton: Trajectory-Relayed On-Policy Distillation Privileged Information Distillation for Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1cba096-2ea0-40f9-8879-457da848da60 · inbound
Not All Tokens Deserve Equal Credit: Counterfactual Sensitivity Credit Reallocation for Long-CoT Reasoning Privileged Information Distillation for Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b30ab58f-eed9-4e45-8784-87573d40aa35 · inbound
Lightning OPD 2.0: Mitigating Style Bias in Cross-Teacher On-Policy Distillation for Large Reasoning Models Privileged Information Distillation for Language Models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f60ec44-9bb8-488f-b329-ee89d846c1d3 · inbound
Is More Privileged Information Better? From Solution Traces to Problem-Solving Structure in Self-Distilled Reasoning Privileged Information Distillation for Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0affb3db-dfba-4460-8545-87dfecc7d1ae · inbound
DAPD: Dual-Anchored Policy Distillation Privileged Information Distillation for Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b9a869b-3f76-4b13-8b96-023b04888c2d · inbound
Instruction-Conditioned Exploration for Reinforcement Learning with Self-Distillation to an Unconditioned Policy Privileged Information Distillation for Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 487ce935-df39-4d34-b567-48054be6a356 · inbound
Instruction-Conditioned Exploration for Reinforcement Learning with Self-Distillation to an Unconditioned Policy Privileged Information Distillation for Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 536e9971-7b0a-4e0d-87eb-1d46b5525292 · inbound
Self-Improving Large Language Models via Progressive Experience Evolution Privileged Information Distillation for Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17bae05d-6cd2-4496-a84d-ee2afd62df82 · inbound
Self-Improving Large Language Models via Progressive Experience Evolution Privileged Information Distillation for Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d0201ab-a43f-49dd-9762-5644e0729f83 · inbound
Learning from Consensus and Disagreement: Unsupervised On-Policy Self-Distillation with Minority-Trajectory Contrast Privileged Information Distillation for Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13844ce5-9c4b-45f4-a5a6-165b9150862c · inbound
BOUND: Brief-Guided Corrective Preference Distillation at Search-Control Boundaries Privileged Information Distillation for Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c974849f-93cb-4426-bf4e-9ba3419b14f9 · inbound
Privileged Solutions or Context-Induced Teacher Behavior? Dissecting On-Policy Self-Distillation Privileged Information Distillation for Language Models
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf81325e-f6c6-49b4-98f6-d0a18497a1a4 · inbound
SR-OPSD: Self-Referenced On-Policy Self-Distillation Privileged Information Distillation for Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d7f7437-2607-476f-be08-137d76aed4b9 · inbound
ADOPD: Reference-Privileged On-Policy Distillation for MLLM-Based Industrial Anomaly Detection Privileged Information Distillation for Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 681049ca-33c3-44f4-b910-6864eea1494a · inbound
Latent On-Policy Self-Distillation Privileged Information Distillation for Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0d6e633-0499-4ac2-b14a-8fe661d1ea9b · inbound
Teach the Magnitude, Not the Direction: Verifier-Bounded Credit Assignment for Multi-Turn Multi-step LLM Agents Privileged Information Distillation for Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.