Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T06:13:56.042393Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 3 inbound Pith citation observations for arXiv:2607.13124.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T06:13:56.042393Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T04:21:09.717262Z
A source-named dated measurement, never combined with another source.
Source: cited_works
44 of 44 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 05495e38-88c2-409c-813c-cf74dbfa49ec · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation The Llama 3 Herd of Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59a18c68-ba1d-424c-ac6d-e2e404c2f5f7 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Qwen3 Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bac50476-99a8-40d7-89f2-d194284c570c · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Llm-pruner: On the structural pruning of large language models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2801b24-01d1-4d08-9a74-b0cdcfe29383 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Slicegpt: Compress large language models by deleting rows and columns
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98a75319-c8e5-4d1d-9272-a1cab40e5f0a · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Shortgpt: Layers in large language models are more redundant than you expect
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 447b6d32-80d7-4faf-b909-c9019dd736cf · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Sheared llama: Accelerating language model pre-training via structured pruning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b65ef1c2-dbac-4242-8ede-ca3324d1dac7 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Compact language models via pruning and knowledge distillation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18d4b32c-1a95-4b5d-a7ed-1b7a59dfe25d · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Sparsegpt: Massive language models can be accurately pruned in one-shot
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 092b3969-111e-4379-833b-01624dcfca21 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation A simple and effective pruning approach for large language models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61d1f87c-c33c-4436-b10a-27bb40d92fcf · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Awq: Activation-aware weight quantization for llm compression and acceler- ation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ca4910d-192a-464a-aa0b-a2d7a2b45f2e · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Fluctuation-based adaptive structured pruning for large language models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90b85378-7099-4928-9953-33e05e129f27 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Shortened LLaMA: Depth Pruning for Large Language Models with Comparison of Retraining Methods
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e203ea91-0583-4939-a257-1781b16f1d00 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation The unreasonable ineffectiveness of the deeper layers
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5951686-c711-4db2-bed7-9eaaf68a9664 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Measuring massive multitask language understanding
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc40ada5-e662-4a22-b5b9-587583dbe7ab · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Hellaswag: Can a machine really finish your sentence? InAnnual Meeting of the Association for Computational Linguistics (ACL), 2019
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a00f185f-1008-4ab3-89cd-12e219de46ee · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation The Benchmark Illusion: Pruned LLMs Can Pass Multiple Choice but Fail to Answer
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af178499-47e5-4b3d-a294-ce4a1197e2b3 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Evaluating Large Language Models Trained on Code
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2de341b-7f6f-458a-9d89-f4783912ea1c · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Training Verifiers to Solve Math Word Problems
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f92c7253-6d76-4f44-9c07-fdd2061fdd7d · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Sequence level training with recur- rent neural networks
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04bb9010-73a1-4e7d-bc22-ef70ddc337e8 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation On-policy distillation of language models: Learning from self-generated mistakes
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dabbd62a-fe92-4b96-a798-9531be65ca9e · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3160d23c-ec95-4040-a91c-803f0b925421 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5313e3eb-834f-42db-b8cf-601a754cafa3 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Sequence-level knowledge distillation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ed40101-740c-4462-900d-3ebc0969e68c · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78aa3d74-6feb-44d6-936a-ff738338db62 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation The curious case of neural text degeneration
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f00dc62-5eb6-49bb-8edf-122584093759 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Learning to break the loop: Analyzing and mitigating repetitions for neural text generation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9798946-738c-48ab-9ab9-21b4c24f6905 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Laco: Large language model pruning via layer collapse
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3eaf089d-7a60-42da-953b-1dfa8d99b9f1 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Shortv: Efficient multimodal large language models by freezing visual tokens in ineffective layers
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b60f8fe-4039-475e-8415-daccc91cbe80 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Everybody prune now: Structured pruning of llms with only forward passes.arXiv preprint arXiv:2402.05406, 2024
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c04183cb-d4b0-440b-9d57-4a61ba1b9793 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation LLM Pruning and Distillation in Practice: The Minitron Approach
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 692bf304-c130-4733-a940-4c62977bc173 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Distilling the Knowledge in a Neural Network
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 063a0a2e-7776-43fb-8c5a-d3186026af8f · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Autoregressive knowledge distillation through imitation learning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 674bed9f-6553-4c2a-bd84-87c1c20744fb · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Minillm: On-policy distillation of large language models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce47cacf-02d3-48b6-8bdb-df892ea4b61b · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation f-divergence minimization for sequence-level knowledge distillation
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cf03823-4f2a-4e9b-8d67-46f5e0252d6d · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Distillm: Towards streamlined distillation for large language models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51638354-1de4-4c03-bc45-513f77a2f90c · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Tulu 3: Pushing frontiers in open language model post-training
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 708f36d5-34e8-43a9-9dde-760c942c83f9 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Dapo: An open-source llm reinforcement learning system at scale
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7f38387-6fa5-4d06-8407-10b71816342c · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Proximal Policy Optimization Algorithms
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bb3807e-8072-4dfd-9f64-8a0cdeb1434c · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Measuring mathematical problem solving with the math dataset
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b774555d-5d2e-45de-9f48-4967f1691e10 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Opencodeinstruct.https://huggingface.co/datasets/nvidia/OpenCodeInstruct, 2024
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 230d03e4-7dc9-450f-bb56-eccd7e75b4b9 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Program Synthesis with Large Language Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d57a98e-3cb1-442d-96d1-d2c17c8ebf9d · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Vicuna: An open-source chatbot impressing gpt-4 with 90% chatgpt quality.https://lmsys.org/blog/2023-03-30-vicuna/, 2023
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b2fd1c5-bada-42fd-b5ce-f41c52d4e1e4 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation Enhancing chat language models by scaling high-quality instructional conversations
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f804196f-7868-49a1-8021-a8980a185687 · outbound
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation ).").")
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2dfe51cd-26ce-4516-8c0d-effef46be78d · inbound
IoU-PD: IoU-Aware Privileged Distillation for Visual Grounding with Multimodal Large Language Models ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation
Reference 137
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61e61b6b-01f6-4e19-8ba0-0a5ea3526c76 · inbound
On-Policy Distillation for LLM Safety: A Routing Approach to Template-Robust Realignment ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7a55579-28d9-413b-b505-3b28b724d7f8 · inbound
Adaptive FastOPD: Progress-Aware Rollout Horizon Expansion for Efficient On-Policy Distillation ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.