Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T03:23:33.135384Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2607.28076.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T03:23:33.135384Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
42 of 42 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8e8172bb-79c1-4e07-b25f-fba0ce655d3f · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc31096d-43f2-4f8e-9935-2b9f4e82ed6a · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning International Conference on Learning Representations , volume=
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff66f73c-6472-40f4-a603-001d30a933bd · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning International Conference on Learning Representations , volume=
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 702411f5-9bf0-409c-8c52-aa436248288b · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning SCOPE: Signal-Calibrated On-Policy Distillation Enhancement with Dual-Path Adaptive Weighting
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecd3a71d-e623-4109-8a4f-04109f6b960b · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning arXiv preprint arXiv:2511.10643 , year=
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6733b28-0330-44a9-a61d-f415abe92211 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning SimCT: Recovering Lost Supervision for Cross-Tokenizer On-Policy Distillation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b46b6cc4-bf9a-4c43-a2cc-4b3c902ec7fc · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Revisiting On-Policy Distillation: Empirical Failure Modes and Simple Fixes
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 253bba33-4b18-48a6-a081-dd3df9188149 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning TIP: Token Importance in On-Policy Distillation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9893034f-faf8-49cc-8406-e8ab0adc43ed · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9a2bbab-cb7c-4910-8365-1ea8ff42b21a · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Reinforcement-aware Knowledge Distillation for LLM Reasoning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d403953d-0e33-4607-bf60-1684d2c52883 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Self-Distilled RLVR
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40dd717c-e0a2-4327-93da-c29ad0f50f90 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Self-Distilled Agentic Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d5cda3f-e3be-4a2a-bfa1-8534a31fdd6f · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c96f8ff-7d79-426f-ad57-19f8d77f4368 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning OPID: On-Policy Skill Distillation for Agentic Reinforcement Learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acc7dedf-2d15-4036-b452-5ec1a670beff · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning ReAct: Synergizing Reasoning and Acting in Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18f0648c-eb94-4cab-b66d-39a6ba55cd0f · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Advances in neural information processing systems , volume=
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60f2afc6-7a34-479a-b357-7e79203781e3 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32638830-3b3c-4b9b-97ec-8cc6bf756421 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning ECHO: Prune To Act, Trace To Learn With Selective Turn Memory In Agentic RL
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3227e63-310a-4bee-9199-0d712b99a2b7 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaa7e3b4-659c-44f7-b5e3-e9842a7e5516 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb8ed1dd-6043-4d29-9610-0bd0ce4ecd10 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Advances in Neural Information Processing Systems , volume=
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e035d40-6089-45b7-8d2a-7ed4e6c53e8b · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 374434a2-dc85-452e-9588-154ec221e931 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning arXiv preprint arXiv:2510.14545 , year=
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d5895cc-3b3c-4899-abb6-c04e930f38c5 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Advances in Neural Information Processing Systems , volume=
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce1318fd-122b-439d-96c4-43d39b2117ed · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 317bd219-c488-45fe-9fcb-8719d4ea1565 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning arXiv preprint arXiv:2603.08754 , year=
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4010f159-17dc-43fa-a917-e3d22f064006 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning RLCSD: Reinforcement Learning with Contrastive On-Policy Self-Distillation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 404fb529-99d6-408e-b12d-1e5c0f10486c · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Reinforcement Learning via Self-Distillation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58e5bc50-b9d5-4850-be04-0ca605e48e6d · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning UniSD: Towards a Unified Self-Distillation Framework for Large Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73831e79-f25f-4256-819f-4233d5cc5bee · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning OGLS-SD: On-Policy Self-Distillation with Outcome-Guided Logit Steering for LLM Reasoning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33257439-8c7f-44d6-bb77-8d48bc9befd5 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning ALFWorld: Aligning Text and Embodied Environments for Interactive Learning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfa5e7ad-9c77-4ef0-954d-56db5f9483dd · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Advances in Neural Information Processing Systems , volume=
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b532ba7-3e9d-4f1e-a0c9-148d7785d855 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Transactions of the Association for Computational Linguistics , volume=
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16265c6e-03a3-414b-ad8e-cb4ab1e53c32 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Proceedings of the 2018 conference on empirical methods in natural language processing , pages=
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0a77114-5bf4-4e4d-9c69-e68211d21594 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58c24a99-b950-4e26-b76d-a5ead428475f · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Proceedings of the 61st annual meeting of the association for computational linguistics (volume 1: Long papers) , pages=
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72c850e6-5948-437d-a8da-a287491891e4 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Proceedings of the 28th International Conference on Computational Linguistics , pages=
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad863716-d0a9-4a4c-b76a-ca7ce78a893a · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Transactions of the Association for Computational Linguistics , volume=
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be90f18d-87b6-4ebc-b118-4b334f34c5c1 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Findings of the Association for Computational Linguistics: EMNLP 2023 , pages=
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a001b18-d8d0-4d52-a2f4-b41507844e65 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning SOD: Step-wise On-policy Distillation for Small Language Model Agents
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 691d5dc1-1590-4148-a46f-56242c5063f8 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning Text Embeddings by Weakly-Supervised Contrastive Pre-training
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74de7f36-81d0-4e79-9e6c-70ff22a79792 · outbound
Group-Reflective Self-Distillation for Agentic Reinforcement Learning 2026 , eprint=
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.