Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:14:31.352194Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 69 of 69 outbound references and 0 inbound Pith citation observations for arXiv:2608.09555.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:14:31.352194Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
69 of 69 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 507b7d8a-38a2-42b3-b2b6-0acf7a468d1a · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents International Conference on Learning Representations , year =
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec0e4c63-4d3c-49c9-9246-db7a4424b54a · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents 2022 , url =
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f414b360-34f2-4511-a4ab-4ad9c7342747 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents 2017 , eprint =
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f3f7cbb-3c51-4871-9536-6cd3e0f3067a · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents International Conference on Learning Representations , year =
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3efc714c-d80d-41c1-91f0-89e9c2ece75e · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents 2025 , url =
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b4e1fdf8-05e5-4121-bce1-fb5acde55492 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents 2025 , eprint =
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92734b7a-db07-4d6d-9659-552b77650e4b · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents 2023 , url =
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bece669-2d40-414f-bb49-228eee50de32 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents International Conference on Learning Representations , year =
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99017870-c196-4656-8f59-c359729bf9e7 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents International Conference on Learning Representations (ICLR) , year =
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e40c3af5-00e0-48ec-bc7a-d4b0a9fe2f33 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Second Conference on Language Modeling , year =
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6dd801f0-3f9e-4817-8d7e-3e1cfa13d9cb · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents 2026 , eprint =
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 735f970c-752e-4b13-9537-801c6497fb0e · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents 2026 , eprint =
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d178ae83-6349-47c7-9064-aaad55c6bdf3 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents GEAR: Granularity-Adaptive Advantage Reweighting for LLM Agents via Self-Distillation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77e13218-f9ee-42fb-94a6-63aad1b95416 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents 2023 , url =
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ea3cc6d0-a8f2-415a-8700-49c02817f0e3 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents 2026 , month = may, eprint =
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation afb1506b-0360-48a2-b815-4fcbb778c4b2 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Proceedings of the 43rd International Conference on Machine Learning , volume =
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9625b209-5e08-4a6b-9e98-aa3ff18ffc87 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Proceedings of the 43rd International Conference on Machine Learning , volume =
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ec4464c1-269d-46b5-87f9-ef545c6507cb · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Advances in neural information processing systems , volume=
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3291273-6664-455b-8734-e00f2be9b645 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Advances in Neural Information Processing Systems , volume=
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fadadc6-3f06-4d56-bdb6-f4c7f8fafa4a · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Forty-third International Conference on Machine Learning , year=
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 94c8467a-5574-4b64-b110-113fa04877b1 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Findings of the Association for Computational Linguistics: EACL 2026 , pages=
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e3d8046d-db4a-4506-925d-baf8001abceb · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Findings of the Association for Computational Linguistics: EACL 2026 , pages=
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c2e3fe46-1ea6-4bd7-a256-7a6315a69c4d · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0fb8bede-4e9f-4bf5-80ef-f88c3eefe78c · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Unresolved cited work
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1f247c1-3cb1-4910-83e1-40bf05a79537 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Unresolved cited work
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7903429c-8806-47f5-980f-186ef96cb48c · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Reinforcement Learning for Long-Horizon Interactive LLM Agents
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4593b1e1-e865-42c5-80b1-38575037517d · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Unresolved cited work
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f45d545-e925-4af2-bfa7-27ec2a10c16d · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Unresolved cited work
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71d40a3b-ce9c-42d6-80d4-a5be85594e34 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Unresolved cited work
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6fa0560f-0a04-4218-8b47-06a2655faad2 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba50a996-3829-4f31-aa03-7ad9e382394e · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a5adf60-4924-4522-bbdc-8762f73a4f17 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents u botter, J.; L \
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 332ac8d8-a995-4f21-9355-a4f861712fe7 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents SoK: Agentic Skills -- Beyond Tool Use in LLM Agents
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2aab8971-1b3e-4fcf-8d37-b6971394b381 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents EDGE-OPD: Internalizing Privileged Context with Evidence Guided On-Policy Distillation
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ebec9a3e-062a-4e55-a20f-6ed588e00157 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Unresolved cited work
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7609e894-bf27-43ad-9b10-a3b69843474f · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents P.; Li, L.; and Li, Y
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 503145b6-beff-4139-ba22-c6a09da9dca3 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a4ff8dd-f86f-4355-890f-0676403ec2fb · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents SkillHone: A Harness for Continual Agent Skill Evolution Through Persistent Decision History
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2730c46-5290-4f27-89a8-d2e32580b505 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents HERO: Hindsight-Enhanced Reflection from Environment Observations for Agentic Self-Distillation
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0c112e7-467d-4bfd-ba97-f50e376d4d6d · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents How Well Do Agentic Skills Work in the Wild: Benchmarking LLM Skill Usage in Realistic Settings
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bc7b0b5-192d-4e4c-bbdf-54a7ffb1cbc1 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Self-Distilled Agentic Reinforcement Learning
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5921ee81-82c8-4cdc-b881-177b858037a4 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acaa1adb-a251-4f1e-9c7a-a594570efc28 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents SkillClaw: Let Skills Evolve Collectively with Agentic Evolver
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fd1d420-36c9-458a-b1cc-3269af44cfce · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc3f5913-16a2-4ea7-bcb1-91dc3ab8e900 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Unresolved cited work
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9135da0c-edf4-4732-9c47-80ad2c89eb9a · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents RLCSD: Reinforcement Learning with Contrastive On-Policy Self-Distillation
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1943922d-9eaf-45a5-892e-17d2320a0630 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents CRISP: Compressed Reasoning via Iterative Self-Policy Distillation
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af4d37ac-b3d8-42e5-b92e-e11cc2b8acb3 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f55933d3-ef4f-4477-af85-e6428f2678e7 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Anti-Self-Distillation for Reasoning RL via Pointwise Mutual Information
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79d3ff01-bf7c-4ee0-8f10-d36bb43efd3a · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5925ede0-620e-408d-85f7-5ffcc4a6cd17 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Unresolved cited work
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 03d315e8-672a-4aaf-9c97-5582db8c0e50 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Unresolved cited work
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65f63379-d7dc-45ae-a4dd-85fe6f5ba98a · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a6e1d17-bafa-4b86-896e-574f0074307b · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents TCOD: Exploring Temporal Curriculum in On-Policy Distillation for Multi-turn Autonomous Agents
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 696b4123-bccf-47e3-9afe-e2687846f7ec · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Unresolved cited work
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a6f46b7f-6602-40b4-bc82-886745447927 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Agent Skills for Large Language Models: Architecture, Acquisition, Security, and the Path Forward
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f746a172-24c2-4525-ab6c-2d609408231b · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Qwen3 Technical Report
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4a4b54e-0ca0-4b12-9818-4f47142fabfa · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Qwen2.5 Technical Report
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53025679-a086-473e-8a08-d6103d765c5c · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Self-Distilled RLVR
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5ef13e7-1098-48c1-96ef-4159d503cbcd · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents OPID: On-Policy Skill Distillation for Agentic Reinforcement Learning
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 406e0b03-90a7-46b6-93ee-5ce3060a1db3 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents OGLS-SD: On-Policy Self-Distillation with Outcome-Guided Logit Steering for LLM Reasoning
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f08f7ea4-21a0-4269-b632-0ab9f03f23e1 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Unresolved cited work
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 88209d9f-c0b7-47df-a118-3fe92ac2f159 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Unresolved cited work
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a9b215d6-3d46-4fa7-bae0-17f50f5065f8 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents SkillEvolver: Skill Learning as a Meta-Skill
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0be07b05-d407-4886-b3e9-ad59563e7207 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents OPSDL: On-Policy Self-Distillation for Long-Context Language Models
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09ad5cc2-8a49-43e4-9f4d-fdf361eb19f7 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents StepOPSD: Step-Aware Online Preference Distillation for Agent Reinforcement Learning
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a11cf51-f63a-403c-aa6c-549e0cf300b3 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents Unresolved cited work
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2782d5f9-46e7-437d-90e8-8929a3f50b8b · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents SCOPE: Signal-Calibrated On-Policy Distillation Enhancement with Dual-Path Adaptive Weighting
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 599e9b4b-7b41-4dfb-b197-fbe4ed3e0544 · outbound
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents SOD: Step-wise On-policy Distillation for Small Language Model Agents
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.