Pith. sign in

Paper Citation Record · LEDGER

Learning by Distilling Context

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 37 inbound Pith citation observations for arXiv:2209.15189.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2209.15189 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 37 of 37 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T23:15:50.848819Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

7
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6315a156-3a7e-4e8f-8087-f5ed01a82869 · inbound

Large Language Models Can Self-Improve cites this paper.

Large Language Models Can Self-Improve Learning by Distilling Context

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-18T17:00:48.263694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-18T17:00:48.167441Z digest=sha256:f4d1817201b07de1911aba6821ce6a89cc627828a260b747a8c0ad311987eb6a

Observation 8a76589d-0a1a-49f6-b3ab-3bce8130e323 · inbound

The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions cites this paper.

The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions Learning by Distilling Context

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:59:30.789754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T10:59:30.728091Z digest=sha256:e1fb4de33eedee2fce632b2db4479fd40b25b489264c885e60067a2fb0ff6866

Observation 7c63712a-c751-46c4-8876-d99ce7a1cfde · inbound

Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models cites this paper.

Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models Learning by Distilling Context

Reference 297

Resolution
verified exact
arxiv_id, observed 2026-05-17T14:43:29.847490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-17T14:43:29.496457Z digest=sha256:a9d99858aabcdb8b4b3970843d6e975dea35385e06fcb001be7880000df2a2aa

Observation 50903d1d-fcc2-4a71-87ca-4b368a71f632 · inbound

TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation cites this paper.

TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation Learning by Distilling Context

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:06:01.325654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-10T16:48:53.092395Z digest=sha256:fdaa784551f9888865bb2c20b680339d1c9b892dc38678c4535043661b37c17c

Observation 6b05a7e0-2cd5-41ea-9ae2-83ab31691c93 · inbound

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning cites this paper.

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning Learning by Distilling Context

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:30:53.811574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T18:28:58.515666Z digest=sha256:ef87f521e280ee0563c04424232604207136006499abebc6ff48ddb086e7c7b2

Observation 390f6538-ff00-4411-b80c-8a8ba392cbc8 · inbound

Tuning Qwen2.5-VL to Improve Its Web Interaction Skills cites this paper.

Tuning Qwen2.5-VL to Improve Its Web Interaction Skills Learning by Distilling Context

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:51:36.168174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:51:00.660022Z digest=sha256:68d1359f689b3ee269bb5d6d1df2e83275264aa4e0c4148e454bb73e2db000ca

Observation bc112bd6-1564-4118-b3f7-1c480e323da7 · inbound

Near-Future Policy Optimization cites this paper.

Near-Future Policy Optimization Learning by Distilling Context

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:54:48.621807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T00:51:36.580600Z digest=sha256:aec24b8364d00368673eda45b644b90eb415fefca9394ad9c71b10ef3720dd16

Observation 23e15da8-ae9f-4e89-a077-959b1c3a9c0f · inbound

Rethinking Dense Sequential Chains: Reasoning Language Models Can Extract Answers from Sparse, Order-Shuffling Chain-of-Thoughts cites this paper.

Rethinking Dense Sequential Chains: Reasoning Language Models Can Extract Answers from Sparse, Order-Shuffling Chain-of-Thoughts Learning by Distilling Context

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T03:50:57.396716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-11T02:11:19.295354Z digest=sha256:5ebf90ac1748df0ce0b031576b59881ab968eb3cb87d97bc830184573f7a31fb

Observation 0d3e8055-9993-4172-b246-d8d7ee956d9c · inbound

CoDistill-GRPO: A Co-Distillation Recipe for Efficient Group Relative Policy Optimization cites this paper.

CoDistill-GRPO: A Co-Distillation Recipe for Efficient Group Relative Policy Optimization Learning by Distilling Context

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:31:26.461256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T00:59:44.364491Z digest=sha256:bd1f88f16690db02e8f2470095104597d436fd3b65edfe36d8dc40836db2755c

Observation 2f30f395-0463-46f1-8409-6420710e6aed · inbound

Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why cites this paper.

Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why Learning by Distilling Context

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:11:27.211928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:28:58.679078Z digest=sha256:b5015f03d4f03fdf83fd1e1714299ade0de488cb0f94ae764faad2180fe02f53

Observation fe2f5de1-d1ce-4a95-88cd-496aadba5cd3 · inbound

From Generic Correlation to Input-Specific Credit in On-Policy Self Distillation cites this paper.

From Generic Correlation to Input-Specific Credit in On-Policy Self Distillation Learning by Distilling Context

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:27:02.004969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:24:34.093541Z digest=sha256:a6ea6e408f90cc70ad6473a2972fddb42cc924f215c6b95dccc602258e759c11

Observation 8d283293-324b-431b-a553-37de5b1c0d88 · inbound

VSPO: Vector-Steered Policy Optimization for Behavioral Control cites this paper.

VSPO: Vector-Steered Policy Optimization for Behavioral Control Learning by Distilling Context

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:43:44.111000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T19:39:55.294398Z digest=sha256:d9038169b25e222c76e03ebbc1d8e8cf68dd2d08502820e413a3b9d8a524d9cd

Observation f2e52020-fd83-41d3-9150-d734f9a25543 · inbound

Self-Supervised On-Policy Distillation for Reasoning Language Models cites this paper.

Self-Supervised On-Policy Distillation for Reasoning Language Models Learning by Distilling Context

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:43:21.837694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-20T14:42:55.368104Z digest=sha256:08cf20bee50bdc450943bd15b1e01c4d96c9076e44ee1658b00a0bce0fd1f2b5

Observation c684b1c3-bc07-4859-bff8-e44a0c0ec5e9 · inbound

Context Memorization for Efficient Long Context Generation cites this paper.

Context Memorization for Efficient Long Context Generation Learning by Distilling Context

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:43:12.621794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T10:39:09.720412Z digest=sha256:5ec5831d8db63edb06bea8461adba8244d89bf6596757b7dc1c1e01aed956867

Observation aead6aed-ce7e-46bc-9f79-d599385a5c80 · inbound

It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs cites this paper.

It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs Learning by Distilling Context

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:09:51.844225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T08:05:44.358256Z digest=sha256:c4e142a85598e401154a20d970eec37182cb3aff8fab273efe09b8173850aa5c

Observation c02f2b1f-98f7-48ae-a4be-57807067c3f1 · inbound

Tailoring Teaching to Aptitude: Direction-Adaptive Self-Distillation for LLM Reasoning cites this paper.

Tailoring Teaching to Aptitude: Direction-Adaptive Self-Distillation for LLM Reasoning Learning by Distilling Context

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-22T08:11:17.794411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T08:06:30.862911Z digest=sha256:ac597827f9d2b08ca62bbaa9b1672a117b2c2aad229d1ca2b64b0d8e1b47fd0f

Observation e0a87535-1268-4e94-a7ce-4f26e4217fc5 · inbound

Do Language Models Need Sleep? Offline Recurrence for Improved Online Inference cites this paper.

Do Language Models Need Sleep? Offline Recurrence for Improved Online Inference Learning by Distilling Context

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-06-29T21:43:59.454603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T21:37:51.638904Z digest=sha256:59c149824ac16ae40ea1082cc7f2379ad0e0f222999bcda09d93dd128cbb9275

Observation 9f489a76-dc34-4a8f-8804-8a86246c3e36 · inbound

ThinkSwitch: Context Distillation with LoRA and Weight Interpolation for Specific-Purpose Reasoning Tasks cites this paper.

ThinkSwitch: Context Distillation with LoRA and Weight Interpolation for Specific-Purpose Reasoning Tasks Learning by Distilling Context

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T18:02:27.323880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-28T17:52:59.979543Z digest=sha256:70d3498933348ce92b81be6a91f89dece77e53907fde10ff10709a99cf866ce6

Observation a27e4b2b-54bf-489c-97b7-9ffcc22ee17e · inbound

Rethinking Continual Experience Internalization for Self-Evolving LLM Agents cites this paper.

Rethinking Continual Experience Internalization for Self-Evolving LLM Agents Learning by Distilling Context

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:56:47.807293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-28T06:29:11.398007Z digest=sha256:704ac6f881fc3d9365957a3503ace9de606b5dc21b928bcf24ba758d5f02aea1

Observation df9f27f8-938f-4796-91dc-a98b8b99672d · inbound

Amortizing Federated Adaptation: Hypernetwork Driven LoRA for Personalized Foundation Models cites this paper.

Amortizing Federated Adaptation: Hypernetwork Driven LoRA for Personalized Foundation Models Learning by Distilling Context

Reference 76

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T13:16:58.682443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-28T01:27:04.241484Z digest=sha256:1b4b3bba740a12fd39ec40ed1981c99c46f8fccd0a299696e6df0852a169e5bd

Observation 2a5df5e2-4795-432a-a14d-fe02e61050d1 · inbound

Task-Aware Structured Memory for Dynamic Multi-modal In-Context Learning cites this paper.

Task-Aware Structured Memory for Dynamic Multi-modal In-Context Learning Learning by Distilling Context

Reference 205

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T09:17:48.415598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-27T10:28:11.440915Z digest=sha256:ca9dbb71e16c6f46135290ab0d112b497403aeca37976481d00481800b207eb4

Observation e8b6433a-3038-4a7e-a9f2-ffffa1b77373 · inbound

Doc-to-Atom: Learning to Compile and Compose Memory Atoms cites this paper.

Doc-to-Atom: Learning to Compile and Compose Memory Atoms Learning by Distilling Context

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:37:57.065542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-27T09:53:35.191873Z digest=sha256:4d85657b7d8330bb0747e4f5fff158dd6fb1933b6107cf2557efdace0a203e91

Observation 59ec0084-5f5e-4a98-b02b-9789f1e586ad · inbound

PRISMR: Overcoming Parse Collapse in Multimodal Listwise Ranking via Parameterized Representation Internalization cites this paper.

PRISMR: Overcoming Parse Collapse in Multimodal Listwise Ranking via Parameterized Representation Internalization Learning by Distilling Context

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:38:29.012031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T07:00:17.978103Z digest=sha256:b6a4ed02b536c399b6ca11f35886275920fdf156afd959d1012253485ef21280

Observation 1e7fea36-569e-4851-a290-89d780eaa74c · inbound

LiteOdyssey: A Lightweight Reasoning AI Agent for Interpretable Rare-Disease Diagnosis cites this paper.

LiteOdyssey: A Lightweight Reasoning AI Agent for Interpretable Rare-Disease Diagnosis Learning by Distilling Context

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-12T13:53:50.979339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T13:53:50.979339Z digest=sha256:6718ef00250e1581a5405cd12e5240227d43ae3f6840bbfa78cea3525e4d250e

Observation 5dd966a1-a03d-4b16-8ed1-aaccc18458a8 · inbound

HMARS: A Hierarchical Multi-Agent Memory System for Long-Context Reasoning cites this paper.

HMARS: A Hierarchical Multi-Agent Memory System for Long-Context Reasoning Learning by Distilling Context

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T11:34:37.890032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-30T11:30:52.871762Z digest=sha256:8f80bbcfd975ab252e31af3bba02d56edcc69a0b8f5051c68d622c2a906ae993

Observation 6ad6097b-b6b7-458b-870b-41f3ce6d15c3 · inbound

A Single Rewrite Suffices: Empirical Lessons from Production Skill Description Optimization cites this paper.

A Single Rewrite Suffices: Empirical Lessons from Production Skill Description Optimization Learning by Distilling Context

Reference 69

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T12:05:43.607060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-07-01T02:32:19.425550Z digest=sha256:fc44bb6880fc3aaae1545a0d538f5a580b8829a214ce20c044b26c1afae42f50

Observation c08528f8-3569-4cad-b504-2a86a7fb4596 · inbound

Distill to Detect: Exposing Stealth Biases in LLMs through Cartridge Distillation cites this paper.

Distill to Detect: Exposing Stealth Biases in LLMs through Cartridge Distillation Learning by Distilling Context

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:36:56.170129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-07-02T12:27:59.332432Z digest=sha256:bfe0c49ed89990379f35ac86c528107f4d95d7888118d25d59810f8226337131

Observation ccca42a2-76c5-4d5d-b4ce-89280ded27cd · inbound

Epistemic Goggles: A Pretrained Module that Induces an Epistemic Frame via Gradient Editing cites this paper.

Epistemic Goggles: A Pretrained Module that Induces an Epistemic Frame via Gradient Editing Learning by Distilling Context

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T14:28:31.104704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-07-03T14:25:39.401131Z digest=sha256:afcdc7545847571a6fc68bba052726c565513c8245d76bd95302dbe08dc9ffb7

Observation 705eea4c-835b-4c44-8dac-9d3fa8fca4d3 · inbound

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents cites this paper.

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents Learning by Distilling Context

Reference 107

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:36:44.163623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-10T01:26:59.421158Z digest=sha256:795d3bf8d387b7ad8a6e443fbb25b246085d0661c2e2ba604533188c7f1b7307

Observation 8f316d1a-e5dd-451f-915c-2364c8772c32 · inbound

Can a Language Model Learn Facts Continually in Its Weights? cites this paper.

Can a Language Model Learn Facts Continually in Its Weights? Learning by Distilling Context

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-14T07:36:43.499258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:36:43.499258Z digest=sha256:a646bf397ea0a6b6308278f2d365d57070897b124f87cd55af29effc23b3e9b5

Observation 7f4310bc-91e5-42bf-8c39-a8b170fc4602 · inbound

Sample-Efficient Learning from Agent Experience cites this paper.

Sample-Efficient Learning from Agent Experience Learning by Distilling Context

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T08:43:21.932131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:43:21.932131Z digest=sha256:690daa0f5249c4afbd41498b3bd8c627f489f970b12c73c3543342c28ee92402

Observation b9d82d3a-535f-425f-8fd3-486c7f062d7e · inbound

Do Modules Stay in Their Lane? Role Drift in Compound LLM Systems cites this paper.

Do Modules Stay in Their Lane? Role Drift in Compound LLM Systems Learning by Distilling Context

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T08:17:26.367813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:17:26.367813Z digest=sha256:317e57cda8ae3bea5be94c7d02803450399eca4cb5e09076079d6fa7b39b2804

Observation e6c8895e-7140-4b0d-9c91-a653f644cb73 · inbound

Listen, Do Not Copy: Internalizing Audio-Grounded Scaffold Context for Robust Omni-Model Speech Understanding cites this paper.

Listen, Do Not Copy: Internalizing Audio-Grounded Scaffold Context for Robust Omni-Model Speech Understanding Learning by Distilling Context

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T06:19:33.044773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:19:33.044773Z digest=sha256:42e827594dd969b6e426f498344604fba84a1433938a3e1a2303ee5ba6f32ef5

Observation 7bad19ac-b647-4b4c-ab95-a94ab1ce032d · inbound

Masked Distillation: Internalizing the Chain-of-Thought in Language Models cites this paper.

Masked Distillation: Internalizing the Chain-of-Thought in Language Models Learning by Distilling Context

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T10:49:24.977221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:49:24.977221Z digest=sha256:de04519c3b7bf1c6f6e501b3c2077f5317d62a59623a535b5961e94d27ce88ce

Observation f34e8e4d-f88c-43da-a22c-1dd43c616bee · inbound

Flux-OPD: On-Policy Distillation with Evolving Contexts cites this paper.

Flux-OPD: On-Policy Distillation with Evolving Contexts Learning by Distilling Context

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-31T20:13:02.033825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T20:13:02.033825Z digest=sha256:968167e7af7a8c9f5f2c82a4187a160d20334c518ff2e13197350094c58c5d75

Observation d702b95c-08f7-4a5f-ae0c-2adbb5ce5d7b · inbound

Learning What to Remember: Test-Time Training via Context Distillation cites this paper.

Learning What to Remember: Test-Time Training via Context Distillation Learning by Distilling Context

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T23:14:10.159095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:14:10.159095Z digest=sha256:60d4d398b0c18fe63b0bdb0af442d2fd0c023dc2f201f020bd5b1861b8904397

Observation ff8d01c5-2bd5-42d1-9e20-50c0d9be4d9a · inbound

Agentic Reinforcement Learning with Self-Distilled Reward Shaping cites this paper.

Agentic Reinforcement Learning with Self-Distilled Reward Shaping Learning by Distilling Context

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T23:15:50.848819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:15:50.848819Z digest=sha256:b40215d80f68a9f4a98a12b59afece3a603a867c712da9b7713571a25d916418