Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T21:54:44.368038Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2608.01743.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T21:54:44.368038Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
57 of 57 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation df80e58d-a917-48ca-be6d-d8d867e8e48b · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c8b89b9c-8b80-427b-bf58-57a1510d7f77 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f20e3a33-f3ed-47aa-af3c-01ea0516b4d9 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d4b0c03d-f000-49f3-a49b-7cb98d1adb80 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning H.; Gonzalez, J.; Zhang, H.; and Stoica, I
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d9356796-39dd-404c-85a1-c7e61bc5c2d1 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 65f5e0f5-bb4a-4284-a5a1-c16010cdfcc8 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74af9aad-919c-4603-a4f7-4284782058ba · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning HybridFlow: A Flexible and Efficient RLHF Framework
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b0d8687-957a-413b-b072-9ba612e3c9c7 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 959a227d-d93c-4771-a638-169caa91aef7 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 495d56fd-0403-4cdb-a2c7-d82029c5ba61 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4c92224e-9ff6-4b52-b7dc-a0a876477ea1 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning 2024 , journal =
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09e37af4-78c6-4be7-863b-ec28a04e474f · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Proceedings of the 29th symposium on operating systems principles , pages=
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f05b61a8-05ed-4c07-954b-c723093d7e39 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Qwen3 Technical Report
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81afb144-f8cc-4f4a-be3e-bc8db4b9f015 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Fine-Tuning Language Models from Human Preferences
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f77e03be-e80f-4784-8276-969b8fc49b18 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Advances in neural information processing systems , volume=
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4614ed2-ffee-4162-a9b8-5f7d2d541dc7 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Advances in neural information processing systems , volume=
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5d7d783-687b-4adf-a1dd-938cddac8947 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2aa3844b-7052-4b9d-9a1e-1610e23bdac9 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning arXiv preprint arXiv:2509.07430 , year=
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a98d7c5b-c8f8-4eba-b5bd-07ff69b01084 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Advances in Neural Information Processing Systems , volume=
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a521534-ce38-4f85-9d5e-09031face8e8 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Implicit Hierarchical GRPO: Decoupling Tool Invocation from Execution for Tool-Integrated Mathematical Reasoning
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4fea440c-0f4f-4aa6-9af2-c869ce602c86 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning arXiv preprint arXiv:2510.03865 , year=
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 51920178-91f2-4f57-aa29-e7275570cb94 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMs
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c5c3b979-95d4-4927-86f1-31f82e92fb49 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning arXiv preprint arXiv:2510.20817 , year=
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a819ad68-2a25-48bb-ba76-c9fe1bf79dfd · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning expo: Exploration-prioritized policy optimization via adaptive kl regulation and gaussian curriculum sampling
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 69b0009d-c43c-4ea1-8c06-e629fe386bee · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Stabilizing Knowledge, Promoting Reasoning: Dual-Token Constraints for RLVR
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51f81fca-0199-4d8e-8479-56237d36f1c1 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b9db30a-cc20-4107-8274-997583ac6163 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Findings of the Association for Computational Linguistics: NAACL 2025 , pages=
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb170a3c-c32c-4233-8b8a-7deb11a1ecc8 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0008add1-0006-4660-b26c-238db2c742fd · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Group Sequence Policy Optimization
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b978fe9-44cc-42d7-996e-074b439cce26 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning SimpleTIR: End-to-End Reinforcement Learning for Multi-Turn Tool-Integrated Reasoning
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61c442a8-efe1-4f94-811d-0f65aa07c352 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning arXiv preprint arXiv:2509.21826 , year=
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation beb7572e-2aad-4fb6-9fb2-415e538b6b4d · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1760ee1a-2a76-4fac-b3b8-d602362fe35f · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11a28c1b-8692-4df2-bc0f-b48642f39907 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Advances in Neural Information Processing Systems , volume=
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 513eb1db-d952-4e0c-abc8-c1e864933d55 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning arXiv preprint arXiv:2507.14783 , year=
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0b85819-9a32-450d-990c-f7710a6cf889 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Can One Domain Help Others? A Data-Centric Study on Multi-Domain Reasoning via Reinforcement Learning
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e48110fa-5932-4e42-a981-6716065cab6a · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning arXiv preprint arXiv:2602.12566 , year=
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78f8ee3c-a707-4ca5-881b-389d9ef6b591 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning arXiv preprint arXiv:2602.02301 , year=
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2de762b3-6690-4c75-a29c-ff5df8769829 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning arXiv preprint arXiv:2505.17508 , year=
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0ab0c6d-c983-4837-ade3-88399ac72e09 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Advances in Neural Information Processing Systems , volume=
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a25a14df-98a3-4de8-be31-241cdd07debb · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e66e8458-af75-4c5b-8783-51bc09ccd8b2 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Advances in Neural Information Processing Systems , volume=
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a5b07d83-b46e-4da2-af82-47e44e99f983 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Measuring Mathematical Problem Solving With the MATH Dataset
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fefc93de-90e7-4a81-a7e9-de149a2a8cef · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 440a5857-0cef-47f0-9e65-01ef9da28d69 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning WildChat: 1M Chat
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d85754cd-3a0f-47a7-96bc-de263b245193 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Advances in Neural Information Processing Systems , volume=
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d3c600e-de81-43fc-b9b7-4ad1eebf5f11 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning arXiv preprint arXiv:2512.15489 , year =
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b129d2e-ab42-4199-8d96-22397aa477a2 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning arXiv preprint arXiv:2509.20357 , year=
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f41ce78-a012-42b9-ab64-9f1f61dd0740 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbd38b4d-117c-43b4-81ca-2cbdc185adab · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning arXiv preprint arXiv:2512.05962 , year=
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 490492f5-7b91-4f81-b8db-d1e281bbd6c6 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Uniform-Correct Policy Optimization: Breaking RLVR's Indifference to Diversity
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 14284321-4222-44e3-864a-f7944fb9fa81 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning arXiv preprint arXiv:2602.19895 , year=
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d40d7bb1-73a0-4e5b-804e-e4a11b18e0cf · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0309255f-307c-41e5-8890-66f275465eba · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning International Conference on Learning Representations , volume=
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ee7e51f3-7ccf-4a53-b5e3-504e3936633f · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 567c4997-dfc7-44bc-90ee-c1b8dd195bdc · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning arXiv preprint arXiv:2507.14843 , year=
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a43e7f0-9217-49ef-97ad-685026069a89 · outbound
Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning Understanding R1-Zero-Like Training: A Critical Perspective
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.