Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T14:33:00.408984Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 2 inbound Pith citation observations for arXiv:2606.02355.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T14:33:00.408984Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-31T22:51:49.158185Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-06-30T07:24:21.967669Z
48 of 48 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d7c3edcb-85a9-4a69-986e-7df18891d504 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Aho and Jeffrey D
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f85de3f7-2069-4e57-a1ac-c69ce3ca98f4 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Unresolved cited work
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5787217c-216d-4f65-806d-db646afb91c5 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Chandra and Dexter C
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation de959fcc-80c6-4cc3-b73f-2844d00fc798 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Scalable training of
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23b150b9-93d5-499b-a65b-2b0d6c573242 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51c9992b-4ebe-493d-a516-7cb1b015b2b1 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Tetreault , title =
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f67183c-b721-4e5f-96f2-f28423392952 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training A Framework for Learning Predictive Structures from Multiple Tasks and Unlabeled Data , Volume =
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e420bb1-fd22-4df5-b64c-bd3c5788ae74 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 24cbd096-6109-4bb2-9f40-25890e9e900d · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Dynamic Dual-Granularity Skill Bank for Agentic RL
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ea0ff959-6359-41b4-8645-9ec4e6f7bc0f · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 00affdbf-b378-48f0-a1a9-3c7d616ce018 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Arise: Agent reasoning with intrinsic skill evolution in hierarchical reinforcement learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 042e0ba2-0df4-40ae-87c8-2083ffa7034b · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Voyager: An Open-Ended Embodied Agent with Large Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1fa1e855-6224-4b7b-a3dc-ac1b4fa03c0d · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Advances in Neural Information Processing Systems , year=
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afe4a27f-2d1d-4260-9714-9a8e970d022e · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Proceedings of the AAAI Conference on Artificial Intelligence , year=
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9bfeafb-591e-4ae8-8461-5c63fa65ed0f · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training International Conference on Learning Representations , year=
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ea5c6d2-e63d-4371-85e8-41be3235031e · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Advances in Neural Information Processing Systems , year=
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 583d742e-a60d-475f-aa98-435e268757e2 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training AppWorld: A Controllable World of Apps and People for Benchmarking Interactive Coding Agents
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 373e9909-1ed6-4f11-81be-f901b1e7e36d · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training International Conference on Learning Representations , year=
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9995315c-f420-4a84-b9bd-b3cf5ecc40e3 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 65cb10c8-1ace-4f2f-a748-6600542717ee · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Tree search for llm agent reinforcement learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2fdd18d3-aac2-43fc-a2a2-7cb9aad31500 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Group-in-Group Policy Optimization for LLM Agent Training
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 24f30cd9-5e82-445f-9802-4b69b2865469 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training arXiv preprint arXiv:2603.08754 , year=
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2af2748a-b2ed-4162-9698-d1db558cb154 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training arXiv preprint arXiv:2603.03078 , year=
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 95c90cbf-0e8c-4441-b627-c857aa5d77a3 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4b69b58f-aa39-40de-b262-02a844b3d7b3 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training MemRL: Self-Evolving Agents via Runtime Reinforcement Learning on Episodic Memory
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d501a210-3cb5-405f-bff5-42b65864e08f · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 83b74d42-2047-4d3b-8f56-9c42c8448584 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training SimpleMem: Efficient Lifelong Memory for LLM Agents
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 55396aae-7a40-42cb-bfd7-0718b1fea377 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b0552a8d-3632-4538-96ea-8e02ce667168 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Advances in neural information processing systems , volume=
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3f5b415-4c40-4d00-a53f-729b798fb1ed · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 567f7c6c-6417-4799-8242-3ce7723d3ab8 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Proceedings of the 61st annual meeting of the association for computational linguistics (volume 1: Long papers) , pages=
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b8ec7d5-0b75-4e95-87f4-518b8a547c21 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training WebGPT: Browser-assisted question-answering with human feedback
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d8d7ca4-61ad-4787-9cb2-05c8e7a187d1 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Advances in neural information processing systems , volume=
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e8dbb42-a2ff-4e2a-b88c-001e7d34f15d · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training nature , volume=
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38b6f1c7-a6f1-429c-b5cf-0c0208c01539 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing , pages=
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0101eaa4-0d34-442b-bd2d-57357d55f2e1 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6638e893-ea36-4086-aeff-aafae6be1d09 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Proceedings of the AAAI Conference on Artificial Intelligence , volume=
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e73b26fe-012c-41a7-8cf0-a8486627ae88 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training ArCHer: Training Language Model Agents via Hierarchical Multi-Turn RL
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6759baf3-9ac4-4d18-a34a-0ade98708a39 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e0c6c9f2-3dda-46d5-8b59-7eb88bd81ba9 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Towards Efficient Online Tuning of VLM Agents via Counterfactual Soft Reinforcement Learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2e23d26e-160b-4cee-8734-d8aaf6893079 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Reinforcement Learning for Long-Horizon Interactive LLM Agents
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0bff4fa0-2046-4936-924d-e42815f36a12 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing , pages=
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e8916c8-2fc7-468c-b68b-306bb611548d · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training From Reasoning to Agentic: Credit Assignment in Reinforcement Learning for Large Language Models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d757079b-5053-4254-a5e2-deba08a82383 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b1d74dbb-ec2c-4ebe-8bd7-71095a4027d2 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training 1998 , publisher=
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3a9b7c5-7764-4e36-aac8-39a7f268ae5b · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Deep Reinforcement Learning in Large Discrete Action Spaces
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6b95f177-45a9-479a-bd6b-1707712d481d · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Proximal Policy Optimization Algorithms
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 453b1ca1-d603-4e5f-bc28-0ff5c8f192a8 · outbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cc2f660e-94ac-46f5-9b76-8877455630b6 · inbound
UCOB: Learning to Utilize and Evolve Agentic Skills via Credit-Aware On-Policy Bidirectional Self-Distillation SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 23a05fea-0bd5-4aee-93f2-bac8687bf5f0 · inbound
From Scoring to Acting: Outcome-Verified Comparative Self-Distillation for LLM Agents SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.