Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-30T21:12:59.453856Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2607.26784.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-30T21:12:59.453856Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T19:53:25.916475Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T19:53:27.058819Z
30 of 30 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c1fe03a0-f688-48a9-bdc8-d2d1679c3821 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Group-in-Group Policy Optimization for LLM Agent Training
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9c878b0-f740-42ec-bfb0-593fc9ed2b90 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fd41044-b57f-451d-942c-c34a313e79b7 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Hierarchy-of-groups policy opti- mization for long-horizon agentic tasks.arXiv preprint arXiv:2602.22817,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f4c6aba-b412-47bf-820e-799e47818d37 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Meta-rl induces explo- ration in language agents.arXiv preprint arXiv:2512.16848,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7576976-63a4-4c83-89a4-269ee79ae7c9 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Self-Distilled Agentic Reinforcement Learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52253690-8d56-4915-ba29-1014e81f4c93 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91a61321-88af-4969-ab41-b190bc5386db · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution ReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76ddb756-05b0-4672-a88d-6b70121e391c · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution SkillOS: Learning Skill Curation for Self-Evolving Agents
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 621de15e-e44f-4f77-91ff-bdeb1b8acde4 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Webrl: Training llm web agents via self-evolving online curriculum reinforcement learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b14af90f-e9c6-4e11-bc14-f2653b3e5646 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Autorefine: From trajectories to reusable expertise for continual llm agent refinement
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 051619d9-698c-41d2-a156-23f3a22854bc · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Proximal Policy Optimization Algorithms
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06dce564-2ccb-414d-864c-37813649e9cd · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d161b1b6-f3e8-44f0-8372-1895d21d9b32 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Milestone-Guided Policy Learning for Long-Horizon Language Agents
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2dc2ff19-2d82-4a25-b66a-44a6ea5b63dd · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e2c48f9-e951-48b3-8d4a-71238d638059 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0700bb3-ee9b-499a-80ae-c4b63f8744ee · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement Learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eac22b0e-9bf5-4cfc-be3a-592faf33dd12 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Qwen3 Technical Report
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b61be434-2d06-46dd-aa2b-719bbe066079 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution SkillOpt: Executive Strategy for Self-Evolving Agent Skills
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5eaf3873-1ce2-4aef-ac20-0b124d42ce0b · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution ReAct: Synergizing Reasoning and Acting in Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2e17167-aedf-42f2-af8b-a9e7e015f472 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Look Before You Leap: Autonomous Exploration for LLM Agents
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a145b9c-9181-430b-acfa-2b13de815d79 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution The Landscape of Agentic Reinforcement Learning for LLMs: A Survey
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c31621b1-d508-40a7-9c59-a565a6f3e871 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 884533dd-6334-41ee-af77-a2f36bc6e103 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce5cf7ee-9bd5-4544-bea0-a63c0043f712 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cb6b808-b81d-4065-a5ec-48e0eaf66b4c · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Reinforcement learning for self-improving agent with skill library
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4af1e024-fd0b-47a0-a885-242921623c03 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f822171c-725d-4a6a-afe9-60b799392731 · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution ALFWorld: Aligning Text and Embodied Environments for Interactive Learning
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a8d2d8c-6f80-4231-979d-dcf2680be10e · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0701f85b-1ea8-4b58-99bc-4f33f420f26e · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Yuxin Chen, Yu Wang, Yi Zhang, Ziang Ye, Zhengzhou Cai, Yaorui Shi, Qi Gu, Hui Su, Xunliang Cai, Xiang Wang, et al
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d9ba3fa-3322-402a-a4ca-e715ccc16e2f · outbound
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Evolving-RL: End-to-End Optimization of Experience-Driven Self-Evolving Capability within Agents
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 711881f7-a651-4888-8f97-894bf50f5a8a · inbound
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.