Pith. sign in

Paper Citation Record · LEDGER

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution

As of 19 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2607.26784.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.26784 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-30T21:12:59.453856Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T19:53:25.916475Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T19:53:27.058819Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c1fe03a0-f688-48a9-bdc8-d2d1679c3821 · outbound

This paper cites Group-in-Group Policy Optimization for LLM Agent Training.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Group-in-Group Policy Optimization for LLM Agent Training

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.359390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.359390Z digest=sha256:6c5b2948d0887f95662251d4357a7078aa19659178873bb735b7d3de7002f3e1

Observation c9c878b0-f740-42ec-bfb0-593fc9ed2b90 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.366710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.366710Z digest=sha256:97919a99bdaa0a493624f49d49ffed9f2dd2ecc33f481a3302b84dd7855792d0

Observation 5fd41044-b57f-451d-942c-c34a313e79b7 · outbound

This paper cites Hierarchy-of-groups policy opti- mization for long-horizon agentic tasks.arXiv preprint arXiv:2602.22817,.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Hierarchy-of-groups policy opti- mization for long-horizon agentic tasks.arXiv preprint arXiv:2602.22817,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.370562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.370562Z digest=sha256:ec70d79c7f34c5b9bf8444e01f03bbdaf03f9ead7740b31b06f48066adf6433c

Observation 2f4c6aba-b412-47bf-820e-799e47818d37 · outbound

This paper cites Meta-rl induces explo- ration in language agents.arXiv preprint arXiv:2512.16848,.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Meta-rl induces explo- ration in language agents.arXiv preprint arXiv:2512.16848,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.373942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.373942Z digest=sha256:670461f70034efea027654b6f31824d421533d935772507c0cf4ff4cfad11ca5

Observation a7576976-63a4-4c83-89a4-269ee79ae7c9 · outbound

This paper cites Self-Distilled Agentic Reinforcement Learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Self-Distilled Agentic Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.377118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.377118Z digest=sha256:1bcf46d030df601338559abb4966ac9c3f19c9674ffa61034ebcc2e7316aceb9

Observation 52253690-8d56-4915-ba29-1014e81f4c93 · outbound

This paper cites Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.381111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.381111Z digest=sha256:c0bf7750c5be5889b91711fd30acfae624746261926460e117f6ee9f569f4fab

Observation 91a61321-88af-4969-ab41-b190bc5386db · outbound

This paper cites ReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution ReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.384535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.384535Z digest=sha256:728b7e60602ed661f275bff9df83045cd7122847a3c6401342b117285eb0ce09

Observation 76ddb756-05b0-4672-a88d-6b70121e391c · outbound

This paper cites SkillOS: Learning Skill Curation for Self-Evolving Agents.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution SkillOS: Learning Skill Curation for Self-Evolving Agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.388053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.388053Z digest=sha256:c45c28acb4bf25f6dcf03193088852221a1719fa73c178bd69a1a5663834bf2c

Observation 621de15e-e44f-4f77-91ff-bdeb1b8acde4 · outbound

This paper cites Webrl: Training llm web agents via self-evolving online curriculum reinforcement learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Webrl: Training llm web agents via self-evolving online curriculum reinforcement learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.391280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.391280Z digest=sha256:03e71ff4bfc44d9abd5913bdb0c221b7d3bf4dbf540d3bc53d786fd844f00495

Observation b14af90f-e9c6-4e11-bc14-f2653b3e5646 · outbound

This paper cites Autorefine: From trajectories to reusable expertise for continual llm agent refinement.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Autorefine: From trajectories to reusable expertise for continual llm agent refinement

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.394453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.394453Z digest=sha256:acb561d42ff60fda5e28a26a3182779292d2842773325fa26e0f035ffca7ad10

Observation 051619d9-698c-41d2-a156-23f3a22854bc · outbound

This paper cites Proximal Policy Optimization Algorithms.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Proximal Policy Optimization Algorithms

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.397336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.397336Z digest=sha256:4eb8241d0721cb49c9e1de000faa3d2d335ee44d5c99ea8a9e211caea2c51339

Observation 06dce564-2ccb-414d-864c-37813649e9cd · outbound

This paper cites Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.403621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.403621Z digest=sha256:fc86a68bdbe3d8ef0a6cc9c0770e90edd3303ffc474fe57b00fccdb67ef37041

Observation d161b1b6-f3e8-44f0-8372-1895d21d9b32 · outbound

This paper cites Milestone-Guided Policy Learning for Long-Horizon Language Agents.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Milestone-Guided Policy Learning for Long-Horizon Language Agents

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.416855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.416855Z digest=sha256:7356f9773398ed7ed3543f04af0f75cb74c1d4230ad02233f1e1a61210c44863

Observation 2dc2ff19-2d82-4a25-b66a-44a6ea5b63dd · outbound

This paper cites SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.425331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.425331Z digest=sha256:5f69a87d06c0a420cc4a73751ac0ada5a05fbffef37f2ec20a627fe5b67ba888

Observation 4e2c48f9-e951-48b3-8d4a-71238d638059 · outbound

This paper cites AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.428554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.428554Z digest=sha256:a0c0ed87d926d86d26f8975edbebc94c06e52d4e91b96219f80441d6642c89c1

Observation f0700bb3-ee9b-499a-80ae-c4b63f8744ee · outbound

This paper cites SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement Learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.431570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.431570Z digest=sha256:12071e91ae579b965bf83f5741e8c57418b49dd17e3ee08f3a8d1326f1ba4c2f

Observation eac22b0e-9bf5-4cfc-be3a-592faf33dd12 · outbound

This paper cites Qwen3 Technical Report.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Qwen3 Technical Report

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.434695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.434695Z digest=sha256:6a8d9b458a42d7e7c2aab5870a1d2ad4e458a368a6e89be31f50b6074ee8f366

Observation b61be434-2d06-46dd-aa2b-719bbe066079 · outbound

This paper cites SkillOpt: Executive Strategy for Self-Evolving Agent Skills.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution SkillOpt: Executive Strategy for Self-Evolving Agent Skills

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.437826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.437826Z digest=sha256:8897558e11dfdf785c5984a295d3ec64c5ddb8b47e6a286d95a5e75e5365a348

Observation 5eaf3873-1ce2-4aef-ac20-0b124d42ce0b · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution ReAct: Synergizing Reasoning and Acting in Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.440916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.440916Z digest=sha256:b5f4dcb388737e83b32cd7955c0a9da0f8a6dad22dfa72d02937faf8d24383f8

Observation b2e17167-aedf-42f2-af8b-a9e7e015f472 · outbound

This paper cites Look Before You Leap: Autonomous Exploration for LLM Agents.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Look Before You Leap: Autonomous Exploration for LLM Agents

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.444076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.444076Z digest=sha256:c8e304f726792aaff776b2dae0a59d4bd296d93c568b062d535a070d8f1bd1e3

Observation 7a145b9c-9181-430b-acfa-2b13de815d79 · outbound

This paper cites The Landscape of Agentic Reinforcement Learning for LLMs: A Survey.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution The Landscape of Agentic Reinforcement Learning for LLMs: A Survey

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.446937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.446937Z digest=sha256:ee029a730fa22c0a4485d01f32957c33604bed12a7d1db210da425bef2ff20f8

Observation c31621b1-d508-40a7-9c59-a565a6f3e871 · outbound

This paper cites MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.449832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.449832Z digest=sha256:4e5e7406db0c8e1be46f73165f566c1e8ade360076b5472280a7b2a5b39d540d

Observation 884533dd-6334-41ee-af77-a2f36bc6e103 · outbound

This paper cites LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.453856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.453856Z digest=sha256:ca8a144e0a1008b468db3b08f618626214f00de7aefbaa02418f3beb99ba89eb

Observation ce5cf7ee-9bd5-4544-bea0-a63c0043f712 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.400414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.400414Z digest=sha256:983c9c5a547b4df72ab4f7fb0de6c0f2c4b051e324fb7f63319bcb1ef440745a

Observation 9cb6b808-b81d-4065-a5ec-48e0eaf66b4c · outbound

This paper cites Reinforcement learning for self-improving agent with skill library.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Reinforcement learning for self-improving agent with skill library

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.410183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.410183Z digest=sha256:0ebc6dbcb58ef7c86eaca18af1d030a6926c85edee9adce29f07034118cbdea9

Observation 4af1e024-fd0b-47a0-a885-242921623c03 · outbound

This paper cites RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.413803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.413803Z digest=sha256:3f853af3de81b61023cd58281d6e02f87a0cc236dca86708b9741e035a087a80

Observation f822171c-725d-4a6a-afe9-60b799392731 · outbound

This paper cites ALFWorld: Aligning Text and Embodied Environments for Interactive Learning.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution ALFWorld: Aligning Text and Embodied Environments for Interactive Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.406582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.406582Z digest=sha256:e95397c8b24f37699a245fd1c5ad22a8fa539e69a29ebfeae1d268d8f352ce21

Observation 7a8d2d8c-6f80-4231-979d-dcf2680be10e · outbound

This paper cites A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.362919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.362919Z digest=sha256:cfb42b43bbd65eba1368b29bd1765de3e766fb7623355c032b2bf072abaae8d8

Observation 0701f85b-1ea8-4b58-99bc-4f33f420f26e · outbound

This paper cites Yuxin Chen, Yu Wang, Yi Zhang, Ziang Ye, Zhengzhou Cai, Yaorui Shi, Qi Gu, Hui Su, Xunliang Cai, Xiang Wang, et al.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Yuxin Chen, Yu Wang, Yi Zhang, Ziang Ye, Zhengzhou Cai, Yaorui Shi, Qi Gu, Hui Su, Xunliang Cai, Xiang Wang, et al

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.351192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.351192Z digest=sha256:5d2e745d305e80e4b67bc7998a7a92489f90f257e415382d3922157d22456a1e

Observation 4d9ba3fa-3322-402a-a4ca-e715ccc16e2f · outbound

This paper cites Evolving-RL: End-to-End Optimization of Experience-Driven Self-Evolving Capability within Agents.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Evolving-RL: End-to-End Optimization of Experience-Driven Self-Evolving Capability within Agents

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.355533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.355533Z digest=sha256:375b8e93e93ca87138fbd31d76fa175000d85cf7cf80f0466c6dbf6cb7daf935

Pith citing papers

Observation 711881f7-a651-4888-8f97-894bf50f5a8a · inbound

AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning cites this paper.

AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution

Reference 66

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T19:53:27.063688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T19:53:25.916475Z digest=sha256:a0b58ae6d0caecce146e669e34122f2c8d9009970ac9f15f1a7f33e4123b5795