Pith. sign in

Paper Citation Record · LEDGER

Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

As of 3 August 2026, this Paper Citation Record lists 6 of 6 outbound references and 34 inbound Pith citation observations for arXiv:2602.01869.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2602.01869 v3

Coverage vector

measured 6 of 6 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T05:35:32.207725Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00

measured 34 of 34 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T02:07:42.304450Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T12:15:01.137692Z

Reference resolution

6 of 6 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 193de6b0-389e-4fe7-99f5-6aeee1d02f7c · outbound

This paper cites an unresolved cited work.

Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T05:35:31.392350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:35:31.392350Z digest=sha256:9e650d9df1b40efb98a81ce37f7cdba18c395bf1aa6a3b5cc35c734e55cb93dc

Observation d94490ce-34cb-42aa-a303-e4657e4a0cce · outbound

This paper cites an unresolved cited work.

Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T05:35:31.506716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:35:31.506716Z digest=sha256:d9b60add52b2ae42e368641f638bc3034110b44da6324bc310679c37fe81bccd

Observation 4202d4c7-abb9-4f07-9966-3bf8adc03256 · outbound

This paper cites •Termination: The first valid move is submitted and initial feedback is received.

Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents •Termination: The first valid move is submitted and initial feedback is received

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T05:35:31.672300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:35:31.672300Z digest=sha256:30434968298b9dc8d9a35d9de4438e27dea55063d82bedebaa0e8cdf7bdf2ebe

Observation aa9f745b-9eb9-489c-8478-4daeece56650 · outbound

This paper cites Add a check for X.

Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents Add a check for X

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T05:35:31.880467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:35:31.880467Z digest=sha256:ead61fea4481034d72520f2e7755e891377c35c4e797eaf01e35ef7a525ab674

Observation e6847a32-8e09-4815-bd39-6752bce422c9 · outbound

This paper cites IF” condition to ensure the skill only starts in valid states. •Policy (π): Update the 3–5 reasoning steps to bypass identified failure modes. •Termination (β): Update the “Stop IF.

Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents IF” condition to ensure the skill only starts in valid states. •Policy (π): Update the 3–5 reasoning steps to bypass identified failure modes. •Termination (β): Update the “Stop IF

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T05:35:32.107876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:35:32.107876Z digest=sha256:1b4db46eb7a7bdaa9fd71a35f59694d85991bd58fe42d34cce900b23e2c5f40e

Observation 27037b2a-8fe7-4864-90df-8af1f6b440c0 · outbound

This paper cites skill_name.

Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents skill_name

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T05:35:32.207725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:35:32.207725Z digest=sha256:fb7bb47bc2345f067504b38b661dca821dd9bb7fe679ee8fb854dc8f82a12253

Pith citing papers

Observation 19fef1f3-3552-44a1-b349-6ab868e31d92 · inbound

From Procedural Skills to Strategy Genes: Towards Experience-Driven Test-Time Evolution cites this paper.

From Procedural Skills to Strategy Genes: Towards Experience-Driven Test-Time Evolution Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-10T10:55:04.021820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-10T10:52:47.760627Z digest=sha256:9acba8e164dade2cc3662c76e23e2c17a99c5d386cffc2f16a2b9b2cef01df35

Observation 4258eae8-cfff-46a6-8abf-66f15ef23617 · inbound

From Procedural Skills to Strategy Genes: Towards Experience-Driven Test-Time Evolution cites this paper.

From Procedural Skills to Strategy Genes: Towards Experience-Driven Test-Time Evolution Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T19:51:48.531280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T19:51:48.531280Z digest=sha256:4e4b125d32450b7e0d96e5e9e4ee1168bdf521418ef37062f96f71bc663fdcf5

Observation 7a4fd127-8aab-40e2-85b0-50d5d8be1649 · inbound

Co-Evolving LLM Decision and Skill Bank Agents for Long-Horizon Tasks cites this paper.

Co-Evolving LLM Decision and Skill Bank Agents for Long-Horizon Tasks Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-10T00:14:46.447134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-10T00:14:07.017420Z digest=sha256:43e9795220029df78b56ae01dcaeea8120961fab6f0fb8ae980eea6c1a11a85d

Observation 48b38cff-4fca-4191-bdd4-b632234a76d1 · inbound

MEMOREPAIR: Barrier-First Cascade Repair in Agentic Memory cites this paper.

MEMOREPAIR: Barrier-First Cascade Repair in Agentic Memory Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T01:40:51.717079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-11T01:39:26.332430Z digest=sha256:f4ea9a2de0fd2e529be0fa76fcacf62ea937817ba73d35fcc97b7251821fc28d

Observation 7bb7ce15-62da-4507-abf7-0074f2037d77 · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-11T04:20:56.979632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-11T01:47:39.926540Z digest=sha256:a52d8357272fa8c3dcdf4a0e71c7795c97c6d6d903b6690802ff5b5360d0711d

Observation de8d5e83-a42e-488f-b143-4c0fc774e373 · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-20T23:19:14.848201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-20T23:15:44.550045Z digest=sha256:8eaafe4be5f10a1f124c90941e1fbf2b5caa6d8d8bcc14923c09b1f3d35b5c7b

Observation da1bfba5-e302-4e74-87a2-1b7794019553 · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:25:07.275896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-30T23:23:42.883286Z digest=sha256:dbd6fdfa1a953cbe32967c3bc0e9cafae9e7d10e05f1efd33d408ce8c599116f

Observation 9d0f598c-3ef0-4dbc-8918-e656370e005e · inbound

SkillLens: Adaptive Multi-Granularity Skill Reuse for Cost-Efficient LLM Agents cites this paper.

SkillLens: Adaptive Multi-Granularity Skill Reuse for Cost-Efficient LLM Agents Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-12T08:21:25.556207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-12T01:12:52.837729Z digest=sha256:62349b432493c7fd0a3640ef68d0faf3d4d7982318fa30e3aec7dc7ece05627b

Observation 6ed8dd37-c8ad-47a9-a70b-834b2a006c71 · inbound

Skill-CMIB: Multimodal Agent Skill for Consistent Action via Conditional Multimodal Information Bottleneck cites this paper.

Skill-CMIB: Multimodal Agent Skill for Consistent Action via Conditional Multimodal Information Bottleneck Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T08:01:28.230686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T01:24:59.721212Z digest=sha256:d992b114c4e39b47b85752bab61aa5bd3a51339e39919fb2f1119c5936b73536

Observation 37c53803-2599-4fae-b568-9d9db4e94bc6 · inbound

Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning cites this paper.

Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-12T07:06:33.576653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-12T03:45:06.199636Z digest=sha256:90bbf9d6320a93d0f362a68b6e113bbde3a4f36ef40dc7ed5adf6cbdaf5b78b6

Observation 6d69f720-e582-485e-bf97-63ff4c98a1da · inbound

Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning cites this paper.

Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-20T22:23:48.304691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-20T22:19:49.016156Z digest=sha256:9eb4025433f22952b396c856be6449d452f17beb03672de6575a7250211e2bfc

Observation 64d11719-9425-486c-8327-5327cd8f025f · inbound

MAP: A Map-then-Act Paradigm for Long-Horizon Interactive Agent Reasoning cites this paper.

MAP: A Map-then-Act Paradigm for Long-Horizon Interactive Agent Reasoning Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-14T19:49:25.815783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-14T19:48:53.213389Z digest=sha256:c6965d3a8b8efc2fd92c815d76f31003e6cfd30aabc5c57c92f6d2620a06df8d

Observation b6806335-5862-4164-98fa-da08dac27af2 · inbound

SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution cites this paper.

SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-20T11:43:15.029937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-20T11:40:45.397038Z digest=sha256:fe522cd66756eadf4aec3a32260a149ce2651b702f4153fc2ba616636056e594

Observation 33adaaef-4515-41ba-9c55-4e9d1c1439a6 · inbound

EvoIR-Agent: Self-Evolving Image Restoration Agentic System via Experience-Driven Learning cites this paper.

EvoIR-Agent: Self-Evolving Image Restoration Agentic System via Experience-Driven Learning Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:51:16.155833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T07:49:15.291339Z digest=sha256:36384692fe523cf63033f7bacde746f3060b362f24b6eeee1eecb90189964ca6

Observation 238e552a-644a-4969-bb10-d809da6210a9 · inbound

EvoIR-Agent: Self-Evolving Image Restoration Agentic System via Experience-Driven Learning cites this paper.

EvoIR-Agent: Self-Evolving Image Restoration Agentic System via Experience-Driven Learning Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-06-30T17:24:57.057323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-30T17:22:36.435074Z digest=sha256:bfcb125129c5b6c738c89c4ed4925680bd3271e91dcd8e89b1d531441b41d704

Observation afc5280c-78d0-4387-b533-aba5625b5b10 · inbound

From Raw Experience to Skill Consumption: A Systematic Study of Model-Generated Agent Skills cites this paper.

From Raw Experience to Skill Consumption: A Systematic Study of Model-Generated Agent Skills Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-25T03:55:20.457433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-25T03:53:56.643075Z digest=sha256:c4bac35eb9e77262e34bd5b8638e03748f2bfe2e7c6635a175901c3c9928fa16

Observation 3cfbaf3c-e1ea-4c7c-a77f-4558e7349d39 · inbound

SkillOpt: Executive Strategy for Self-Evolving Agent Skills cites this paper.

SkillOpt: Executive Strategy for Self-Evolving Agent Skills Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-25T03:56:36.893749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-25T03:51:49.812710Z digest=sha256:1e22795c8fc096d7f502183945f2d478bcbfb454d965604d7d7660b675726328

Observation 51c09df6-9169-4d7b-b5f0-0343cd0febb0 · inbound

SkillOpt: Executive Strategy for Self-Evolving Agent Skills cites this paper.

SkillOpt: Executive Strategy for Self-Evolving Agent Skills Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:35:12.842968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-30T16:26:23.130418Z digest=sha256:8f983fafe0279f9610ff250cf70268427e6bc660a88efa2ad4bab4b5295cf068

Observation 20bd4c45-8216-4088-8122-e0f4cf1d07ac · inbound

ElasticMem: Latent Memory as a Learnable Resource for LLM Agents cites this paper.

ElasticMem: Latent Memory as a Learnable Resource for LLM Agents Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-06-29T00:22:51.816491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-28T23:06:57.377183Z digest=sha256:46be73d167f8af93da07540273475804fc53b2c2ad0711e7324a09a5cd6c4087

Observation b5c8730b-9e81-4701-b55a-045d1a32d7dd · inbound

SkillRevise: Improving LLM-Authored Agent Skills via Trace-Conditioned Skill Revision cites this paper.

SkillRevise: Improving LLM-Authored Agent Skills via Trace-Conditioned Skill Revision Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-06-28T17:42:24.840642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-06-28T17:37:31.081762Z digest=sha256:b247e1fb8440361494cf1cee40b66023fc542e25042e494252d2a57ea6013de9

Observation d50b2e55-fb36-4925-88bc-92a7534abfd5 · inbound

SkillPyramid: A Hierarchical Skill Consolidation Framework for Self-Evolving Agents cites this paper.

SkillPyramid: A Hierarchical Skill Consolidation Framework for Self-Evolving Agents Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T03:16:33.461204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-06-28T10:15:03.274182Z digest=sha256:afc9a2d6e34576e8278367d9e54e1ae3dfd2cb26eefc4d3679cad3eb364b2ab5

Observation 4b598877-4146-4006-a466-6af2f89edeb5 · inbound

SkillJuror: Measuring How Agent Skill Organization Changes Runtime Behavior cites this paper.

SkillJuror: Measuring How Agent Skill Organization Changes Runtime Behavior Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-03T10:17:57.313157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-27T10:12:03.985549Z digest=sha256:dfc304837206a8cc07fc81e1a6b8de3d0822dea11711a5bd6d948aeb010e0358

Observation d6c60a18-c9a4-421a-a523-91721422df55 · inbound

Organize then Retrieve: Hierarchical Memory Navigation for Efficient Agents cites this paper.

Organize then Retrieve: Hierarchical Memory Navigation for Efficient Agents Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-07-03T10:17:57.783648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-27T10:06:59.839645Z digest=sha256:bf389e4690b30fb6866379fadf3e6bb498fc074e43a038bfd3d40e99b48493c7

Observation c3e04f58-2b8a-40f6-ae1f-f639b524ab25 · inbound

PreAct: Computer-Using Agents that Get Faster on Repeated Tasks cites this paper.

PreAct: Computer-Using Agents that Get Faster on Repeated Tasks Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-07-03T20:58:57.438121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-27T01:06:26.643487Z digest=sha256:68980855d8a2de00d248a7d38aac2d99bd946fa597ccb8b51fad669dc73dfdec

Observation 2a6cccdd-3991-4a64-903e-1dc4ee7015f3 · inbound

Metis: Bridging Text and Code Memory for Self-Evolving Agents cites this paper.

Metis: Bridging Text and Code Memory for Self-Evolving Agents Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T16:29:56.683203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-26T00:42:55.758578Z digest=sha256:17c3f3f3b1d26dc79d9b8aaa459defded93b535572d206dbaa90a861cb459c1e

Observation af1bef39-b102-45a6-9224-45e2b514305e · inbound

UCOB: Learning to Utilize and Evolve Agentic Skills via Credit-Aware On-Policy Bidirectional Self-Distillation cites this paper.

UCOB: Learning to Utilize and Evolve Agentic Skills via Credit-Aware On-Policy Bidirectional Self-Distillation Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T07:24:22.348730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-06-30T07:15:44.387347Z digest=sha256:fdbaa4e6e14237575efbd87e88a947a29321bf69c0e971bcfa5d61aa838df700

Observation 0f0a5d9f-79fd-4c3e-86d4-2601cfd7ee42 · inbound

LLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via State Proprioception cites this paper.

LLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via State Proprioception Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T10:46:46.380037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T10:46:46.380037Z digest=sha256:2a36d1416f699317d1f7d2ac692bb184487cb69aa870d270b8670e43d09f0655

Observation eb408c66-2785-45b2-ac2d-50964b7b290c · inbound

LLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via State Proprioception cites this paper.

LLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via State Proprioception Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T09:38:42.966270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:38:42.966270Z digest=sha256:8b62b9f398b99792507a7bb0fc8a72b292e1d7fc9020a1030d200af5727480cf

Observation 59ada0ac-c954-4da0-8c1d-783aa83c2fed · inbound

LLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via State Proprioception cites this paper.

LLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via State Proprioception Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T02:07:42.304450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:07:42.304450Z digest=sha256:3be010db4679f616259ede5a49fb285839cf632c8e2da86cb7a41b5202dce8b2

Observation 66b0aea3-0daa-4bc5-8200-ee51091e9c3c · inbound

EvoAgentBench: Benchmarking Agent Self-Evolution via Ability Transfer cites this paper.

EvoAgentBench: Benchmarking Agent Self-Evolution via Ability Transfer Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-07T23:54:18.645734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-07-07T23:48:16.309103Z digest=sha256:9e2e550f82a014cbb451ecfce857a6868bda1fed2fbd249bc4669db4f28c8b5d

Observation df9cfc06-f91c-4eb0-b324-8f7ea915dd0c · inbound

Memory as a Controlled Process: Learned Adaptive Memory Management for LLM Agents cites this paper.

Memory as a Controlled Process: Learned Adaptive Memory Management for LLM Agents Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T04:50:36.427905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:50:36.427905Z digest=sha256:6d30c41c8a5da0aeb68f45433f8427dffcb1bc8a51aaee4f5aeb649ebbe19346

Observation d3f9d975-5668-41f8-88d9-e14b2bc7d137 · inbound

Experience Memory Graph: One-Shot Error Correction for Agents cites this paper.

Experience Memory Graph: One-Shot Error Correction for Agents Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T03:29:10.739890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:29:10.739890Z digest=sha256:37dcc6a80e3695ece81c056ea7fd8928ea8a541553e6a95048004b496dd369b2

Observation c4bc95f6-2965-4b6a-829b-8c0d1c04ab9f · inbound

From Memory to Skills: Evidence-Grounded Co-Evolution Governance for Long-Horizon LLM Agents cites this paper.

From Memory to Skills: Evidence-Grounded Co-Evolution Governance for Long-Horizon LLM Agents Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T20:27:33.860778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:27:33.860778Z digest=sha256:33166463b3a56dafd99aec2e1ae1d8ac12bc84578e945bb63c8161805dcc8f6e

Observation 4d74f686-e46e-4e04-80d0-371b74be435c · inbound

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills cites this paper.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:47.064528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:47.064528Z digest=sha256:006cd19811cb8569e0a88bcb8a97e8fc915fa621feb864d7e82a53d7ad267f25