Pith. sign in

Paper Citation Record · LEDGER

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills

As of 6 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2607.22529.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.22529 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T04:28:49.464315Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0aa6daad-b7c5-4289-8cdf-61e05fa1dbdc · outbound

This paper cites Ultraif: Advancing instruction following from the wild.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Ultraif: Advancing instruction following from the wild

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:44.927193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:44.927193Z digest=sha256:bc8fee1c3b44a8f521c389c9941a0387784f89b4fce07438432d6b0dec545a2d

Observation a535cdee-e078-49be-b9bc-64cd775d6504 · outbound

This paper cites SkillJuror: Measuring How Agent Skill Organization Changes Runtime Behavior.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills SkillJuror: Measuring How Agent Skill Organization Changes Runtime Behavior

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:45.144472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:45.144472Z digest=sha256:077391e32fe5695248d4882212440737a461e8ef516db4343063fef4500323f1

Observation c82e338f-656d-468f-ac41-3df830a9e4b3 · outbound

This paper cites I want to plan a dinner for my friends this weekend in Portland. We’re looking for a nice Italian place that doesn’t cost more than $30 per person.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills I want to plan a dinner for my friends this weekend in Portland. We’re looking for a nice Italian place that doesn’t cost more than $30 per person

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:49.404490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:49.404490Z digest=sha256:83a1c859ee65886c14d91d39a81715f3db2e785c9bc81949d6824360d174f2e8

Observation 3b8b6ba3-07ec-41c1-8899-32f9e7845c05 · outbound

This paper cites From self-evolving synthetic data to verifiable-reward rl: Post-training multi-turn interactive tool-using agents.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills From self-evolving synthetic data to verifiable-reward rl: Post-training multi-turn interactive tool-using agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:45.458502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:45.458502Z digest=sha256:645697f7d0a801465bfbd8bb27612e6caed837dffa32aa51f0cc329c9bddb836

Observation bba3d182-65cb-4f0c-a6a5-429435763b84 · outbound

This paper cites RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:45.552110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:45.552110Z digest=sha256:9ef94627e9b1131807afe5ccf15d2f647eae0c06dd29a58774d331094a247ce3

Observation 2efc86d2-a72e-4ffe-a41c-9dce29985975 · outbound

This paper cites Gems: Agent-native multimodal generation with memory and skills.arXiv preprint arXiv:2603.28088,.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Gems: Agent-native multimodal generation with memory and skills.arXiv preprint arXiv:2603.28088,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:45.639596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:45.639596Z digest=sha256:86a7e75d4bf2a8809d5664ca770ec2a299604ffdac98f09508e72e3bb7c9883d

Observation d0f755cd-4170-43dd-ac7a-4442881dbc90 · outbound

This paper cites R-Zero: Self-Evolving Reasoning LLM from Zero Data.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills R-Zero: Self-Evolving Reasoning LLM from Zero Data

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:45.732564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:45.732564Z digest=sha256:cd13ade0eb154fcabf97ed3ea2d733431f20cf558a807156efa20a8fe5dae87a

Observation 1187567d-5631-46c9-8bae-981a4ff62159 · outbound

This paper cites Bowen Jiang, Taiwei Shi, Ryo Kamoi, Yuan Yuan, Camillo J Taylor, Longqi Yang, Pei Zhou, and Sihao Chen.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Bowen Jiang, Taiwei Shi, Ryo Kamoi, Yuan Yuan, Camillo J Taylor, Longqi Yang, Pei Zhou, and Sihao Chen

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:45.841654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:45.841654Z digest=sha256:9145c5a341d8e59058e2038b6595776d2f88307f08164f796a3cd28bc9603131

Observation 64db138e-10e3-4c85-9ef0-86e5ca864eb2 · outbound

This paper cites DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:45.954722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:45.954722Z digest=sha256:0083cbcdc4c84f53f90ebcf796d4fd2c769c3d39af1910e8d683bee62630d794

Observation 434e22b1-81ff-4212-a799-f31a0f7c68af · outbound

This paper cites R-diverse: Mitigating diversity illusion in self-play llm training.arXiv preprint arXiv:2602.13103, 2026a.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills R-diverse: Mitigating diversity illusion in self-play llm training.arXiv preprint arXiv:2602.13103, 2026a

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:46.184528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:46.184528Z digest=sha256:f07662f248ccd1b918623df81f434dd83268ff1cf04b638bed6ae83ca15d68b4

Observation bde93b8d-d7c9-49f3-b762-451245bfa63e · outbound

This paper cites SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:46.307495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:46.307495Z digest=sha256:0b4de8caa5ad8ddcba0feb9b8afb2235bc9c37c23d0cc2b08c6c284ef1e23ef8

Observation e96f8b1d-10c6-4b08-bdf5-ad7be70090cf · outbound

This paper cites Agent skills: A data-driven analysis of claude skills for extending large language model functionality.arXiv preprint arXiv:2602.08004,.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Agent skills: A data-driven analysis of claude skills for extending large language model functionality.arXiv preprint arXiv:2602.08004,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:46.439202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:46.439202Z digest=sha256:ae99893e90c0a8755de2ff3cfe0569e3551e436ba0efbe571c013799f3503c6c

Observation 0f867384-7fb6-4afc-a086-205fb99b4c50 · outbound

This paper cites Ministral 3.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Ministral 3

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:46.545451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:46.545451Z digest=sha256:dbb63bdfc75ef781f7bacee6a78d6296fb61bc2faf910be6ce8fec4f02afc5d6

Observation 9521f4d7-2342-4664-8f4d-1a88f3503ef3 · outbound

This paper cites SkillRevise: Improving LLM-Authored Agent Skills via Trace-Conditioned Skill Revision.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills SkillRevise: Improving LLM-Authored Agent Skills via Trace-Conditioned Skill Revision

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:46.657137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:46.657137Z digest=sha256:a67e45006144e9498c9d7d0436cb18e235a9e95319d8970e15f11353c4419c04

Observation e0bef9a0-5c8e-4d82-9306-3bcf9227c841 · outbound

This paper cites Search Self-play: Pushing the Frontier of Agent Capability without Supervision.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Search Self-play: Pushing the Frontier of Agent Capability without Supervision

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:46.810522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:46.810522Z digest=sha256:1fe0ef652e4d44e40db53824e26b2ebf85fcd5d43f94cea6fbccc819dc746f34

Observation d173e1cc-a592-4c26-a775-d1566ce09ab8 · outbound

This paper cites SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:46.952475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:46.952475Z digest=sha256:d535c0e0af41f9478485ef07c4319f6717bdf5141aa31bf0897b59f20b64d1bf

Observation 4d74f686-e46e-4e04-80d0-371b74be435c · outbound

This paper cites Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:47.064528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:47.064528Z digest=sha256:006cd19811cb8569e0a88bcb8a97e8fc915fa621feb864d7e82a53d7ad267f25

Observation 91c76fc2-5099-44d7-848c-2a4df08b4823 · outbound

This paper cites Better alignment with instruction back-and-forth translation.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Better alignment with instruction back-and-forth translation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:47.217091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:47.217091Z digest=sha256:9ee6a2252df923acf6d6fb94b88315b62d2a18cc807183ea26143df8a79073b7

Observation c88c5bb2-7079-4a47-a693-2a4933a77eea · outbound

This paper cites Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:47.335859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:47.335859Z digest=sha256:b04b12890c808343a390826fd24062b854be4aab3bdfd11ab0d2e589c2b11aee

Observation 15964e59-f02a-42c4-85d8-0bad85c80133 · outbound

This paper cites Large Language Models Can Self-Improve At Web Agent Tasks.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Large Language Models Can Self-Improve At Web Agent Tasks

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:47.436292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:47.436292Z digest=sha256:151949cd9f8041aba63025638b65430494bfc1f0d7754362621de3357d0812de

Observation 9f32634f-2afc-489b-866f-bf837084d20a · outbound

This paper cites Infobench: Evaluating instruction following ability in large language models.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Infobench: Evaluating instruction following ability in large language models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:47.552765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:47.552765Z digest=sha256:d9b4e6f47f483c3858c7343820f8271a4ef2e2b75e0fb1b348e7de77ffe1ba93

Observation 3f8aee4f-fa73-4196-8831-2197fb3ef672 · outbound

This paper cites Sentence-bert: Sentence embeddings using siamese bert-networks.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Sentence-bert: Sentence embeddings using siamese bert-networks

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:47.668921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:47.668921Z digest=sha256:a4b45de153f6f87ec8d386fea71e4f7433f32fab2212c9eaef5488c537c5410c

Observation 8ef2ee6f-18ea-4836-bf27-e9b421aa3821 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:47.878934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:47.878934Z digest=sha256:244d444e23a670b08f93cbde6479ba6d120ada7f61d33b4e2b733578f65cc12a

Observation 9d9e1a79-94e1-4133-bf5f-48e89bcb0001 · outbound

This paper cites SKILLFOUNDRY: Building Self-Evolving Agent Skill Libraries from Heterogeneous Scientific Resources.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills SKILLFOUNDRY: Building Self-Evolving Agent Skill Libraries from Heterogeneous Scientific Resources

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:48.022530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:48.022530Z digest=sha256:ac856391ec3867f011a8fef4fbb856e92aad604a208265a40724c26e158d28cd

Observation 01050069-831c-4d3c-bc99-413d35f3e41b · outbound

This paper cites A Survey on Self-Evolution of Large Language Models.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills A Survey on Self-Evolution of Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:48.253415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:48.253415Z digest=sha256:8b27c650ad83fb5ef721b5ad0c21fc3d276114a020047288303edb23f0710144

Observation 89ede134-3d16-42ad-ada1-d4ac6ac62866 · outbound

This paper cites Code-a1: Adversarial evolving of code llm and test llm via reinforcement learning.arXiv preprint arXiv:2603.15611, 2026a.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Code-a1: Adversarial evolving of code llm and test llm via reinforcement learning.arXiv preprint arXiv:2603.15611, 2026a

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:48.304531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:48.304531Z digest=sha256:fa568f4a18b07d9384ccc4abd7a3ef88ff38ee4af7fec01794c244e6545b9f3c

Observation 36cfbba1-346e-4c22-a86c-f7d320e895ff · outbound

This paper cites Toward Training Superintelligent Software Agents through Self-Play SWE-RL.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Toward Training Superintelligent Software Agents through Self-Play SWE-RL

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:48.430179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:48.430179Z digest=sha256:11b06ab39da513dc95304f15eec25b4fd52e6e769d65095ffef1d5f181ee7496

Observation d0351b3c-df92-4b77-b934-d95962ece4e1 · outbound

This paper cites Propose, solve, verify: Self-play through formal verification.arXiv preprint arXiv:2512.18160,.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Propose, solve, verify: Self-play through formal verification.arXiv preprint arXiv:2512.18160,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:48.564168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:48.564168Z digest=sha256:db0ac63260c050e73e2d8a3f4cf01a3a9c22470245e30dba73ff4a74e9d742cb

Observation 6c223b5c-34cd-47a6-9283-f3f904b07b23 · outbound

This paper cites Wizardlm: Empowering large pre-trained language models to follow complex instructions.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Wizardlm: Empowering large pre-trained language models to follow complex instructions

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:48.732010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:48.732010Z digest=sha256:dc8b0897668f32f0dceefa2cc1277a1abfdc888684eca273ba6e4c50ae738f06

Observation 7d81a9a8-731e-4485-b7f3-514beb42fd7a · outbound

This paper cites OpenSkill: Open-World Self-Evolution for LLM Agents.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills OpenSkill: Open-World Self-Evolution for LLM Agents

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:48.825254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:48.825254Z digest=sha256:36bb370c778caa31c094138d63ee4ddd4eb0cb0d4f09b1fc1677aa93be04a2ce

Observation 1f323bd6-937c-496f-bcba-61d5d76ad2fc · outbound

This paper cites Qwen3 Technical Report.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Qwen3 Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:48.911778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:48.911778Z digest=sha256:399316322125cc4aae2fa84ef1e88a028565754f42c1c05e17d214918be38eb6

Observation 0e14d102-216a-41ad-a1c2-4695e8cd40b1 · outbound

This paper cites CoT-Self-Instruct: Building high-quality synthetic prompts for reasoning and non-reasoning tasks.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills CoT-Self-Instruct: Building high-quality synthetic prompts for reasoning and non-reasoning tasks

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:48.960183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:48.960183Z digest=sha256:f431d3c34201467d294e278b75f7d16d32f1cf53e4054fbc6047e558ad1328e7

Observation 35acc60b-73d9-4046-8363-e7e252ca2fd6 · outbound

This paper cites Learning to Reason without External Rewards.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Learning to Reason without External Rewards

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:49.009194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:49.009194Z digest=sha256:4992708723cc060d65ae82be862d014ce9d9d8517e634e7fb447ec7c60e7f7f1

Observation f798bf16-0179-4e87-afc2-09e4b6e5797d · outbound

This paper cites SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:49.067525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:49.067525Z digest=sha256:e786755916bc80954a58e9e5a2ddd21bd7e10f0024c4db65665bc5dde7ef3601

Observation 6cf073f6-9c81-4cd5-8578-a9cf77f80205 · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Instruction-Following Evaluation for Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:49.133550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:49.133550Z digest=sha256:2060c18e38e9f3e5f100e67728fe2c21e331442e6d9cfc4f14102bb0774bb1fe

Observation bc7c0064-307a-4281-b5c6-0dc9d1a4f021 · outbound

This paper cites Webarena: A realistic web environment for building autonomous agents.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Webarena: A realistic web environment for building autonomous agents

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:49.194675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:49.194675Z digest=sha256:0e50080433cc8b00a6c9f9836d8c0e8b8897132d166ba85c6449c7cfe49d8e6c

Observation defb9909-59c5-48be-a39f-08b213d06789 · outbound

This paper cites SkillGenBench: Benchmarking Skill Generation Pipelines for LLM Agents.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills SkillGenBench: Benchmarking Skill Generation Pipelines for LLM Agents

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:49.274654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:49.274654Z digest=sha256:330d0649bc500e5b9c431cd8c50485b159f15c74e305991f8fba6a32baa0f1a2

Observation 36b44770-e84a-47f1-98f5-5dafee72db1f · outbound

This paper cites project updates.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills project updates

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:49.329638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:49.329638Z digest=sha256:09afe6c0f6add856660d5ae1e35b2ff5f335bd1151fc4cfdd4fee99743211fc5

Observation f2175526-f46e-4dc3-90fa-b52d953befb4 · outbound

This paper cites schedule a meeting.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills schedule a meeting

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:49.464315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:49.464315Z digest=sha256:174c16a711bf757d9b8cca895481c4fb8098031d7a67d82e52e260bd02519607

Observation 3d419ccf-b96f-4864-bf73-32ada2d1824f · outbound

This paper cites EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:48.135993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:48.135993Z digest=sha256:5015506801e9473befbb300301b5da9dc4d5796f651233c83334dbb084be7168

Observation 3b32beb6-35e5-4885-bc22-505a1fea3841 · outbound

This paper cites Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:47.805036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:47.805036Z digest=sha256:5356b6331c9dd0ded58481cb6abb62214399b115f9de185d80084264a577ff77

Observation 27ae9a1c-8fde-48eb-a40e-cb7097003ffc · outbound

This paper cites Language self-play for data-free training.arXiv preprint arXiv:2509.07414,.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Language self-play for data-free training.arXiv preprint arXiv:2509.07414,

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:46.086644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:46.086644Z digest=sha256:fdf535fc4fb958d199051387e8d10cd4dbec5fd41374ddfedb134a1e3a76692b

Observation 064ca75f-2f59-4c0e-88c8-753332ffd3d5 · outbound

This paper cites SPaR: Self-Play with Tree-Search Refinement to Improve Instruction-Following in Large Language Models.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills SPaR: Self-Play with Tree-Search Refinement to Improve Instruction-Following in Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:45.255280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:45.255280Z digest=sha256:d9cceeb971827af698e8f45a521ce53e34d72691c15e21b5feb93dc86d8c28cd

Observation dd69018d-9cdf-4f1b-b65d-2d2cc94df016 · outbound

This paper cites Skill-RM: Unifying Heterogeneous Evaluation Criteria via Agent Skill.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Skill-RM: Unifying Heterogeneous Evaluation Criteria via Agent Skill

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:45.019999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:45.019999Z digest=sha256:2e7d2f5c13094e570e02bdf55bc9713a75fa1fd9263858742667261a85859364

Observation 232f5c1d-59d6-4eaa-a412-8853239a5168 · outbound

This paper cites Self- play with execution feedback: Improving instruction-following capabilities of large language models.

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Self- play with execution feedback: Improving instruction-following capabilities of large language models

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-01T04:28:45.367391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:28:45.367391Z digest=sha256:f20a6cd732856fd79ffca16f8bcfeadee403d50c5e2693af284fb6380ec590d6

Pith citing papers

No inbound Pith citation observations are available.