Pith. sign in

Skillvla: Tackling combinatorial diversity in dual-arm manipulation via skill reuse

4 Pith papers cite this work. Polarity classification is still indexing.

4 Pith papers citing it

citation-role summary

background 2

citation-polarity summary

years

2026 4

roles

background 2

polarities

background 2

representative citing papers

Skill Reuse as Compression in Agentic RL

cs.LG · 2026-05-29 · unverdicted · novelty 5.0

ReuseRL augments agentic RL with an MDL-based compression penalty on skill reuse, proves a PAC-Bayes bound, and reports higher in- and out-of-distribution success on ALFWorld, TextWorld-Cooking, and Countdown-Stepwise versus GRPO and round-length baselines.

Code as Agent Harness

cs.CL · 2026-05-18 · accept · novelty 5.0

A survey that organizes existing work on LLM-based agents around code as the central harness, structured in three layers of interfaces, mechanisms, and multi-agent scaling, with applications across domains and listed open challenges.

citing papers explorer

Showing 4 of 4 citing papers.

  • See Selectively, Act Adaptively: Dual-Level Structural Decomposition for Bimanual Robot Manipulation cs.RO · 2026-06-11 · unverdicted · none · ref 26

    A VLA policy using view-selective visual routing and interaction-aware action MoE improves average success by 27.7% in simulation and 43.3% in real-world bimanual tasks over monolithic baselines.

  • TAMEn: Tactile-Aware Manipulation Engine for Closed-Loop Data Collection in Contact-Rich Tasks cs.RO · 2026-04-08 · unverdicted · none · ref 5

    TAMEn supplies a cross-morphology wearable interface and pyramid-structured visuo-tactile data regime that raises bimanual manipulation success rates from 34% to 75% via closed-loop collection.

  • Skill Reuse as Compression in Agentic RL cs.LG · 2026-05-29 · unverdicted · none · ref 6

    ReuseRL augments agentic RL with an MDL-based compression penalty on skill reuse, proves a PAC-Bayes bound, and reports higher in- and out-of-distribution success on ALFWorld, TextWorld-Cooking, and Countdown-Stepwise versus GRPO and round-length baselines.

  • Code as Agent Harness cs.CL · 2026-05-18 · accept · none · ref 112

    A survey that organizes existing work on LLM-based agents around code as the central harness, structured in three layers of interfaces, mechanisms, and multi-agent scaling, with applications across domains and listed open challenges.