Pith. sign in

Title resolution pending

5 Pith papers cite this work. Polarity classification is still indexing.

5 Pith papers citing it

fields

cs.RO 4 cs.LG 1

years

2026 4 2024 1

verdicts

UNVERDICTED 5

representative citing papers

Diffusion Policy Policy Optimization

cs.RO · 2024-09-01 · unverdicted · novelty 6.0

DPPO fine-tunes diffusion policies via policy gradients and outperforms prior RL approaches for diffusion policies and PG-tuned alternatives on robot benchmarks while enabling stable training and hardware deployment.

TacCoRL: Integrating Tactile Feedback into VLA via Simulation

cs.RO · 2026-06-10 · unverdicted · novelty 5.0

TacCoRL integrates tactile feedback into VLA policies via real-aligned simulation co-training and RL, raising average success from 50% to 72.5% on four bimanual contact-rich tasks with direct real-robot transfer.

An Agency-Transferring Model-Free Policy Enhancement Technique

cs.LG · 2026-06-08 · unverdicted · novelty 5.0

A model-free RL method arbitrates between a functional baseline policy and a learning policy, transferring agency over time to yield a standalone policy with high goal-reaching rates and competitive returns on continuous-control tasks.

citing papers explorer

Showing 5 of 5 citing papers.

  • DexCompose: Reusing Dexterous Policies for Multi-Task Manipulation with a Single Hand cs.RO · 2026-06-26 · unverdicted · none · ref 36

    DexCompose achieves 77.4% average success on 16 composite dexterous tasks by using role-aware residual composition with explicit finger ownership to combine pretrained policies without destructive interference.

  • AnyBody: Free-Form Whole-Body Humanoid Control from Arbitrary Keypoint Guidance cs.RO · 2026-06-28 · unverdicted · none · ref 42

    AnyBody distills a privileged teacher tracker into a latent unit-sphere representation and uses a masked transformer to drive humanoid control from arbitrary keypoint subsets.

  • Diffusion Policy Policy Optimization cs.RO · 2024-09-01 · unverdicted · none · ref 3

    DPPO fine-tunes diffusion policies via policy gradients and outperforms prior RL approaches for diffusion policies and PG-tuned alternatives on robot benchmarks while enabling stable training and hardware deployment.

  • TacCoRL: Integrating Tactile Feedback into VLA via Simulation cs.RO · 2026-06-10 · unverdicted · none · ref 48

    TacCoRL integrates tactile feedback into VLA policies via real-aligned simulation co-training and RL, raising average success from 50% to 72.5% on four bimanual contact-rich tasks with direct real-robot transfer.

  • An Agency-Transferring Model-Free Policy Enhancement Technique cs.LG · 2026-06-08 · unverdicted · none · ref 15

    A model-free RL method arbitrates between a functional baseline policy and a learning policy, transferring agency over time to yield a standalone policy with high goal-reaching rates and competitive returns on continuous-control tasks.