Pith. sign in

From foresight to forethought: Vlm-in-the-loop policy steering via latent alignment

9 Pith papers cite this work. Polarity classification is still indexing.

9 Pith papers citing it

fields

cs.RO 8 cs.CV 1

years

2026 9

verdicts

UNVERDICTED 9

representative citing papers

Robot Critics that Sweat the Small Stuff

cs.RO · 2026-06-19 · unverdicted · novelty 6.0

Fine-tuning VLMs with pairwise progress supervision from policy rollouts improves fine-grained failure detection and boosts robot manipulation success by 11% real-world and 5.9% in simulation.

DREAM-Chunk: Reactive Action Chunking with Latent World Model

cs.RO · 2026-06-17 · unverdicted · novelty 6.0

DREAM-Chunk uses test-time sampling and latent-world-model rollouts to select robust action chunks from chunking-based VLA policies, improving performance under stochastic dynamics on simulation and hardware tasks.

Position: Good Embodied Reward Models Need Bad Behavior Data

cs.RO · 2026-05-31 · unverdicted · novelty 4.0

Embodied reward models systematically over-reward unsafe, suboptimal, and shortcut robot behaviors due to training on successful data only, and modest inclusion of bad behavior data improves alignment with human preferences.

citing papers explorer

Showing 9 of 9 citing papers.