Pith. sign in

REVIEW 4 cited by

Dense Policy: Bidirectional Autoregressive Learning of Actions

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2503.13217 v1 pith:W6VEZHNF submitted 2025-03-17 cs.RO cs.CVcs.LG

classification cs.ROcs.CVcs.LG
keywords autoregressivepolicieslearningpolicyactiondensegenerativeholistic
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Mainstream visuomotor policies predominantly rely on generative models for holistic action prediction, while current autoregressive policies, predicting the next token or chunk, have shown suboptimal results. This motivates a search for more effective learning methods to unleash the potential of autoregressive policies for robotic manipulation. This paper introduces a bidirectionally expanded learning approach, termed Dense Policy, to establish a new paradigm for autoregressive policies in action prediction. It employs a lightweight encoder-only architecture to iteratively unfold the action sequence from an initial single frame into the target sequence in a coarse-to-fine manner with logarithmic-time inference. Extensive experiments validate that our dense policy has superior autoregressive learning capabilities and can surpass existing holistic generative policies. Our policy, example data, and training code will be publicly available upon publication. Project page: https: //selen-suyue.github.io/DspNet/.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. H$^3$DP: Triply-Hierarchical Diffusion Policy for Visuomotor Learning

    cs.RO 2025-05 conditional novelty 6.0 of 10

    H3DP couples depth-layered, multi-scale visual features to coarse-to-fine denoising stages, reporting a +27.5% relative success-rate improvement over DP3 across 44 simulation tasks.

  2. Motion Before Action: Diffusing Object Motion as Manipulation Condition

    cs.RO 2024-11 conditional novelty 6.0 of 10

    MBA is a plug-in module that generates robot actions by first diffusing predicted object pose sequences, then conditioning action diffusion on those poses, improving imitation-learning success rates across simulation ...

  3. RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design

    cs.RO 2026-03 conditional novelty 5.0 of 10

    A benchmark and modular policy show that explicit memory components substantially improve robotic manipulation on tasks requiring recall of past observations.

  4. OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation

    cs.RO 2025-05 conditional novelty 5.0 of 10

    OpenHelix shows that a frozen vision-language model with a prompt-tuned token and an auxiliary action-prediction head beats full fine-tuning on CALVIN language generalization while training far fewer parameters.

Pith tools