Pith. sign in

hub Canonical reference

Spatial forcing: Implicit spatial representation alignment for vision- language-action model

Canonical reference. 71% of citing Pith papers cite this work as background.

34 Pith papers citing it
Background 71% of classified citations

hub tools

citation-role summary

background 5 baseline 1 method 1

citation-polarity summary

years

2026 33 2025 1

representative citing papers

Geometric Action Model for Robot Policy Learning

cs.RO · 2026-06-15 · unverdicted · novelty 6.0

GAM splits a geometric foundation model to enable language-conditioned future geometry prediction and action decoding for robot policies, claiming superior performance on manipulation benchmarks.

A Pragmatic VLA Foundation Model

cs.RO · 2026-01-26 · unverdicted · novelty 6.0

LingBot-VLA is a VLA foundation model trained on massive real robot data that shows superior generalization across tasks and platforms with fast training throughput.

Learning Action Priors for Cross-embodiment Robot Manipulation

cs.RO · 2026-06-24 · unverdicted · novelty 5.0

A two-stage framework pretrains an action module with temporal motion priors from unconditioned trajectories using flow-matching, then transfers it to VLA training via decoder reuse and distillation, yielding better performance on cross-embodiment tasks.

citing papers explorer

Showing 34 of 34 citing papers.