Embodied Interpretability: Linking Causal Understanding to Generalization in Vision-Language-Action Models

· 2026 · cs.RO · arXiv 2605.00321

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

open full Pith review browse 1 citing papers arXiv PDF

abstract

Vision-Language-Action (VLA) policies often fail under distribution shift, suggesting that decisions may depend on spurious visual correlations rather than task-relevant causes. We formulate visual-action attribution as an interventional estimation problem. Accordingly, we introduce the Interventional Significance Score (ISS), an interventional masking procedure for estimating the causal influence of visual regions on action predictions, and the Nuisance Mass Ratio (NMR), a scalar measure of attribution to task-irrelevant features. We analyze the statistical properties of ISS and show that it admits unbiased estimation, and we characterize conditions under which action prediction error provides a valid proxy for causal influence. Experiments across diverse manipulation tasks indicate that NMR predicts generalization behavior and that ISS yields more faithful explanations than existing interpretability methods. These results suggest that interventional attribution provides a simple diagnostic approach for identifying causal misalignment in embodied policies.

representative citing papers

Bridge-WA: Predicting Where and How the World Changes for Robotic Action

cs.RO · 2026-07-02 · unverdicted · novelty 4.0

Bridge-WA introduces a lightweight distillation-based world-action model that uses future-change priors to improve robotic task success and robustness without deployment-time dense rollouts.

citing papers explorer

Showing 1 of 1 citing paper after filters.

Bridge-WA: Predicting Where and How the World Changes for Robotic Action cs.RO · 2026-07-02 · unverdicted · none · ref 49 · internal anchor
Bridge-WA introduces a lightweight distillation-based world-action model that uses future-change priors to improve robotic task success and robustness without deployment-time dense rollouts.

Embodied Interpretability: Linking Causal Understanding to Generalization in Vision-Language-Action Models

fields

years

verdicts

representative citing papers

citing papers explorer