Pith. sign in

hub Mixed citations

Vila: On pre-training for visual language models

Mixed citation behavior. Most common role is background (60%).

19 Pith papers citing it
Background 60% of classified citations

hub tools

citation-role summary

background 3 baseline 2

citation-polarity summary

representative citing papers

OpenVLA: An Open-Source Vision-Language-Action Model

cs.RO · 2024-06-13 · unverdicted · novelty 6.0

OpenVLA achieves 16.5% higher task success than the 55B RT-2-X model across 29 tasks with 7x fewer parameters while enabling effective fine-tuning and quantization without performance loss.

What Limits Vision-and-Language Navigation ?

cs.RO · 2026-05-13 · unverdicted · novelty 5.0

StereoNav reaches new benchmark highs on R2R-CE and RxR-CE and improves real-robot reliability by supplying persistent target-location priors and stereo-derived geometry that stay stable under lighting changes and blur.

citing papers explorer

Showing 19 of 19 citing papers.