Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-22T18:02:23.305313Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 93 of 93 outbound references and 100 inbound Pith citation observations for arXiv:2504.16054.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-22T18:02:23.305313Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T14:42:32.635806Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
93 of 93 outbound references displayed
External citation measurements
0
pith, observed 2026-08-05T02:28:24.338817Z
Observation 3f3e3f5d-95c4-4dba-80a1-873ce12d6d39 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 78d66ea7-8622-4c83-a5a6-a95c2d5e91be · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d45c2814-20a1-40ab-bbf5-4fccddb9e364 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Minivla: A better vla with a smaller footprint
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1222a5a0-d94e-4da3-9bf8-337b50151eb9 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization RT-H: Action Hierarchies Using Language
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 89e3696b-734e-4753-bf03-c3f2e0044103 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization PaliGemma: A versatile 3B VLM for transfer
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 328a572c-4a08-401a-a4f5-d2da11190dcc · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Roboagent: Generalization and efficiency in robot manip- ulation via semantic augmentations and action chunking
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 10b90701-b7af-4b2b-9ebe-eac35d252f1e · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5e556176-8e55-4d4a-865a-26b45d0f82cf · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b1510de5-0c53-457a-8cbe-ebc43ff4fc69 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization RT-1: Robotics Transformer for Real-World Control at Scale
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4f1d3e48-3b07-4bb0-ba34-beeac5598e02 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1b3ceba0-937f-4688-9dce-193331f2c144 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Automating Robot Failure Recovery Using Vision-Language Models With Optimized Prompts
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 75e1d32f-7ffe-4658-b2fc-ef8ea69bf78d · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Microsoft COCO Captions: Data Collection and Evaluation Server
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 64d3b1c8-7ccc-47a6-9849-95ad8e8a46b4 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d5300036-5645-4478-ae0d-b91b5a3f4a5d · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Universal manipulation interface: In- the-wild robot teaching without in-the-wild robots
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d24f5f49-717c-4aa9-8545-1695fc521456 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Racer: Rich language-guided failure recovery policies for imitation learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c6b63635-db63-4cd0-9250-900f99b42ee8 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Robonet: Large-scale multi-robot learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 018921d4-d02f-441f-862b-aefff6bb0a18 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization An unbiased look at datasets for visuo- motor pre-training
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4903c79d-2e7a-4715-ba10-a683ddc771bf · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d2b55c9e-f7d7-4036-804b-37bbcbe7668e · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Reviews-consumer technology
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation fe721ed6-4894-4136-b725-f82d08392f76 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Bert: Pre-training of deep bidirec- tional transformers for language understanding
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5bdf743e-2472-4524-8b28-1ed0ca327ba8 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Scaling cross-embodied learning: One policy for manipulation, navigation, locomotion and aviation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation fc9d53e5-909f-454f-8b1a-a95d15c0e966 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization PaLM-E: An Embodied Multimodal Language Model
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation cb2b7158-1d5a-45ee-b1c1-b61fcee8e625 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Manipulate-Anything: Automating Real-World Robots using Vision-Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 50ce92dc-96f8-4969-851c-c9724a6bcf1b · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Bridge Data: Boosting Generalization of Robotic Skills with Cross-Domain Datasets
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a416a26a-ea1d-41c0-a3d2-31ec3d1d38ab · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization SPOC: Imitating Shortest Paths in Simulation Enables Effective Navigation and Manipulation in the Real World
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation caba3e9e-ae05-49b6-9d2f-545baa5da75c · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Scaling rectified flow transformers for high-resolution image synthesis
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 60b016cb-36e7-43bc-b150-aef360bd979a · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Robot Utility Models: General Policies for Zero-Shot Deployment in New Environments
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7c5746f0-1f9f-4cd6-b3f6-99c53ab6217c · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Anygrasp: Robust and efficient grasp perception in spatial and temporal domains
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 41def0df-85e7-496b-bdca-b1d8718a43fe · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Rh20t: A comprehensive robotic dataset for learning diverse skills in one-shot
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 679b4fb0-974f-476b-ae0c-70c04fc373c0 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Navigating to objects in the real world
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e9b44355-9293-4a62-a9a3-8778dca84f67 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Making the V in VQA matter: Elevating the role of image understanding in visual question answering
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0fb3d65f-0631-40bd-ba3a-172e9030b655 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Robot learning in homes: Improving generalization and reducing dataset bias
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8db603b6-6d48-443b-818d-bdbbf9c95b76 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Deep residual learning for image recognition
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ac1b73af-7c00-4fa6-9631-9c4996e1c3eb · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Masked autoencoders are scalable vision learners
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2f1cd2c7-f65e-4034-bc60-c64512065632 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8c61f841-1fe9-4e09-b874-027d4d476f79 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Otter: A vision-language-action model with text-aware visual feature extraction
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6a2ccd2c-0e9d-4f83-9356-a237e9600aeb · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Language models as zero-shot planners: Extracting actionable knowledge for embodied agents
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 03668529-d3c6-4f04-9439-a58281f01c53 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization OpenAI o1 System Card
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2913e9ca-7bf9-4806-ba41-f2caa95b040e · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Robots at the tipping point: the road to irobot roomba
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 45fad8e0-6958-4439-ad36-0978a75fc72b · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 24e088ea-b9ca-4b63-bfa1-28a8ce078cbd · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization OpenVLA: An Open-Source Vision-Language-Action Model
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation abe1abcd-4e08-4c5e-97c1-6b415b892c55 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Segment Anything
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2c2695fc-1af4-4a13-9d6d-61c7a2daf9ec · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Interactive task planning with language models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9dfaf2b2-2e11-45fb-91d2-6a8693942608 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 729510c9-4d9a-4b2f-adcf-1e77007432be · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization LLaRA: Supercharging Robot Learning Data for Vision-Language Policy
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 01f4aa74-3f39-4664-973e-0c22a3171904 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a4cf5d20-fc53-4f46-9d29-77fbbc64c6c6 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Code as policies: Language model programs for em- bodied control
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c659049c-3b15-4da0-a082-7049d7d19bb9 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Data Scaling Laws in Imitation Learning for Robotic Manipulation
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 70863e49-e320-4d4f-b285-dba65fec6b93 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Flow Matching for Generative Modeling
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 857188e5-2369-4e7e-90b5-0be4799080f7 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Moka: Open-vocabulary robotic manipulation through mark-based visual prompting
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f0ad5648-5195-416f-a61c-ac5f80899851 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 10654187-49b0-4ca9-87d4-c064cd22d9e1 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization OK-Robot: What Really Matters in Integrating Open-Knowledge Models for Robotics
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bb45dc09-c860-4ab2-80b8-b86a53be9b6e · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Rectified Flow: A Marginal Preserving Approach to Optimal Transport
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4fffcbe7-bf8f-45d5-b2a9-eaf52fde9fa1 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 33ff78da-1e73-4c88-9a77-25401e3fa3ad · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Dex-Net 2.0: Deep Learning to Plan Robust Grasps with Synthetic Point Clouds and Analytic Grasp Metrics
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d44ca5df-f078-4a0c-996e-dff2de4529d1 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Where are we in the search for an artificial visual cortex for embodied intelligence? Advances in Neural Information Processing Systems, 36:655–677
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 91f068c7-a5ce-4092-832a-9a0b0e443802 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization R3m: A universal visual representation for robot manipulation
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 753df86c-10a5-4030-87b2-d0e9b3d2b2d4 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization PIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMs
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ef0e990f-e717-404d-9658-4cb6820b5e1a · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Autonomously learn- ing to visually detect where manipulation will succeed
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c6264bba-281b-41fd-b8b8-bb7f49518a39 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0e527388-67c9-4bff-ac4b-1bcc16ebcf55 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Octo: An open-source generalist robot policy
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2ecd5116-f2e7-4cb9-afdb-72ec8e93560a · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Open X-Embodiment: Robotic Learning Datasets and RT-X Models
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b4860028-eeea-43b3-b91e-49406634ace1 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization FAST: Efficient action tok- enization for vision-language-action models
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7c1f1b22-cc76-4d44-ac4e-35010933035c · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Open-vocabulary Mobile Manipulation in Unseen Dynamic Environments with 3D Semantic Maps
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 41548a38-1c30-4e65-a077-217fd00ee66e · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Learning transferable visual models from natural lan- guage supervision
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3a83061e-6c76-489f-8989-d1d738bef472 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization On Bringing Robots Home
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3d6d025c-d723-4c03-ac61-baa72c34fb15 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Gnm: A general navigation model to drive any robot
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e8dedf86-1f6a-465f-977c-2584ea469877 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization ViNT: A Foundation Model for Visual Navigation
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f97ab4cc-78e7-4a97-a4b3-034f09c98d90 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization BUMBLE: Unifying Reasoning and Acting with Vision-Language Models for Building-wide Mobile Manipulation
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4b94d6f4-8050-42ca-a421-12760bde3724 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Yell At Your Robot: Improving On-the-Fly from Language Corrections
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4ca952a4-2f8c-4c18-bd64-911acd1d9eb5 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1705a5b8-196e-46ab-a1ed-5d3bc823352c · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Progprompt: Generating situated robot task plans using large language models
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8a97b39f-b308-47ce-a50e-cf8997c5a9b2 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Open-World Object Manipulation using Pre-trained Vision-Language Models
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a427c21b-8a0f-4bcb-9ca1-c34cf5463855 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ec8058ee-b505-4fbc-aa4c-de5c78b0bc54 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Gemini Robotics: Bringing AI into the Physical World
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4690b45b-0cdf-4d00-b5c2-84c85f709926 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Cambrian-1: A fully open, vision-centric exploration of multimodal llms
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f1fb22f4-1a26-4e4b-a80d-ab3f1c1ff213 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization LLaMA: Open and Efficient Foundation Language Models
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 052b0a18-4771-4866-b28b-8f3c723cceca · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Attention is all you need
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e12937bf-8415-49a7-bbc9-cb24d6cb5fea · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization BridgeData v2: A dataset for robot learning at scale
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f85ce96f-b335-442a-85ab-1699d27a5493 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Llmˆ 3: Large language model-based task and motion planning with motion failure reasoning
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a670d574-06e4-49f3-8ad0-99e646053eec · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Chain-of-thought prompting elicits reasoning in large language models
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c710c50e-2bcc-4709-b055-dcdbd03afaef · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 18f4c3ef-2dc0-42c0-afa1-29edfaefb4a4 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation cf3f0499-aec2-4f22-bca2-0f57d8c90e19 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Masked Visual Pre-training for Motor Control
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation cc67c135-cb35-427a-b9e7-6bacdcf30490 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Magma: A Foundation Model for Multimodal AI Agents
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 72890800-e9f2-4b87-9548-89018afbe567 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Capsfusion: Rethinking image-text data at scale
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 828e06b8-607d-44b8-aecb-d3192bd68c6c · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Robotic control via embodied chain-of-thought reasoning
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 59a2ec82-c40d-4496-b57e-16411d91d8e9 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Cot-vla: Visual chain- of-thought reasoning for vision-language-action models
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2dba29be-db06-4c0a-8c69-5dff826f3490 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c380f0ec-d388-4497-84a5-7521cbedfed3 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Closed-Loop Open-Vocabulary Mobile Manipulation with GPT-4V
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4361a7fb-9b7d-44a6-b812-36d8e3f10e43 · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization put the scissors in the drawer
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4cb6fad6-37d1-4747-a575-fcdc96a4e21a · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization Unresolved cited work
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d4b8922d-a915-44b9-8e76-20c26a58ab8a · outbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization action expert
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1bff4d0c-4c89-49f2-abc0-fa8959b304d5 · inbound
What Matters in Building Vision-Language-Action Models for Generalist Robots $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation db3bf548-5062-47a9-a496-ea1fd0b8bd69 · inbound
DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c289f6b5-265d-4bab-9790-cadc7468ced2 · inbound
GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b67af7ca-89d7-4017-8dcd-a9aed805bb63 · inbound
Seed1.5-VL Technical Report $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6ced7795-e819-4a0d-b6cb-46f522532940 · inbound
DreamGen: Unlocking Generalization in Robot Learning through Video World Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b4e42aa6-311a-4dc5-861e-cf83f6f7fa2d · inbound
DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 857dad4d-0e7d-4cc9-9770-0099d54f9e38 · inbound
Interactive Post-Training for Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 45587987-2e58-4c9b-ba7f-001bd17d22eb · inbound
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e5a96561-63e1-4239-81a0-1a43cb8bac19 · inbound
Real-Time Execution of Action Chunking Flow Policies $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 72dc4682-8985-4280-81fd-d359e9130445 · inbound
V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 101ca8ae-75dc-4e00-ab6f-89e1ace28cb3 · inbound
WorldVLA: Towards Autoregressive Action World Model $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 03d12432-816d-4351-be1d-ffe37b8cf870 · inbound
A Survey on Vision-Language-Action Models: An Action Tokenization Perspective $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 127
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d617f648-b2df-43ba-8faa-50399d09bc84 · inbound
DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8e86f715-3493-4a84-9516-e8def77facdf · inbound
A Careful Examination of Large Behavior Models for Multitask Dexterous Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c250a78b-7295-4aad-a829-6045709cac23 · inbound
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2dea83ce-9844-4114-a00b-0f60d39b7f72 · inbound
Vidar: Embodied Video Diffusion Model for Generalist Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5d20a68c-29ed-4dcd-b978-f74f5fee51da · inbound
GR-3 Technical Report $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e151758a-4ebc-43a5-9c26-778274baeca7 · inbound
Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 204
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 24cbae40-cbb2-42ce-af58-84229a872e20 · inbound
CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cb59a10-a635-46a0-9e1c-776520621a03 · inbound
Prompt-to-Product: Generative Assembly via Bimanual Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a699971-f082-47c4-8c80-c75de34631a8 · inbound
Galaxea Open-World Dataset and G0 Dual-System VLA Model $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2b0527b-41a6-4c37-98c2-d8d7da7e52fe · inbound
Align-Then-stEer: Adapting the Vision-Language Action Models through Unified Latent Guidance $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eee92cf9-71f0-43df-9c11-9aaba1001183 · inbound
TA-VLA: Elucidating the Design Space of Torque-aware Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41f9ae23-47df-4cfa-8699-e0e5f19bc2b6 · inbound
RoboChemist: Long-Horizon and Safety-Compliant Robotic Chemical Experimentation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ccdd464-d875-4dd1-89a7-4b8d3fe74e44 · inbound
A Survey of Reinforcement Learning for Large Reasoning Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 224
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f895124d-e9f9-4b5a-bd36-8e0cde952dff · inbound
SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 317d6c8c-a1f1-45cc-8389-fde0af5ed0f4 · inbound
Ask-to-Clarify: Resolving Instruction Ambiguity through Multi-turn Dialogue $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 127896f9-22f6-4885-8a3f-34a7382a5108 · inbound
Training Agents Inside of Scalable World Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6fed43e7-171c-4653-83b6-24924e3dbcb8 · inbound
A Systematic Study of Large Language Models for Task and Motion Planning With PDDLStream $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4da73bb-5af4-4e59-8536-a7b8193d8678 · inbound
INSIGHT: INference-time Sequence Introspection for Generating Help Triggers in Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f88509f-7fdf-4324-b089-902bbd28b144 · inbound
LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f15583d8-99a2-4077-bb3b-05299ba444b6 · inbound
LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13847b29-9fb7-4c7a-8aea-0ed219205bcd · inbound
GAE: Unleashing Physical Potential of VLM with Generalizable Action Expert $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5c10fdb-c9c9-4880-9576-50d882d82c86 · inbound
Humanoid Everyday: A Comprehensive Robotic Dataset for Open-World Humanoid Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4e1984a-b5e8-4747-85e0-461ca5f17ead · inbound
Ctrl-World: A Controllable Generative World Model for Robot Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7e26644c-764d-4401-a13c-73b1c64c5931 · inbound
DriveVLA-W0: World Models Amplify Data Scaling Law in Autonomous Driving $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2c2d1d15-49a7-4538-98a7-2f3b8694c86e · inbound
QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05942d54-c760-462d-aec7-289e9565d1b2 · inbound
What Questions Should Robots Be Able to Answer? A Dataset of User Questions for Explainable Robotics $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd8a25d4-d20e-4821-b7ed-6d7c424de79e · inbound
RESample: A Robust Data Augmentation Framework via Exploratory Sampling for Robotic Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 70eb54f7-2acf-45d7-aa92-c5deb405cecb · inbound
A Compositional Paradigm for Foundation Models: Towards Smarter Robotic Agents $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d19ea9a-fc35-4144-92dd-8c6d4e76cb84 · inbound
Principled RL for Flow Matching Emerges from the Chunk-level Policy Optimization $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f22d87d0-674b-409d-8883-68ec379d1e4a · inbound
Alpamayo-R1: Bridging Reasoning and Action Prediction for Generalizable Autonomous Driving in the Long Tail $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3220ba0f-cba3-45c1-9046-468f63db3259 · inbound
XR-1: Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8734c05c-1182-44e2-9a35-0fa48370b3ec · inbound
XR-1: Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12199c74-cd1d-47a9-8c57-0bec1d2b0098 · inbound
SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e5343eb-1f60-4702-a82a-269dbe94dcd3 · inbound
ActiveGrasp: Information-Guided Active Grasping with Calibrated Energy-based Model $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ba93a51-b85a-458b-88db-d0aa7b7143c1 · inbound
AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5e9c034e-c047-46c3-a72b-9110abd22f73 · inbound
DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b1fd7e4e-a5f9-4452-8d6e-e485c52fed88 · inbound
Unify Robot Actions in Camera Frame $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2276f108-a3ec-49d7-bcaf-a04833d25917 · inbound
RynnVLA-002: A Unified Vision-Language-Action and World Model $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11066cb0-96ae-43b5-81dc-de0f302b5d77 · inbound
VLASH: Real-Time VLAs via Future-State-Aware Asynchronous Inference $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 959b51c7-0685-4d7b-b8ea-6f7a0fb70425 · inbound
Transport Discrepancy as a Reliability Signal for Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2877a3d8-5e91-4fb3-b1a8-b40025c437f9 · inbound
IGen: Scalable Data Generation for Robot Learning from Open-World Images $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 95cf4eea-589e-48d0-b830-58bc45045b29 · inbound
HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 02e4d1ea-30fb-4843-9d24-274fc9f9f59d · inbound
ImplicitRDP: An End-to-End Visual-Force Diffusion Policy with Structural Slow-Fast Learning $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3856cc85-aa78-4f76-a532-ad5fcb631af5 · inbound
VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b47a868-7136-425e-9eb9-6fb24fb95377 · inbound
Read or Ignore? A Unified Benchmark for Typographic-Attack Robustness and Text Recognition in Vision-Language Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da3fe36e-dd70-4f7a-8681-45877e354712 · inbound
MindDrive: A Vision-Language-Action Model for Autonomous Driving via Online Reinforcement Learning $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d63030b4-20dc-40c1-b027-472928688250 · inbound
mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 49176935-da59-4e4c-a172-d0113ba9cccb · inbound
Learning Semantic Atomic Skills for Multi-Task Robotic Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba07debc-8ef5-439a-8209-c1b2b6bdb2a3 · inbound
StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bc8d88f-6cda-44f0-a3c2-8788ec401683 · inbound
Clutter-Robust Vision-Language-Action Models through Object-Centric and Geometry Grounding $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f633d5bf-2293-42d2-bf5d-ae6e706e68b7 · inbound
Learning to Feel the Future: DreamTacVLA for Contact-Rich Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 32694325-e753-4d5d-9de2-3a90dc290465 · inbound
Learning to Feel the Future: DreamTacVLA for Contact-Rich Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb7779e4-cd41-46bb-84a2-f7208358222d · inbound
Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c554426d-5ef4-47c7-937a-f1f8de4fc835 · inbound
Genie Sim 3.0 : A High-Fidelity Comprehensive Simulation Platform for Humanoid Robot $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dfb145c8-acff-429b-8e15-964018781195 · inbound
Genie Sim 3.0 : A High-Fidelity Comprehensive Simulation Platform for Humanoid Robot $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d3eeb53-4de4-49ff-9163-060000a59307 · inbound
CycleVLA: Proactive Self-Correcting Vision-Language-Action Models via Subtask Backtracking and Minimum Bayes Risk Decoding $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c7bdf54-4323-4445-87c4-d80772fa3baf · inbound
VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3befa379-7fdd-4311-a3ce-4bb9df6d84b2 · inbound
PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 559cc542-074e-403f-bace-651f1620e661 · inbound
CLARE: Continual Learning for Vision-Language-Action Models via Autonomous Adapter Routing and Expansion $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ba848dd2-a97e-4270-b92f-4a85e068cd45 · inbound
SilentDrift: Exploiting Action Chunking for Stealthy Backdoor Attacks on Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2574a6b-878b-45b6-8c80-5736e89514f3 · inbound
Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6f4d1fdb-fa58-41f1-9dde-fb0e10ffac64 · inbound
TouchGuide: Inference-Time Steering of Visuomotor Policies via Touch Guidance $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ff30f187-bf5b-4e2a-b0ea-69750a3f8f9d · inbound
CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation cba26dfb-47e4-4a94-8d5f-35dae4d9bcfe · inbound
PLanAR: Planning-Language-Grounded Agentic Reasoning for Robot Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec841025-6bab-4d75-be70-b62660958626 · inbound
MobileManiBench: Simplifying Model Verification for Mobile Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8efcab4-8756-4f76-a201-a4a26be4ae40 · inbound
Action Hallucination in Generative Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 822fa9cc-27c0-42aa-804c-b12f3cc2c5b0 · inbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f429bd4f-dda0-41d3-87bd-1f735a17ee52 · inbound
Consensus-based optimization (CBO): Towards Global Optimality in Robotics $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbe83c48-fda6-4a53-9f4c-0bd5983d5603 · inbound
VGAS: Value-Guided Action-Chunk Selection for Few-Shot Vision-Language-Action Adaptation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d84481d5-0f37-4a81-bff2-b68a4b21982d · inbound
TwinRL: Digital Twin-Driven Reinforcement Learning for Real-World Robotic Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 453b2280-79e1-4fb6-a597-2683c9e4ee6b · inbound
SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2e14a9a8-b0db-46fd-917c-51ac4d7493bd · inbound
SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ce62710-b8d4-451e-a343-9dc33d815d60 · inbound
Towards Long-Lived Robots: Continual Learning VLA Models via Reinforcement Fine-Tuning $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c06f1fc0-9b34-480a-ad17-ba7a970d86c0 · inbound
AugVLA-3D: Depth-Driven Feature Augmentation for Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 62cccc27-cd54-4ada-aebf-f88eef954206 · inbound
RISE: Self-Improving Robot Policy with Compositional World Model $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 12fc5a08-103d-474b-a7c4-9b4e8fa421b1 · inbound
ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ac8d9279-251b-419a-bea3-ce423503ecbe · inbound
Robot-DIFT: Correspondence-Sensitive Diffusion Features for Contact-Rich Robot Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fdfede4-577f-473e-b6bb-34c1cfcf2ae9 · inbound
LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfe368ce-6ede-4d57-a0e1-890224b89277 · inbound
Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33274cb5-09d9-4684-af41-3ef5e931b8f8 · inbound
WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3859444-12dd-4cc5-8c70-f4e5eab3d792 · inbound
World Action Models are Zero-shot Policies $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0a3d9840-ce2f-4f5c-9780-6474f13409af · inbound
When Vision Overrides Language: Evaluating and Mitigating Counterfactual Failures in VLAs $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92826e6a-2113-45d5-840c-7794690bd75e · inbound
Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Efficient Robotic Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4c857a25-5d50-4912-b536-efb2ad166f3f · inbound
QuantVLA: Scale-Calibrated Post-Training Quantization for Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9733372d-3292-4d28-92cd-2715eac8a56c · inbound
Notes-to-Self: Scratchpad Augmented VLAs for Memory Dependent Manipulation Tasks $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53349c06-8d18-40a0-9b86-0ec37d997b72 · inbound
CableRobotGraphSim: A Graph Neural Network for Modeling Partially Observable Cable-Driven Robot Dynamics $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd2cf00a-f745-4b5c-a5a1-479e0e8936f7 · inbound
Unleashing the Potential of Diffusion Models for End-to-End Autonomous Driving $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a842e040-2780-437b-9a7c-ac5160d82f15 · inbound
InCoM: Intent-Driven Perception and Structured Coordination for Mobile Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.