Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-13T18:18:27.211034Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 100 inbound Pith citation observations for arXiv:2403.09631.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-13T18:18:27.211034Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T10:08:58.446528Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
62 of 62 outbound references displayed
External citation measurements
14
pith, observed 2026-08-05T02:28:24.338817Z
Observation 97c26ff7-2daf-42b0-9f32-c61a0597ce06 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Flamingo: a visual language model for few-shot learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 54c519dd-d490-47a3-84ab-5477256c9f96 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ecbe30c9-510a-4a0b-91f1-8e058c01ba94 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Zero-shot robotic manipulation with pretrained image-editing diffusion models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 26525cfe-1dc5-4f41-9bed-18f1fcf5a0e2 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model RT-1: Robotics Transformer for Real-World Control at Scale
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation fb41528e-1b26-46b0-ae0e-888efea28365 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0f4e72b5-d1fc-4267-b9e7-f2ecc317f2b0 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f82a101a-3f98-4ca4-9c88-9ec18aa98a34 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Playfusion: Skill acquisition via diffusion from language-annotated play
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 94c34a7a-f505-4621-a41e-235d45706b52 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Ll3da: Visual interactive instruction tuning for omni-3d understanding, reasoning, and planning, 2023 b
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4ec0708d-5dad-472e-beaf-36a3579bf0aa · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model X., Savva, M., Halber, M., Funkhouser, T., and Nießner, M
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2acc2055-fb9e-4b7f-8f8e-6bdad4863da6 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model M., Fidler, S., Furnari, A., Kazakos, E., Moltisanti, D., Munro, J., Perrett, T., Price, W., et al
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e56ab7a1-7359-4dc6-ad3d-3cb9d452bd56 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9696f091-f581-41ee-ac2f-5be77b0e1e27 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Objaverse: A universe of annotated 3d objects
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b84cd66f-e925-47e8-99c8-681c07eb4b41 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model DreamLLM: Synergistic Multimodal Comprehension and Creation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 79465bc0-9128-4873-b79b-f6626afb40a9 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model PaLM-E: An Embodied Multimodal Language Model
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 28a5589f-a5a9-4c7a-870c-92474b4b305c · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a1e5f78b-b147-426a-a44b-18ff132741de · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Structure and content-guided video synthesis with diffusion models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2ebb3c51-960b-47df-8712-0be93cd09567 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model RH20T: A Comprehensive Robotic Dataset for Learning Diverse Skills in One-Shot
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9372c22d-f714-43ae-b152-3b0be8d7a963 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Finetuning Offline World Models in the Real World
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation daef5360-dc86-4303-9daf-02ee05aec76d · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Point-bind & point-llm: Aligning point cloud with multi-modality for 3d understanding, generation, and instruction following
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6c53d0dd-a18c-4c88-9f4c-840e83f7097a · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model 3D-LLM: Injecting the 3D World into Large Language Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 91595000-032b-4048-ab83-af3e71f081f2 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model MultiPLY: A Multisensory Object-Centric Embodied Large Language Model in 3D World
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6bb457c2-2154-4c7a-be26-4f239908b6ec · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model and Montani, I
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 713f92ed-29cc-46f3-af21-25a6bee88c14 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model LoRA: Low-Rank Adaptation of Large Language Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3468c4cc-5af2-4be1-aead-87c78eb6d5ef · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Chat-3d v2: Bridging 3d scene and large language models with object identifiers, 2023 a
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation fe44a321-56a2-416f-8c0e-a8cbd7e22ba0 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model An Embodied Generalist Agent in 3D World
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 413872a8-85d0-4228-b6ab-d2a98e89f293 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Language Is Not All You Need: Aligning Perception with Language Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6dd03766-d22d-41dc-b1cf-64b51615841e · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model R., and Davison, A
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 628055ba-4871-4010-9d8b-5a332ecff3d7 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Bc-z: Zero-shot task generalization with robotic imitation learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a379add8-27ec-4f8e-89cb-b7ef2806394d · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Auto-Encoding Variational Bayes
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 715ce97a-70ca-4f6f-9b17-991f49e24ccb · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ab28b99f-416d-4fc3-9526-e3528ecc82bd · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model CoVLM: Composing Visual Entities and Relationships in Large Language Models Via Communicative Decoding
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5981af40-678c-4683-9992-f97fd2844b89 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3b80b2e0-cc5a-429b-b8a9-1f735fda131a · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model 3dmit: 3d multi-modal instruction tuning for scene understanding
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 78a9492e-e41a-4cf4-bb1c-368f1dfb832c · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Visual Instruction Tuning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 13a4c3ac-e92d-4788-b118-19e30c4b12ca · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Hoi4d: A 4d egocentric dataset for category-level human-object interaction
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9e6ad27d-fc14-4499-8d7f-3e24c337c29a · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model UNIFIED - IO : A unified model for vision, language, and multi-modal tasks
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1ce1ef94-0f93-4ec5-99a0-c2f9185fd7b6 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Language Conditioned Imitation Learning over Unstructured Data
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0f286e5d-f3ce-4600-92a8-f1d175c6dac4 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Interactive language: Talking to robots in real time
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 07e492e3-b632-43ce-a334-cc76f3f9dce7 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Scaling robot supervision to hundreds of hours with roboturk: Robotic manipulation dataset through human reasoning and dexterity
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 215447fc-f02c-40bf-a3e2-2f7d50dffe5a · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model and Chater, Nick , year =
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 83e2b69b-3d7c-46e1-807e-2d6ab947a696 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Calvin: A benchmark for language-conditioned policy learning for long-horizon robot manipulation tasks
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a72ae54b-0011-465a-8eab-dfc30cf77d7a · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Grounding language with visual affordances over unstructured data
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 702edcdb-dfc9-4d5e-be54-6c7d9ccab732 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Point-E: A System for Generating 3D Point Clouds from Complex Prompts
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 13af462d-d4de-422b-95f2-4a6e7a6c5680 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Open X-Embodiment: Robotic Learning Datasets and RT-X Models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1bef82e3-21cf-4661-a399-b3e06604126e · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model The effects of contextual scenes on the identification of objects
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1016d80c-7344-4cfd-8dec-c1e3aa05f3c2 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Kosmos-2: Grounding Multimodal Large Language Models to the World
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e6884e3c-673a-4d5c-bda9-b7a6d21dbcaf · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Seeing and Visualizing: It's Not What You Think
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1d1f9053-699f-49f2-88c8-43a5177d1300 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Gpt4point: A unified framework for point-language understanding and generation
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 09d79b3b-787a-4f41-89fa-4a843446cb97 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model K., Gokaslan, A., Wijmans, E., Maksymets, O., Clegg, A., Turner, J., Undersander, E., Galuba, W., Westbury, A., Chang, A
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f35d3ee3-8890-43dd-adcc-e608f402d12f · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Grounded sam: Assembling open-world models for diverse visual tasks
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a039cc22-5a82-4cec-a9a6-7697428d927b · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model High-resolution image synthesis with latent diffusion models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 25e1f738-1cce-4d0f-a042-dce8a4a66f31 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Playing with food: Learning food item representations through interactive exploration
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f18b2b71-bf5c-46c7-bd02-07188b7f11e6 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model RoboVQA: Multimodal Long-Horizon Reasoning for Robotics
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation bf507b6f-0f39-4043-8a16-6330bdf8de9b · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model On Bringing Robots Home
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b80b00c1-fc0c-404e-b5b0-488b7d0b8e56 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model MUTEX : Learning unified policies from multimodal task specifications
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a66d1196-1bd7-4435-aafe-0978d513753d · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Lancon-learn: Learning with language to enable generalization in multi-task manipulation
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4664baef-2491-44d6-8a1d-9ab5f4c7e00d · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model and Deng, J
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5b8db1bf-1f28-4e29-b7c2-40499e9029e1 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model R., Black, K., Zhao, T
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation db948bda-a4a1-4bf3-85e3-84dacd27d500 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model NExT-GPT: Any-to-Any Multimodal LLM
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f6750f1f-752a-4b20-9c80-744927bceddd · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Pointllm: Empowering large language models to understand point clouds
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9859893a-4dce-486d-b210-b2128a11b2df · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model Uni3d: Exploring unified 3d representation at scale
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 650d54a1-714d-4bd0-a627-b48d4a576c78 · outbound
3D-VLA: A 3D Vision-Language-Action Generative World Model MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a41e3deb-0fa0-4796-8413-d4d595ae3519 · inbound
DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2d975812-61d1-4c09-b249-95e8f1228012 · inbound
OpenVLA: An Open-Source Vision-Language-Action Model 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d9f196ac-444b-4a81-99da-629611871e77 · inbound
Generalist Virtual Agents: A Survey on Autonomous Agents Across Digital Platforms 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 123
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c8a6d4a-033b-4043-b8b2-31194363a28f · inbound
ShowUI: One Vision-Language-Action Model for GUI Visual Agent 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83f99ab1-5a8a-4afa-a952-dbb517c0ee36 · inbound
LSceneLLM: Enhancing Large 3D Scene Understanding Using Adaptive Visual Preferences 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a3c3806-e811-4803-a045-c40be984e3f8 · inbound
What Matters in Building Vision-Language-Action Models for Generalist Robots 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4312bfa2-5b9b-490a-921c-217ff4d7ae90 · inbound
Bridging Adaptivity and Safety: Learning Agile Collision-Free Locomotion Across Varied Physics 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d10f1bf-b5df-40ff-abd3-8cefb998b810 · inbound
Imagine while Reasoning in Space: Multimodal Visualization-of-Thought 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ff6da187-d1c0-4359-9aea-5a10f721b7dd · inbound
RoboReflect: A Robotic Reflective Reasoning Framework for Grasping Ambiguous-Condition Objects 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74032957-6ba3-4cf5-b1ba-8fb2cc4d88bd · inbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3999e930-803c-42fe-ae4e-cee7ca3af88e · inbound
GeoManip: Geometric Constraints as General Interfaces for Robot Manipulation 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b9472d6-1d08-4b76-b0b4-0eed4be64b16 · inbound
Generative Physical AI in Vision: A Survey 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 244
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73dd6cfd-4145-4c5a-acc3-baefae20e169 · inbound
UP-VLA: A Unified Understanding and Prediction Model for Embodied Agent 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e311560-b342-4516-8f30-a641bd5605d6 · inbound
DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 71c95e1a-7b16-4ade-90ca-b0ae8150b446 · inbound
RoboBERT: An End-to-end Multimodal Robotic Manipulation Model 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 169c83f6-dd58-4877-beda-9487749e1334 · inbound
Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c656c790-2eb5-4a55-b911-72ad627b1760 · inbound
GR00T N1: An Open Foundation Model for Generalist Humanoid Robots 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c310e0ce-5b0d-49ce-8724-05321ed2f72e · inbound
CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2dba29be-db06-4c0a-8c69-5dff826f3490 · inbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0a805211-7440-4d80-a92c-34c9e0a3afd7 · inbound
Generative AI in Embodied Systems: System-Level Analysis of Performance, Efficiency and Scalability 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 159917a2-244a-47dc-afd8-8c86931c3136 · inbound
Masked Point-Entity Contrast for Open-Vocabulary 3D Scene Understanding 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39e33cb3-5677-4d04-a67a-8035344bc37b · inbound
TesserAct: Learning 4D Embodied World Models 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 745f23fe-ccc6-4bb1-bd3a-9a7f00823540 · inbound
Robotic Visual Instruction 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5972453f-af9c-4f0e-8a91-c0ec38829814 · inbound
GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a857750e-8bfe-40c8-b84b-249362800f93 · inbound
VLAs are Confined yet Capable of Generalizing to Novel Instructions 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e7086f8d-467e-4228-b22c-bc9786ff61fd · inbound
DenseGrounding: Improving Dense Language-Vision Semantics for Ego-Centric 3D Visual Grounding 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5095a52b-d1a3-4946-b666-ab2ad5062b46 · inbound
CLTP: Contrastive Language-Tactile Pre-training for 3D Contact Geometry Understanding 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30055d34-7ab0-4b3e-9974-361af1ee686c · inbound
Training Strategies for Efficient Embodied Reasoning 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cff7998-77ee-4db3-8159-e59f5b25357d · inbound
VTLA: Vision-Tactile-Language-Action Model with Preference Learning for Insertion Manipulation 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eebb4f1c-5abf-4e3f-9713-304fc97b826c · inbound
DataMIL: Selecting Data for Robot Imitation Learning with Datamodels 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3ad2b89-8e9e-4b9e-9570-0e14eb8bb404 · inbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d2c2cdd-fcc1-4865-ab40-e0e70f9af25c · inbound
FLARE: Robot Learning with Implicit World Modeling 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 058a3be1-e52f-4b8d-b071-11be389466cb · inbound
ReFineVLA: Reasoning-Aware Teacher-Guided Transfer Fine-Tuning 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae070d1c-4485-46b8-ab14-35feabada817 · inbound
Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38e1098c-acb6-48fe-a4e1-8b5bb8bb162a · inbound
DSG-World: Learning a 3D Gaussian World Model from Dual State Videos 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f26de07-ff57-449c-b422-21cc47b85286 · inbound
Real-Time Execution of Action Chunking Flow Policies 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7601f167-343c-4e52-a9cd-4ace54fb07e0 · inbound
AntiGrounding: Lifting Robotic Actions into VLM Representation Space for Decision Making 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 101
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28d62628-12bb-4d40-bd52-e4a219ccefa7 · inbound
CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 064c76d7-2151-4ef2-b5fb-b740272f4835 · inbound
GAF: Gaussian Action Field as a 4D Representation for Dynamic World Modeling in Robotic Manipulation 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 598224b6-42f0-4c7a-81d7-4efecd310df9 · inbound
DyNaVLM: Zero-Shot Vision-Language Navigation System with Dynamic Viewpoints and Self-Refining Graph Memory 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 336692a1-eda4-4520-9bc3-b1f0b07728e7 · inbound
Beyond Syntax: Action Semantics Learning for App Agents 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b7875404-7e74-441d-8324-22f30dab737a · inbound
WorldVLA: Towards Autoregressive Action World Model 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation dc0d5e19-50a6-4363-b5c3-eef93dbdf956 · inbound
A Survey on Vision-Language-Action Models for Autonomous Driving 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 161
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21d4acac-a0e5-447a-8c71-a4e49754acb8 · inbound
A Survey: Learning Embodied Intelligence from Physical Simulators and World Models 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50648549-db24-47ed-9cd1-4360a95c363b · inbound
VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6560375-45a2-4982-b1a7-492b871e2c5a · inbound
MoGe-2: Accurate Monocular Geometry with Metric Scale and Sharp Details 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c0dbc718-15b3-4fef-ba94-993a9eaf9604 · inbound
Move to Understand a 3D Scene: Bridging Visual Grounding and Exploration for Efficient and Versatile Embodied Navigation 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c3c555a-d500-47c1-aa25-5d7a32955703 · inbound
DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation cbe25985-f876-4a60-ab34-47c1b18f445c · inbound
Reconstructing 4D Spatial Intelligence: A Survey 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fae05599-aee1-4385-b567-8459311861d5 · inbound
Exploring the Link Between Bayesian Inference and Embodied Intelligence: Toward Open Physical-World Embodied AI Systems 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72328781-969b-4e9c-89dc-4916df0347ee · inbound
H-RDT: Human Manipulation Enhanced Bimanual Robotic Manipulation 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 072f4e33-2712-412c-9145-b74cca030574 · inbound
Perceiving and Acting in First-Person: A Dataset and Benchmark for Egocentric Human-Object-Human Interactions 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 147
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc1f00d7-649d-4957-920c-df764718e823 · inbound
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b6ffc71-a3ec-4fc6-a7d2-f446f491920b · inbound
Source Component Shift Adaptation via Offline Decomposition and Online Mixing Approach 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cbb195e-7fb0-4695-8fa5-abad61c589d2 · inbound
Leveraging OS-Level Primitives for Robotic Action Management 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71c88cdc-4890-4d1d-9fb9-ab781a9d3a60 · inbound
Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 232
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bc92a44-f2b0-4e57-97a0-8e62412361e2 · inbound
Grounding Actions in Camera Space: Observation-Centric Vision-Language-Action Policy 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b8aa446-dc3f-4c22-bed2-46c7a8f76ea9 · inbound
Enhancing Reliability in LLM-Integrated Robotic Systems: A Unified Approach to Security and Safety 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09dbc511-d9ba-488f-bf02-af7993566b9b · inbound
Manipulation as in Simulation: Enabling Accurate Geometry Perception in Robots 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bbaa614-04ae-4443-bada-3041e9009be2 · inbound
Mind Meets Space: Rethinking Agentic Spatial Intelligence from a Neuroscience-inspired Perspective 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 228
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55e265b9-81e5-4d9d-a14f-a7bbc814150a · inbound
QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe397185-f5b2-41a5-9be7-3b448b93b77c · inbound
BridgeEQA: Virtual Embodied Agents for Real Bridge Inspections 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 91e3baa7-afcb-44f5-84ed-f9500e5b0f16 · inbound
LISA-3D: Lifting Language-Image Segmentation to 3D via Multi-View Consistency 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bea9803e-44d8-4f44-8903-8701d55a4814 · inbound
VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 662f0ae5-f798-498d-9708-7c14a6018916 · inbound
StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0db86ef-5a47-47d0-8270-c69c21a01831 · inbound
GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 96d4c22e-17a2-424f-9628-7540bcd9a511 · inbound
PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 150
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation eed17cf9-dad7-460c-a9ee-fb795ffe05e1 · inbound
AugVLA-3D: Depth-Driven Feature Augmentation for Vision-Language-Action Models 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 68c16829-e20e-4cf0-b57a-7cc02c56c00c · inbound
Learning Native Continuation for Action Chunking Flow Policies 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b2b36cf2-8633-4e4f-916f-cd52925ae079 · inbound
UniLACT: Depth-Aware RGB Latent Action Learning for Vision-Language-Action Models 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d1a4d5aa-079b-4c3b-bb8a-141773ed4568 · inbound
Notes-to-Self: Scratchpad Augmented VLAs for Memory Dependent Manipulation Tasks 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 110a41a6-daaa-4e9d-af88-49196de51f9d · inbound
VLA Knows Its Limits: Adaptive Execution Horizons for Robot Policies 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcdc9a41-8ff9-429c-8b92-e08d87448d9d · inbound
What if? Emulative Simulation with World Models for Situated Reasoning 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 124
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d9a6d77-5115-48e0-9b8a-a9e383dea878 · inbound
ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5ba0b4dc-7668-4075-93d2-5be0d3cfad58 · inbound
Redefining End-of-Life: Intelligent Automation for Electronics Remanufacturing Systems 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 183
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 85c95fa6-d8a6-4794-bd6b-08754ebac289 · inbound
CoEnv: Driving Embodied Multi-Agent Collaboration via Compositional Environment 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4922074b-c631-45c4-a225-c19cc523f6e9 · inbound
Action Images: End-to-End Policy Learning via Multiview Video Generation 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b9ec0899-39dd-410d-a5d3-1e1a2e9d8bf1 · inbound
R3D: Revisiting 3D Policy Learning 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation fcbf9ec7-030e-40b3-bb66-c11208c6e426 · inbound
${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 705b8e60-503a-4d9c-a00a-3b90167ef080 · inbound
ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3606e2d6-207c-4c55-a974-c29958e5db91 · inbound
ST-$\pi$: Structured SpatioTemporal VLA for Robotic Manipulation 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation dcd18759-44dd-4e78-8fd1-56c7538958a7 · inbound
dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d9edb622-9300-427f-b3b8-ba68b530cf62 · inbound
Vision-Language-Action in Robotics: A Survey of Datasets, Benchmarks, and Data Engines 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 07d746d4-9c3f-4554-baa2-9b13bd8a0739 · inbound
LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2dd1db0e-1f53-4ca7-b3c5-3ca150d7c0ec · inbound
LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 890f35fa-c7d8-4d09-8233-f94583bb73b1 · inbound
Affordance Agent Harness: Verification-Gated Skill Orchestration 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d61860be-bb22-407c-9489-0dda0ebb58fd · inbound
Affordance Agent Harness: Verification-Gated Skill Orchestration 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b9f682cb-cc59-4279-872f-df65e9dd9ae5 · inbound
ConsisVLA-4D: Advancing Spatiotemporal Consistency in Efficient 3D-Perception and 4D-Reasoning for Robotic Manipulation 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5f242772-84dc-4baa-9c5c-4cfbb9bd33d0 · inbound
One Token Per Frame: Reconsidering Visual Bandwidth in World Models for VLA Policy 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9a95271f-d267-452b-90a4-e6e7f090c6eb · inbound
One Token Per Frame: Reconsidering Visual Bandwidth in World Models for VLA Policy 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 717ff208-ad5a-4ddf-9604-cf969d7289fb · inbound
One Token Per Frame: Reconsidering Visual Bandwidth in World Models for VLA Policy 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d462be3f-9b37-4448-b91f-275895efa924 · inbound
VEGA: Visual Encoder Grounding Alignment for Spatially-Aware Vision-Language-Action Models 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 821e4050-6e8b-45fb-aaeb-19d9b37ac222 · inbound
Nautilus: From One Prompt to Plug-and-Play Robot Learning 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 11ba71ca-a685-4913-9c87-de0bb4496f54 · inbound
Nautilus: From One Prompt to Plug-and-Play Robot Learning 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31f76f2d-395f-4460-b1b8-0475b3c65bcb · inbound
World Action Models: The Next Frontier in Embodied AI 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 270
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d652e784-ca52-4d06-9e2c-d8711ced4f10 · inbound
Towards Robotic Dexterous Hand Intelligence: A Survey 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 112
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ce8d7db6-c8d6-40a3-960d-686f48c24b00 · inbound
Evo-Depth: A Lightweight Depth-Enhanced Vision-Language-Action Model 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation cdc1cc04-548e-49d3-950a-9ed717e5c448 · inbound
DyGRO-VLA: Cross-Task Scaling of Vision-Language-Action Models via Dynamic Grouped Residual Optimization 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0cfabf81-4168-472c-ae58-194d1fa6af2f · inbound
ECG-WM: A Physiology-Informed ECG World Model for Clinical Intervention Simulation 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 404c8e18-f088-4e1a-9456-61cb54ff6eb9 · inbound
GaussianDream: A Feed-Forward 3D Gaussian World Model for Robotic Manipulation 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.