Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 100 inbound Pith citation observations for arXiv:2307.15818.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T21:12:55.725471Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
270
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 96658184-8130-456b-a56a-c52d5da317bc · inbound
Cognitive Architectures for Language Agents RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ef194d76-9d83-44ad-b50e-1b4cf63ca8b8 · inbound
GPT-Driver: Learning to Drive with GPT RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f50b1ee2-ebe6-4ea7-b375-ed4d48be7db3 · inbound
Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 649fbcd2-8058-433b-a436-24504b3c9aae · inbound
Learning Interactive Real-World Simulators RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 23bba126-a8fd-4f01-b6e9-09e8094c89de · inbound
Open X-Embodiment: Robotic Learning Datasets and RT-X Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 38c972c4-232c-4ccb-9a83-c2208992c99e · inbound
Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3fd271d6-e777-4d79-844f-8ab2ab44ba53 · inbound
TD-MPC2: Scalable, Robust World Models for Continuous Control RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 156
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 21e3ab03-a12d-4689-9be7-9794c12146e1 · inbound
Vision-Language Foundation Models as Effective Robot Imitators RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e5f92c29-2975-4e70-9586-65035dd4c01b · inbound
Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 73ce41a8-5064-48c6-80d9-0e4fd8e93486 · inbound
AppAgent: Multimodal Agents as Smartphone Users RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation bef27680-7526-4a59-889c-cf316c3614fa · inbound
Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 916b2d1a-9167-419b-aad0-b5a2b05d2e1a · inbound
Agent AI: Surveying the Horizons of Multimodal Interaction RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 120
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f022c4e6-8fd8-4677-aec9-aeda741ec060 · inbound
DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b58cb528-c895-49c0-87a0-a7c36819cb45 · inbound
NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 30ab2bec-9c11-49af-ba4c-a82f90499660 · inbound
RT-H: Action Hierarchies Using Language RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fb41528e-1b26-46b0-ae0e-888efea28365 · inbound
3D-VLA: A 3D Vision-Language-Action Generative World Model RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 37a39ef4-f617-4a6d-abd2-6c000a0eaf42 · inbound
OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4882d3c3-7a9d-4410-9107-ade4035e05b6 · inbound
The Platonic Representation Hypothesis RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 219
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation cdb17367-46f1-4777-8115-4007acb99935 · inbound
A Survey on Vision-Language-Action Models for Embodied AI RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 69f71b86-68fc-4236-957a-6ae74cb5af71 · inbound
OpenVLA: An Open-Source Vision-Language-Action Model RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 54bf625c-a741-4d60-81c8-4c46b7298ba1 · inbound
LongVILA: Scaling Long-Context Visual Language Models for Long Videos RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8158301a-499c-477b-ba24-38f401b22131 · inbound
ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 107
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6bb2cd43-b995-43ca-9603-ddb731b32873 · inbound
TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a130f85b-dd09-4f7b-9b58-01a4f0dcb0df · inbound
Training Language Models to Self-Correct via Reinforcement Learning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 150
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 668a0c42-cfd3-4db6-ab7e-484a94654fc8 · inbound
Towards Robust Surgical Automation via Digital Twin Representations from Foundation Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation cb94fd1d-303c-4ee1-8144-d70b041977cc · inbound
GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f9ac1071-0598-4bd7-bc05-8314ca9f4cb9 · inbound
Language Conditioned Multi-Finger Dexterous Manipulation Enabled by Physical Compliance and Switching of Controllers RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9e164021-00cd-4c88-bc85-f7c0beeb88f3 · inbound
OS-ATLAS: A Foundation Action Model for Generalist GUI Agents RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0ead8dff-ba03-423b-94af-1257f7b9ddca · inbound
$\pi_0$: A Vision-Language-Action Flow Model for General Robot Control RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ea61dd21-5e13-46dc-98ab-88fb0eb5a5a4 · inbound
DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8ff72640-c87b-4848-8c52-9699682fb933 · inbound
DART-LLM: Dependency-Aware Multi-Robot Task Decomposition and Execution using Large Language Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6caea9b3-b158-49f0-9e63-62b1e085c5b6 · inbound
ClevrSkills: Compositional Language and Visual Reasoning in Robotics RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a77a7c43-8faa-4108-ae38-69fb5e90ff21 · inbound
VidMan: Exploiting Implicit Dynamics from Video Diffusion Model for Effective Robot Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afa4cfbc-ac71-4591-86be-8a12bd8caafb · inbound
VeriGraph: Scene Graphs for Execution Verifiable Robot Planning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 48753607-abc0-47e3-afbd-49236c0fdff5 · inbound
Generalist Virtual Agents: A Survey on Autonomous Agents Across Digital Platforms RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0b3482d-54be-4f3a-b3bc-955f090c82cb · inbound
Generative Timelines for Instructed Visual Assembly RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f109449-c97b-4e95-a2f2-ca3fef466c4b · inbound
I Can Tell What I am Doing: Toward Real-World Natural Language Grounding of Robot Experiences RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91d73b72-41e1-486c-8197-770a754d7160 · inbound
Tra-MoE: Learning Trajectory Prediction Model from Multiple Domains for Adaptive Policy Conditioning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 245c755d-aa35-49be-a360-233464c34a6a · inbound
RoCoDA: Counterfactual Data Augmentation for Data-Efficient Robot Learning from Demonstrations RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a76e26a7-79d4-4e8a-83e5-5a2e6a124367 · inbound
ShowUI: One Vision-Language-Action Model for GUI Visual Agent RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41614331-65d7-47e0-8eeb-976ab87d2f51 · inbound
Prediction with Action: Visual Policy Learning via Joint Denoising Process RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe84e8a9-6e0b-499a-9e62-6641a5e5dfe5 · inbound
Embodied Red Teaming for Auditing Robotic Foundation Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0194a111-af39-4318-8b9b-20d84a1e7633 · inbound
GRAPE: Generalizing Robot Policy via Preference Alignment RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 929fd749-adcc-45c6-a506-7b00056469fe · inbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4561561e-a90b-4a51-905b-8febd9491325 · inbound
On Foundation Models for Dynamical Systems from Purely Synthetic Data RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 728e2d13-d0aa-4087-a869-d3d9acaae1ef · inbound
STEVE-Audio: Expanding the Goal Conditioning Modalities of Embodied Agents in Minecraft RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31b84b8e-6667-41da-a7ae-07a1a67e814a · inbound
Quantization-Aware Imitation-Learning for Resource-Efficient Robotic Control RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7da82f87-4651-4a8b-8fc2-a7b3a5b42c37 · inbound
Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation abfa9331-448a-47a8-98cf-831b0e642247 · inbound
RoboMD: Uncovering Robot Vulnerabilities through Semantic Potential Fields RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d68b3fe0-05ac-4fab-8d7e-8ada58bf5b1b · inbound
Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27a61242-1ec9-4906-bc25-7f33746c3a3d · inbound
AIpparel: A Multimodal Foundation Model for Digital Garments RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2b8ee52-f021-44c8-9733-42242a134d02 · inbound
Dissociating Artificial Intelligence from Artificial Consciousness RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 801fa62b-b580-48a0-9527-024e954c2582 · inbound
A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09ae7cd0-ac43-4720-b14d-2d730d95ad82 · inbound
Can Large Language Models Help Developers with Robotic Finite State Machine Modification? RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5457b199-a646-4e3c-8f33-a7ca63a66dbb · inbound
Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8083658f-daf3-49e9-8445-d586424e8d39 · inbound
World knowledge-enhanced Reasoning Using Instruction-guided Interactor in Autonomous Driving RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b519844a-7607-41f4-8b74-1cd4eb869f43 · inbound
CARP: Visuomotor Policy Learning via Coarse-to-Fine Autoregressive Prediction RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da14a212-209f-4beb-88aa-6134e084f8d5 · inbound
The Synergy of LLMs & RL Unlocks Offline Learning of Generalizable Language-Conditioned Policies with Low-fidelity Data RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f51e188b-b378-4b15-b5b3-e540e52c8e9b · inbound
From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb39e0f2-581a-4627-b27e-05764f095e0d · inbound
Learning Novel Skills from Language-Generated Demonstrations RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b262064-4d81-44c2-b0f3-38a88d99d9f2 · inbound
Owl-1: Omni World Model for Consistent Long Video Generation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14a40144-bb53-4fc7-8b60-7c81d6f01980 · inbound
Learning Camera Movement Control from Real-World Drone Videos RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7e826cf-3740-4b27-a9a6-83344279e1bd · inbound
Doe-1: Closed-Loop Autonomous Driving with Large World Model RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d00c172-07b8-4636-b490-b7b01f03a7f9 · inbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f09baac6-b13e-4ac8-9ffd-647ba224fe8e · inbound
TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1732c799-21f6-4b38-b370-1f9ef5864090 · inbound
GROOT-2: Weakly Supervised Multi-Modal Instruction Following Agents RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce788729-3152-4c03-b513-db31a68808d6 · inbound
Modality-Driven Design for Multi-Step Dexterous Manipulation: Insights from Neuroscience RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cc30af0-d5ea-43f4-8c55-f0707043de11 · inbound
Emma-X: An Embodied Multimodal Action Model with Grounded Chain of Thought and Look-ahead Spatial Reasoning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69986738-ce32-4f7d-a3b5-9c5142ebea0b · inbound
Meta-Controller: Few-Shot Imitation of Unseen Embodiments and Tasks in Continuous Control RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d68e3086-cdbd-443c-bf8d-f0aa26cbe441 · inbound
Task-Parameter Nexus: Task-Specific Parameter Learning for Model-Based Control RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2f43dcf-3eab-497d-80b0-bbe697b6ac30 · inbound
Efficient Diffusion Transformer Policies with Mixture of Expert Denoisers for Multitask Learning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 635f13ad-a235-49ee-8cee-bfc845328e1a · inbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f70d6d8-e7a0-4264-823a-6d06bd4319d0 · inbound
What Matters in Building Vision-Language-Action Models for Generalist Robots RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation efb69738-553e-4497-a1e3-e7834093d8a5 · inbound
The One RING: a Robotic Indoor Navigation Generalist RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 870883cb-c452-47dd-b7f6-86fa504fe8fa · inbound
Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7e666c42-3e23-4fbd-8ecd-6c47303b384c · inbound
AutoLife: Automatic Life Journaling with Smartphones and LLMs RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 462edb99-56b0-4f08-b2a2-3ab34c1da5ce · inbound
System-2 Mathematical Reasoning via Enriched Instruction Tuning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d4422b2-3183-4872-b068-6c6730d3f162 · inbound
MMFactory: A Universal Solution Search Engine for Vision-Language Tasks RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddba8c1b-5791-42cf-9a91-2b6cc081af45 · inbound
VLABench: A Large-Scale Benchmark for Language-Conditioned Robotics Manipulation with Long-Horizon Reasoning Tasks RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce96daae-8f69-4ae3-b931-b06f74a29a5f · inbound
Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6322b4e8-2a9f-46d4-ad72-c8609b6e963c · inbound
FaGeL: Fabric LLMs Agent empowered Embodied Intelligence Evolution with Autonomous Human-Machine Collaboration RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afa221d4-4a07-4cff-aeff-66825466c7bf · inbound
CoA-VLA: Improving Vision-Language-Action Models via Visual-Textual Chain-of-Affordance RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c5ee29e-1ccd-49a0-9ad1-fff95276bad7 · inbound
Enhanced Vision-Language Models for Diverse Sensor Understanding: Cost-Efficient Optimization and Benchmarking RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec9c6166-5c3c-4088-9934-5f9b2747f732 · inbound
OmniManip: Towards General Robotic Manipulation via Object-Centric Interaction Primitives as Spatial Constraints RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3bb3328-d326-4921-952d-9afe55280f7d · inbound
Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db453c77-7842-449e-8aad-968d3c912f7a · inbound
Robotic Programmer: Video Instructed Policy Code Generation for Robotic Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14aca93b-3052-4b89-83b0-feb16f436a09 · inbound
Integrating LLMs with ITS: Recent Advances, Potentials, Challenges, and Future Directions RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 273
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b557a1db-8b23-4c04-b92a-d3b8486721ca · inbound
Beyond Sight: Finetuning Generalist Robot Policies with Heterogeneous Sensors via Language Grounding RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12713fc5-4687-4d6c-baa2-32f687342492 · inbound
UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14c7e728-3a75-436b-9d65-cdf6aa305cc5 · inbound
Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b1adefa-12e1-468b-903b-79afe464c0f0 · inbound
From Screens to Scenes: A Survey of Embodied AI in Healthcare RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 289d7746-07bd-40cb-9d58-e247f5c2a9cd · inbound
LAMS: LLM-Driven Automatic Mode Switching for Assistive Teleoperation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 447861ab-e23e-4abc-8f8a-4e8759b184d7 · inbound
Generative Planning with 3D-vision Language Pre-training for End-to-End Autonomous Driving RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d489a10d-73b8-4a93-985f-b51d0b174dad · inbound
Embodied Scene Understanding for Vision Language Models via MetaVQA RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0b41df9-e0d6-4477-9494-b451fa7e1ab9 · inbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0f9a2674-08bf-4b9d-a39a-624b30438b75 · inbound
GeoManip: Geometric Constraints as General Interfaces for Robot Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 065db545-3594-4e8d-a719-9ef9786643fc · inbound
Universal Actions for Enhanced Embodied Foundation Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5768a73a-6f9f-4afe-a8fe-95db8df26036 · inbound
Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5535967-32e9-4051-8d77-8d402f89e2af · inbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf8d3fb4-5f69-4029-a1e2-c59cca2d2d0c · inbound
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.