Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-11T08:52:31.686474Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 73 of 73 outbound references and 100 inbound Pith citation observations for arXiv:2501.09747.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-11T08:52:31.686474Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T20:38:42.966861Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
73 of 73 outbound references displayed
External citation measurements
0
pith, observed 2026-08-05T02:28:24.338817Z
Observation bae3ca87-2444-40dc-b1c3-d16c9b308870 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Dis- crete cosine transform
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e0d2a305-fe01-4fb1-88d8-8aa637f3478b · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2c402382-742d-4233-a846-ace70612bfdd · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Minivla: A better vla with a smaller footprint
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7c4b5521-4ff1-435b-bf32-e5ac0ff4fe6a · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models RT-H: Action Hierarchies Using Language
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bc76d6e0-6103-431e-84e1-79558827d5e8 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models PaliGemma: A versatile 3B VLM for transfer
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4a6108e2-a318-4763-9cb0-824cf8f55913 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Roboagent: Generalization and efficiency in robot manip- ulation via semantic augmentations and action chunking
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0600c0bc-a69d-4080-8a38-1d103ef5b0e3 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4ed60879-415d-45e0-b15b-cdbf7d349585 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models RT-1: Robotics Transformer for Real-World Control at Scale
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e0b41df9-e0d6-4477-9494-b451fa7e1ab9 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation fc482c4c-3273-44e8-85fd-cce4d26bac5b · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0fd8edf6-6bcf-447c-a786-d6886f7d1814 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models BEATs: Audio Pre-Training with Acoustic Tokenizers
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 836504b2-6e5d-48a7-9dc3-fa941533a835 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e63e21ac-9087-4699-806b-214448dc7f00 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Open-TeleVision: Teleoperation with Immersive Active Visual Feedback
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e854caf7-f3ec-4203-b09a-61cc28501a40 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Dif- fusion policy: Visuomotor policy learning via action diffusion
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b526abab-ecb1-49be-b41f-bd5255d464d2 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Universal manipulation interface: In- the-wild robot teaching without in-the-wild robots
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f542c32f-c3fa-4e53-993b-2248a17a4a93 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models An algorithm for the machine calculation of complex fourier series
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c62ca3ab-b25a-4109-9967-2a984199dc1f · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Keypoint action tokens enable in-context imitation learning in robotics
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2ce0f7c7-e664-495d-ad2e-176afc93a187 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Scaling cross-embodied learning: One policy for manipulation, navigation, locomotion and aviation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a54155e5-b5b8-4145-9dc9-ca23203faeab · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models PaLM-E: An Embodied Multimodal Language Model
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e4f9e417-9a48-4503-bd8c-7f4dbf7b5d84 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Tam- ing transformers for high-resolution image synthesis
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d99d2fd7-e9c8-47e6-b7f8-9b1e24cab9ab · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Qi, Yin Zhou, Zoey Yang, Aur’elien Chouard, Pei Sun, Jiquan Ngiam, Vijay Vasudevan, Alexander McCauley, Jonathon Shlens, and Dragomir Anguelov
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0aee9769-5477-46f4-aeba-c69181368476 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Rh20t: A comprehensive robotic dataset for learning diverse skills in one-shot
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 74ea73da-19d1-4b11-8546-fccdad4d667c · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Moka: Open-world robotic manipulation through mark-based visual prompting
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1de2e5bb-5534-4c44-add6-9afae8b7ec89 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Humanplus: Humanoid shadowing and imitation from humans
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 139b7318-649a-4636-997b-a44d25034a95 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models A new algorithm for data compression
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9323ea76-cf37-4d3c-88f2-e2d062098c6d · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Multilingual Language Processing From Bytes
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1887f98b-1110-4843-aaf3-33a3be1c7998 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Singing voice graph modeling for singfake detection
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d374ff46-e885-4174-8c12-8a6ca347d6e5 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Bridging the Human to Robot Dexterity Gap through Object-Oriented Rewards
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f3ff5b79-7cd0-47cf-b94b-fc97f1c58d63 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models UMI on legs: Making manipulation policies mo- bile with manipulation-centric whole-body controllers
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1951f2e3-cbfb-4a6f-bd29-26b41b2ad0f8 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 930bdf6b-cfc2-4353-b922-9366a8ef917a · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 70bc0a3e-e3ef-4560-86b2-de67a24c6e60 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Efficient Long Video Tokenization via Coordinate-based Patch Reconstruction
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ca0b98e2-7f33-4f5e-b31a-8f295255e97a · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models DexMimicGen: Automated Data Generation for Bimanual Dexterous Manipulation via Imitation Learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 17831756-63ec-4a95-9488-fbbc611cb21f · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Beyond Sight: Finetuning Generalist Robot Policies with Heterogeneous Sensors via Language Grounding
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5d086ce1-2226-47e6-bb46-f50689720485 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Pris- matic vlms: Investigating the design space of visually- conditioned language models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d11cc28c-9a22-44b3-b2f5-50a8e4853b89 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 88da5c10-2c5c-450b-adec-a4abfa1220ff · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 01f6c475-9287-4e2a-950d-bbd4ac6a2fbe · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Action chunking as conditional policy compression
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5603573c-d0d0-49bb-aada-44aa7a7651f6 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Behavior Generation with Latent Actions
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 773e604a-216c-4d02-8af2-93adc7485486 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Learning Visuotactile Skills with Two Multifingered Hands
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5d35112b-3e22-4330-a9ef-823018900679 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Libero: Benchmarking knowledge transfer for lifelong robot learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f40da83b-c9a4-4e13-b4c3-1df1f3ac1ec0 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Visual instruction tuning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 56bfbb44-3817-4bd2-ac75-64ef67e9f81a · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Decoupled Weight Decay Regularization
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 98c3cb93-c186-4589-95dd-d5c95f24ccfc · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Serl: A software suite for sample-efficient robotic reinforcement learning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 87371680-dead-4f95-92a2-3784f2bcfecb · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Roboturk: A crowdsourcing platform for robotic skill learning through imitation
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 559e37a9-ac78-4c21-a367-b1a0e981aad3 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Finite scalar quantization: Vq- vae made simple
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ca6e0a4b-3e52-4a6f-aca0-5036622043c5 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models QueST: Self-Supervised Skill Abstractions for Learning Continuous Control
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c9591172-98da-4c6b-981a-ac2101d7d1e0 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Pivot: Iterative visual prompting elicits actionable knowledge for vlms
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a516bdb4-14b9-47b2-86c5-abd31b7ec9f3 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Octo: An open-source generalist robot policy
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 36b4bd67-6172-4190-ad77-550a7bb25457 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Open X-Embodiment: Robotic Learning Datasets and RT-X Models
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4a4e7967-2e5d-4712-818d-959e767d0ece · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Byte latent transformer: Patches scale better than tokens
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 788e0cc3-a08a-4a69-b4d0-b08df8f32c1e · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models In-Hand Object Rotation via Rapid Motor Adaptation
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5ab9d296-bab5-404f-8188-319f34c3b3e7 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Language models are unsupervised multitask learners
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1f02c590-d58a-4d6e-84e0-6815ecfb47a5 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models A generalist agent
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 30d41e1b-2759-45e0-acd8-0a23b752066e · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Neural Machine Translation of Rare Words with Subword Units
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5ce49807-3e8e-46a9-87c7-5ab79d51082a · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Hand-Object Interaction Pretraining from Videos
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 136b9880-42d6-40fa-ac61-9d3b0c7f47a5 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Neural discrete representation learning
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7ac2cf8b-b9e1-4fdb-81cd-fbd253706109 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Neural Discrete Representation Learning
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5d493849-b62b-42ef-9b98-f67b871fe005 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models BridgeData v2: A dataset for robot learning at scale
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 62f4316b-b1af-4955-81aa-85293122d50e · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models The jpeg still picture compression standard
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0ef39012-e3bb-40f7-88b8-b72e348e123f · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Scaling proprioceptive-visual learning with hetero- geneous pre-trained transformers
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c5216e5f-b39a-4f44-83d4-f866473f92fc · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b1eb95a1-be23-421e-96f6-83b3bedfc13b · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models ElasticTok: Adaptive Tokenization for Image and Video
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 57cfbfcb-8c20-4ce7-95b9-c5e853cb1f70 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Latent Action Pretraining from Videos
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b8ac0df3-9642-4073-aa10-fa8153dde81c · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models MAGVIT: Masked Generative Video Transformer
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4f83859b-9f94-42aa-8527-aa8ca9156103 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Robotic control via embodied chain-of-thought reasoning
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 79955a72-0acc-47f7-a02a-8a740e4e8185 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models SoundStream: An End-to-End Neural Audio Codec
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c6a7f4cc-755b-43a4-b328-2f3589bc4b7f · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c58671ea-6b25-4182-a6a0-aada05b60156 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models ALOHA Unleashed: A Simple Recipe for Robot Dexterity
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 74032957-6ba3-4cf5-b1ba-8fb2cc4d88bd · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1b72a23b-1c38-46e7-b57a-0998e8f618de · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6e150253-6d71-4ff0-ba0b-2ab9035b8ed9 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Autonomous im- provement of instruction following skills via foundation models
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b6d45ae7-00c2-46b6-840f-b972d00b3a48 · outbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models Compression of individ- ual sequences via variable-rate coding
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3c17340d-0043-4f36-a7e6-03d4a668d1f0 · inbound
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c177e2c1-4370-407b-86d3-c32632cc9b07 · inbound
DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 432d5666-6801-4c84-88bc-4e89fa89b527 · inbound
Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e1190dcc-6d96-4cf3-96ff-2a9be3cf739c · inbound
Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f78da1bd-03f3-4eb1-9027-821ab3f6391b · inbound
SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0b31a95a-8506-4798-b973-785e19acce9d · inbound
HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 20008719-ae60-4e23-baf5-b47dee0db4db · inbound
NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 74dd22fd-a4a0-41d6-a649-ec3271b9cf9c · inbound
Policy Contrastive Decoding for Robotic Foundation Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a6e56e9e-e96c-40e7-83bf-857197eb6c07 · inbound
FLARE: Robot Learning with Implicit World Modeling FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a523c01d-a6c5-48eb-a79e-506e999fe7c4 · inbound
Interactive Post-Training for Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6990683a-0047-4f84-bd40-d34b0f706960 · inbound
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 14c9ae98-a376-491f-8152-377b49fbaa14 · inbound
SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e8512285-7216-41cf-9287-43949bdc2555 · inbound
Real-Time Execution of Action Chunking Flow Policies FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 553badf2-5ebb-4667-9690-cfadccfa0be3 · inbound
AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4330f603-7bf6-4bb5-93a0-158667a91944 · inbound
WorldVLA: Towards Autoregressive Action World Model FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a189148d-5bb7-4167-b865-d57482b0d284 · inbound
A Survey on Vision-Language-Action Models: An Action Tokenization Perspective FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 126
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation aa644d30-0dab-443b-9c0f-b9cda949ddd7 · inbound
DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ddd62206-4c0b-46d1-95b6-60a2f4ee14bc · inbound
AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5130156d-0f88-40e8-bed6-72e2c2798d39 · inbound
GR-3 Technical Report FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation df4247d9-d154-4a04-b7b0-c36570efb474 · inbound
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 957ef696-5bca-4bed-9f3c-ac4dd53f3b93 · inbound
Source Component Shift Adaptation via Offline Decomposition and Online Mixing Approach FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd557c63-484d-40f1-856b-a2b5012b244e · inbound
Leveraging OS-Level Primitives for Robotic Action Management FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61a8e01b-d9bf-4685-9cb5-d77a6979d8e9 · inbound
Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 146
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e33c551-7fab-42c5-bc95-867a4a125250 · inbound
Robotic Manipulation via Imitation Learning: Taxonomy, Evolution, Benchmark, and Challenges FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c7d7538-4911-4813-b2db-7f81eeaad77d · inbound
MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ecbac68c-3f86-4700-a0b6-05c1aa1d1bf1 · inbound
CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99d155fb-1462-4ecf-86ab-dfdaf40e4073 · inbound
Mechanistic interpretability for steering vision-language-action models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d71fa393-e759-4444-b8d8-e12ec46b78aa · inbound
Galaxea Open-World Dataset and G0 Dual-System VLA Model FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87117c1b-5aed-4d9f-a546-1a2317f3df49 · inbound
Align-Then-stEer: Adapting the Vision-Language Action Models through Unified Latent Guidance FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 439805ee-919d-4ebb-ad27-10464cb11206 · inbound
FLOWER: Democratizing Generalist Robot Policies with Efficient Vision-Language-Action Flow Policies FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f748370-c075-4b89-bc33-0ce4e1404976 · inbound
F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 46dee853-bb34-49b8-b5b4-1db200593787 · inbound
TA-VLA: Elucidating the Design Space of Torque-aware Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cff1b3eb-0556-428f-9437-8e359a293b07 · inbound
RoboChemist: Long-Horizon and Safety-Compliant Robotic Chemical Experimentation FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90e33f91-4bd9-483b-87a4-aa566715af39 · inbound
Ask-to-Clarify: Resolving Instruction Ambiguity through Multi-turn Dialogue FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e7e6cae-fae1-4465-9325-d976b76eb960 · inbound
VLBiMan: Vision-Language Anchored One-Shot Demonstration Enables Generalizable Bimanual Robotic Manipulation FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 193d796d-07f2-4d6b-aee5-002dfeeb6fd2 · inbound
World-Env: Leveraging World Model as a Virtual Environment for VLA Post-Training FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0de6fb67-aa8b-45f4-8da8-a62bb926f755 · inbound
RobustVLA: On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 287b5301-89a9-43ae-9f5d-424e6a6fde02 · inbound
INSIGHT: INference-time Sequence Introspection for Generating Help Triggers in Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebc4c4b4-72dd-4335-bef4-e9d2bf973cff · inbound
Contrastive Representation Regularization for Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79e7768f-f511-4db8-a540-5ded61693a00 · inbound
Verifier-free Test-Time Sampling for Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62bff911-0d1e-4e43-bc97-ef5521b9c79c · inbound
Humanoid Everyday: A Comprehensive Robotic Dataset for Open-World Humanoid Manipulation FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edf9e0ab-06cc-4ebb-9615-0da54e91ebae · inbound
Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e68e9b60-6f35-4c6c-8d4f-e658c700d96d · inbound
Ctrl-World: A Controllable Generative World Model for Robot Manipulation FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f0ab2c21-24d0-405a-a98e-ed8f47ca68e0 · inbound
DriveVLA-W0: World Models Amplify Data Scaling Law in Autonomous Driving FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 39461b7e-6354-4ea3-aa17-8eaaa2a13d30 · inbound
InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a1a160eb-83f9-4afc-b7b9-b8bc5f6f352d · inbound
Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72f02667-d3bd-4e7a-9ff9-d0d050a2785e · inbound
Speculative Coupled Decoding for Training-Free Lossless Acceleration of Autoregressive Visual Generation FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 90750e1c-e854-4de1-8a6f-05946be50bdd · inbound
LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Manipulation FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0b6dfa7c-d79a-441e-b759-72be7aa7a3ba · inbound
SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 00c6d782-afb3-41f0-96ca-d67350f92265 · inbound
AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5e0fcc0a-341c-4338-bf45-cb8dd4e1cd8f · inbound
SPEAR-1: Scaling Beyond Robot Demonstrations via 3D Understanding FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dc291951-9c29-49b9-b360-0ed466e93021 · inbound
RynnVLA-002: A Unified Vision-Language-Action and World Model FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2dd1631e-8d20-4d39-a5ff-adaca0501b58 · inbound
AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 28178c42-aeb1-40bc-8033-01a9d1997d73 · inbound
Mixture of Horizons in Action Chunking FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8884fcb-8c89-4f58-b5b6-ee2dd654999b · inbound
StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e649f92-2d80-4e7b-b914-4925b27d864e · inbound
Clutter-Robust Vision-Language-Action Models through Object-Centric and Geometry Grounding FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c65aa1f7-b389-4017-92a7-465d9e25a3f6 · inbound
VLA-Arena: An Open-Source Framework for Benchmarking Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0065bab7-4218-4845-b536-8c89d63e06e6 · inbound
Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d790f23-2631-4d05-b551-82bdf4dfa8d7 · inbound
VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 063ca814-7ea0-4958-ac76-c7a1005d8d41 · inbound
Stable Language Guidance for Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2ac26e8b-edd2-4aed-bd47-aa1b694e3992 · inbound
DextER: Language-driven Dexterous Grasp Generation with Embodied Reasoning FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f6e8c49b-1f4f-41e9-9f9b-b548a1219b0d · inbound
Supervised Mixture-of-Experts for Surgical Grasping and Retraction FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9c8d4159-35f2-4fb4-848f-eb7f6e40c2e7 · inbound
MetricAnything: Scaling Metric Depth Pretraining with Noisy Heterogeneous Sources FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca02669e-59e7-4cf2-a841-51dbd3fae137 · inbound
Towards Long-Lived Robots: Continual Learning VLA Models via Reinforcement Fine-Tuning FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ae4d161f-6ec2-4318-80e5-328138cf44c1 · inbound
ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b3be4d23-8927-472f-9f06-f2782764af10 · inbound
Learning Native Continuation for Action Chunking Flow Policies FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d38a70cd-7f3f-4f91-ab79-10108d364c10 · inbound
Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation efdb2e1d-c66e-4073-a413-99ed4ad9b070 · inbound
When Vision Overrides Language: Evaluating and Mitigating Counterfactual Failures in VLAs FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2198c2b-e888-4d1f-b497-b45c68da367f · inbound
VLANeXt: Recipes for Building Strong VLA Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 185ec468-d8ee-497a-bc6c-2be3224e419c · inbound
PoseVLA: Universal Pose Pretraining for Generalizable Vision-Language-Action Policies FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 58dc68ba-727b-4dc8-a9e4-3b0ed97a74f7 · inbound
PoseVLA: Universal Pose Pretraining for Generalizable Vision-Language-Action Policies FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23e97b66-5306-47be-8230-9a484a30460b · inbound
Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Efficient Robotic Manipulation FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9a2440b7-8291-444f-a112-09c142415ee9 · inbound
QuantVLA: Scale-Calibrated Post-Training Quantization for Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 47cc75fc-f237-437a-b673-4da9cd804bdd · inbound
Learning Physics from Pretrained Video Models: A Multimodal Continuous and Sequential World Interaction Models for Robotic Manipulation FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ac54e185-26f9-448e-932c-55ea746034b7 · inbound
KERV: Kinematic-Rectified Speculative Decoding for Embodied VLA Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9e779ea9-0eea-4086-b075-03cba721c6c7 · inbound
Neural Implicit Action Fields: From Discrete Waypoints to Continuous Functions for Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab770e7a-bba9-4b6e-b0b4-0a257bf42ddf · inbound
Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5b0b7430-c71b-4e2b-be27-aa5080bb5a42 · inbound
AR-VLA: True Autoregressive Action Expert for Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ff145d29-385c-4ae7-a8bb-486a5b596d7e · inbound
vla-eval: A Unified Evaluation Harness for Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 80c3c169-1c8d-4e58-a86c-25698f7a17ba · inbound
OxyGen: Unified KV Cache Management for VLA Inference under Multi-Task Parallelism FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dc73b356-41ad-4121-83b3-3aaaf8f56e93 · inbound
Towards Generalizable Robotic Manipulation in Dynamic Environments FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a856bf57-2118-4f9b-98fb-4d89109fc848 · inbound
Towards Generalizable Robotic Manipulation in Dynamic Environments FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cea80e3-bb25-4814-a574-cc457e8bee22 · inbound
HeiSD: Hybrid Speculative Decoding for Embodied Vision-Language-Action Models with Kinematic Awareness FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f4d48a6c-510e-4923-9cfa-1be29594313c · inbound
Generative Control as Optimization: Time Unconditional Flow Matching for Adaptive and Robust Robotic Control FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4dcfb5fd-2318-4d8f-8512-5ccfaf5b04fb · inbound
FASTER: Rethinking Real-Time Flow VLAs FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2004ca27-2f34-428f-a72d-21042e7c79e6 · inbound
FASTER: Rethinking Real-Time Flow VLAs FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation fc8b8e5b-b726-4976-9e72-b1e8451f21fd · inbound
LaMP: Learning Vision-Language-Action Policy with 3D Scene Flow as Latent Motion Prior FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bccb7a51-0e66-4057-89eb-0d29352d9c81 · inbound
Fast-dVLA: Accelerating Discrete Diffusion VLA to Real-Time Performance FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 51321e68-5091-4560-9106-27ef63b7aff6 · inbound
DFM-VLA: Iterative Action Refinement for Robot Manipulation via Discrete Flow Matching FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e024f619-ddaa-41a3-a3c6-f57b528a210e · inbound
DIAL: Decoupling Intent and Action via Latent World Modeling for End-to-End VLA FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bcbff7ca-3b7d-463a-9102-b1afe759f26b · inbound
Multi-View Video Diffusion Policy: A 3D Spatio-Temporal-Aware Video Action Model FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1226feed-329d-45ac-919c-5dd5640eb28d · inbound
The Compression Gap: Why Discrete Tokenization Limits Vision-Language-Action Model Scaling FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6167e5e1-afb0-4904-8d8b-fe4b664ff974 · inbound
Hierarchical Planning with Latent World Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 68b4e89c-26f5-4668-8238-97e961dcdcac · inbound
Adaptive Action Chunking at Inference-time for Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 08970cd1-31fa-481d-b91e-e80f4dcba141 · inbound
Hierarchical SVG Tokenization: Learning Compact Visual Programs for Scalable Vector Graphics Modeling FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 349dd5ee-56fe-419a-b699-aa196cc02fc1 · inbound
VLA-InfoEntropy: A Training-Free Vision-Attention Information Entropy Approach for Vision-Language-Action Models Inference Acceleration and Success FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2136fea0-43eb-4595-9015-dac234295ab2 · inbound
A1: A Fully Transparent Open-Source, Adaptive and Efficient Truncated Vision-Language-Action Model FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7c5f000f-ae8e-4009-b3e7-b2eafafa7314 · inbound
AssemLM: A Spatial Reasoning Multimodal Large Language Model for Robotic Assembly FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bdd9e340-079f-4dbb-9eef-f6cce84d9422 · inbound
AssemLM: A Spatial Reasoning Multimodal Large Language Model for Robotic Assembly FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73d523bc-9056-4ba4-af8a-e9afb3b666d4 · inbound
VAG: Dual-Stream Video-Action Generation for Embodied Data Synthesis FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.