Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-14T15:19:27.489381Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2607.09818.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-14T15:19:27.489381Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1b267005-5ab1-44ed-97b8-954325acfd22 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edc2f0b6-8ead-423b-8773-7cadbacbc3cd · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging LLaDA-VLA: Vision Language Diffusion Action Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 859a8458-897c-4e3e-b624-649a9597d353 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging OpenVLA: An Open-Source Vision-Language-Action Model
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e9b0bf7-589f-4912-a635-7616040e92fc · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Diffusion policy: Visuomotor policy learning via action diffusion,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1a644a8-9ad8-4ef4-9c83-6870ec0aaf79 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 369bdd64-4844-48fb-95e8-985d5619ca7b · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Multimodal Diffusion Transformer: Learning Versatile Behavior from Multimodal Goals
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1259e6dc-5590-49af-98f1-7a7f4d5846e3 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Vla-adapter: An effective paradigm for tiny-scale vision-language-action model,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 706c131b-1f93-4bdd-85d6-3befd6e6e406 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging InterMask: 3D Human Interaction Generation via Collaborative Masked Modeling
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07a52544-01ca-4b0f-9804-8ac9042403f7 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Libero: Benchmarking knowledge transfer for lifelong robot learning,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79606a1f-34e9-497b-bbbf-e86757255dc6 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Calvin: A benchmark for language-conditioned policy learn- ingforlong-horizonrobotmanipulationtasks,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00ffe3d2-333d-4c64-8e7c-9990893e3afa · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c193481-5430-49ad-9a74-b8a516e399ca · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging RT-1: Robotics Transformer for Real-World Control at Scale
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b82ab7c-12e0-4319-aff2-f7a6811707e9 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Rt-2: Vision- language-action models transfer web knowledge to robotic control,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d20bc52c-5272-486a-8d8c-0dc5ed4240f6 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Structured denoising diffusion models in discrete state-spaces,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f08fef00-7267-48db-8903-a0cc860cd472 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Argmax flows and multinomial diffusion: Learning categori- cal distributions,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33b84dd6-d4dc-4b8f-a1a8-7120febb38bc · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Vector quantized diffusion model for text-to- image synthesis,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74ee7fd2-4233-423c-a90e-60b1b5a7c823 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Diffusion-lm improves controllable text gener- ation,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c19a742a-c009-4406-af5c-2505fd47970f · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Maskgit: Masked generative image transformer,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ec63cf4-bed7-4917-8670-0bea5fe5b8fc · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Large Language Diffusion Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e9f574d-00ba-4c30-8cb5-3eeffc71803b · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging DINOv2: Learning Robust Visual Features without Supervision
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7eec16c-baa7-4a21-b11e-8dc943a30131 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Sig- moid loss for language image pre-training,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd2566e3-8b05-4789-bede-072094f40b39 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Qwen2.5-Coder Technical Report
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b27e478-b6d8-4e74-836d-5d0a8d330349 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Flowvla: Thinking in motion with a visual chain of thought,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac4b013c-ed50-4fd0-9e15-e7a256802667 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Cot-vla: Visual chain-of- thought reasoning for vision-language-action models,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d5d561e-cbf3-44ed-9d84-650fbcb13f37 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec5eb7c4-13bc-43c4-90c1-7ce0a8b08a25 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging UniVLA: Learning to Act Anywhere with Task-centric Latent Actions
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8daba2e4-ffd5-4a45-a0fa-89c32a412ad7 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05c4d5c3-5759-4335-9d8b-b72751d51077 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Towards Synergistic, Generalized, and Efficient Dual-System for Robotic Manipulation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 756c7481-332b-4c0b-a611-265b8f4fbe84 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 679555cc-f140-41bf-872f-17c2adabb100 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f67087ff-e350-4716-9512-a028039cd180 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2217ddb-d8c4-43e1-ae7b-9aa2aeb76629 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65a9fb3d-95c6-47ff-a54b-f1815e42c368 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Deer-vla: Dynamic inference of multimodal large language models for efficient robot execution,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da9a171c-ec69-422b-b275-2573ef587e77 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Vision-Language Foundation Models as Effective Robot Imitators
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95876727-8414-4d8d-9cfd-a16fda10e621 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fae95702-3314-47fe-92ec-32a640721d30 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3541f0cc-245d-4df8-a6af-7f0f93129804 · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging VLA-OS: Structuring and Dissecting Planning Representations and Paradigms in Vision-Language-Action Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fde7b3c-846c-4c76-a20f-43311929f11f · outbound
TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging Efficient Diffusion Transformer Policies with Mixture of Expert Denoisers for Multitask Learning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.