Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T04:27:01.485780Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2608.08839.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T04:27:01.485780Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
28 of 28 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9c4cdf7b-882d-406c-b4ef-208ea7513bda · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models Motus: A Unified Latent Action World Model
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45bff85b-e0d3-4f0d-a057-d9580aacb76a · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models π0.5: a vision-language-action model with open-world generalization
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 94a0bf8e-e965-4e35-bea9-f4996e576e80 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models WorldVLA: Towards Autoregressive Action World Model
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45301073-61e7-4430-a514-1d0e169dd403 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41ef3d8a-b30a-4df4-a51e-6c1d3f278632 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models LTX-Video: Realtime Video Latent Diffusion
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65d62c5e-809e-4f91-b6a6-4f2e9e8bfd07 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c097634-f4f9-47eb-a59c-167333b337e8 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models Rynnvla-001: Using human demonstrations to improve robot manipulation.arXiv preprint arXiv:2509.15212,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5647a053-0c1a-4c39-bb99-f593781aed6e · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f83a792-6cb8-4ded-8659-693c7593dfbd · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models Unified Video Action Model
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5fa2be7-2580-4e88-9baa-cdbd2005f8e1 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0d199e1-abd0-4e4c-80c8-412f57e1ed56 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models Depth Anything 3: Recovering the Visual Space from Any Views
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73a90e82-f850-41f5-a373-264846f5716c · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 416ca311-f861-4115-88b2-4b45c2ff0c0c · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models Mask World Model: Predicting What Matters for Robust Robot Policy Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f851e50-8e62-45e1-a400-2bce99d00a32 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models Dit4dit: Jointly modeling video dynamics and actions for generalizable robot control.arXiv preprint arXiv:2603.10448,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df5a5daf-c291-4eb0-8c32-af1a286e8981 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98d7f6f8-d680-4a1a-a310-45b741db355f · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27bcc08b-6877-4a20-ad7d-935a2ad0db66 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models S-VAM: Shortcut video-action model by self-distilling geometric and semantic foresight.arXiv preprint arXiv:2603.16195,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c20bc1a-e422-464d-878c-7855c0698104 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models MaskWAM: Unifying Mask Prompting and Prediction for World-Action Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f19e0cf9-3b5d-42c8-8b9a-e98d39b885b3 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models Fast-WAM: Do World Action Models Need Test-time Future Imagination?
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea6ecb71-d1cd-4df9-b9cb-08a5ea330434 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models FlowVLA: Visual chain of thought-based motion reasoning for vision-language-action models.arXiv preprint arXiv:2508.18269,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01a340b4-28fb-446a-a0ba-9d9f01afd41b · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models DualCoT-VLA: Visual-linguistic chain of thought via parallel reasoning for vision-language-action models.arXiv preprint arXiv:2603.22280,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3d68a9d-5b10-45b2-81f9-802e6f4f655e · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models Unified World Models: Coupling Video and Action Diffusion for Pretraining on Large Robotic Datasets
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e357fa73-6940-4db3-bb28-3a4ac4759dd9 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models DSWAM: A Dual-System World Action Foundation Model for Fine-Grained Robot Manipulation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 6a50ebe8-ed3b-405d-baff-8aefda09b651 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cca0c34-2788-4da8-bdce-8b05e6dbe381 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models UniVLA: Learning to Act Anywhere with Task-centric Latent Actions
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1b0234c-48e2-4d0c-8b5f-f201e595275f · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models LIBERO-Plus: In-depth Robustness Analysis of Vision-Language-Action Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e761f15-8856-4ccf-af6c-764da0415a20 · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c97199d-61b9-4aae-a5e6-e6f88464c71c · outbound
SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models Causal World Modeling for Robot Control
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.