Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:39:40.605712Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 0 inbound Pith citation observations for arXiv:2608.11671.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:39:40.605712Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
59 of 59 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2f5c8d6a-811d-4c38-ab78-5befcc6692e4 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Qwen3-VL Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67d75d5f-0648-4b58-b217-e545c44d0264 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models RT-H: Action Hierarchies Using Language
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dac024d-934a-4e63-9c28-17108811c050 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Motus: A Unified Latent Action World Model
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d188d7a-3dcb-41a8-9f41-95057c8b2db6 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc553b3e-ccae-4b9a-9985-c34c2ecf9c15 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models UniVLA: Learning to Act Anywhere with Task-centric Latent Actions
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d1af0ab-83cc-4257-9d54-d35208203a99 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models See Once, Then Act: Vision-Language-Action Model with Task Learning from One-Shot Video Demonstrations.arXiv preprint arXiv:2512.07582, 2025
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35e919bf-d023-467d-b3ed-a0256c2a742f · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78dbd86f-6df9-4060-87f0-f9c36e3b9296 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models From Intention to Execution: Probing the Generalization Boundaries of Vision-Language-Action Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92387e10-3135-4277-ae10-736ecc3fef1a · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models LIBERO-Plus: In-depth Robustness Analysis of Vision-Language-Action Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59b30ab7-0fa0-4397-a48c-9c111f559777 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models See what matters: Differentiable grid sample pruning for generalizable vision-language-action model
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e1295b98-35fa-4107-820e-4a881ae8b493 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models WAM-TTT: Steering World-Action Models by Watching Human Play at Test Time
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f09b947f-1824-459f-b2ee-31d585b4c5ce · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Icrt: In-context imitation learning via next-token prediction
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5ca88b4a-50bf-46f0-afc4-e195fcc9caa2 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b74b918f-6ff3-4a17-bc9a-f7d4e5c201ab · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Hancock, Xindi Wu, Lihan Zha, Olga Russakovsky, and Anirudha Majumdar
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc03fbe7-39c9-4a07-a4cb-08641716db5b · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Motion dynamics learning for few-shot embodied adaptation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation dc9ad96b-87e0-4072-bc9b-3499977baaba · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Thinkact: Vision-language- action reasoning via reinforced visual latent planning.Advances in Neural Information Processing Systems, 38: 82782–82802, 2026
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 21ad72bd-a67f-4dfd-a627-4ab833b294fc · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Seeing Realism from Simulation: Efficient Video Transfer for Vision-Language-Action Data Augmentation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 289bd257-aba7-4d12-a147-a9f0cf32edeb · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd1e365e-fbf5-4751-a3d7-ed85140f453d · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Ra-vla: Retrieval- augmented vla for test-time adaptation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 767bc788-6f2e-467b-b80d-9b6644aeec2e · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Ra-vla: Retrieval- augmented vla for test-time adaptation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 66fbaafc-a41f-4c2b-b8aa-67c6d9b5983c · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models RoboTTT: Context Scaling for Robot Policies.arXiv preprint, 2026
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ebb4701a-c999-4648-a9d6-2ed844b4d514 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b433c465-28c0-4c40-9039-47a72fc38fa3 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0faafd41-c780-4e71-95a5-70a99bf761fd · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models MolmoAct: Action Reasoning Models that can Reason in Space
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af116d28-9cb8-4d60-be99-bb45c4d5b95f · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81313a33-46db-4747-aafb-d1ca6837f4af · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Evo-Depth: A Lightweight Depth-Enhanced Vision-Language-Action Model
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a9848be-5e94-4b18-97ae-ed3eaedd743d · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models LA4VLA: Learning to Act without Seeing via Language-Action Pretraining
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b40be1e-1cc9-498f-b0fc-8bb09341e4e2 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cea6bdc0-b60c-4b96-8748-1bd43bed470a · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models LocoFormer: Generalist Locomotion via Long-Context Adaptation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation cb272567-7abf-4b62-a6ba-ad96b84fc08e · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 984c62fa-cf3d-49b1-b3d5-6a27c331b914 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Behavior Prompting Policy: Demonstrations as Prompts for Manipulation
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c8b74a6f-d11d-4d31-bba9-b692ebb2c443 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Action-aware dynamic pruning for efficient vision-language-action manipulation
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 85f6ee72-9c7b-4f32-9dcb-9b5b11063e0c · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93bb0d3b-252d-4ab8-bff1-f5be796a5a06 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b5fe059-9135-45c7-92f3-d01d315ab801 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70f45c2d-7d60-4706-93b6-f443ff7cef90 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc00cbe3-e0d0-45ba-b9f1-25cf18ce645f · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0b64ab9-d276-4788-96aa-e8b824aea016 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b68db90-0d21-4b5d-8502-9ea65dd1f2be · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models RICL: Adding In-Context Adaptability to Pre-Trained Vision-Language-Action Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1865d81-e0a2-4a80-8d75-fcf68bc8ea9b · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19f7e96f-fdaa-4a2a-988f-5868b448f60d · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Test-Time Training with Self-Supervision for Generalization under Distribution Shifts
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41f76f0c-886c-4559-8c30-c280a22a1732 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models X-OP: Cross-Morphology Whole-Body Teleoperation via MPC Retargeting
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 479b66dc-ecf4-4bec-b458-164d4580e070 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models From Foundation to Application: Improving VLA Models in Practice
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70ca2af7-44b9-4716-b5d6-e4d845a1c7b2 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models A V A-VLA: Improving Vision-Language-Action Models with Active Visual Attention
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation afe0bddc-4a58-4ea8-9541-71736e7cf7f9 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Towards efficient embodied reasoning: Mixture-of-depth compute allocation for vision- 17 language-action model
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 878b0789-b711-4571-8299-f330b4f86d39 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Vla-cache: Efficient vision- language-action manipulation via adaptive token caching.Advances in Neural Information Processing Systems, 38:164448–164473, 2026
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45094728-647b-4228-89c2-1813c683421f · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Affordance field intervention: Enabling vlas to escape memory traps in robotic manipulation
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ce9f1612-a1dc-4988-9b10-dde7eb44ebcf · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Latent Action Pretraining from Videos
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60d41e58-ce4f-472d-984a-7d2b30795eae · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Robotic Control via Embodied Chain-of-Thought Reasoning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc9939bf-886c-425e-a7e4-3c06e6301528 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Hancock, Mingtong Zhang, Tenny Yin, Yixuan Huang, Dhruv Shah, Allen Z
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b74d118-0022-4d72-b5a7-8e64f05164a0 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models VLA-Arena: An Open-Source Framework for Benchmarking Vision-Language-Action Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17528a3f-09df-4915-b63c-4472aee2cf7d · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Revisiting Parameter Redundancy in Vision-Language-Action Models: Insights from VLM-to-VLA Adaptation
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c9c63866-d224-4bf7-840f-dddfeabfb829 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 106b0c91-6b38-467a-8957-6146547487af · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d009941a-08dc-4f66-b5a6-f5cbe72df6a1 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Retrieval-VLA: Training-Free In-Context Adaptation for Vision-Language-Action Models
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation af0c0057-cb24-4949-a60b-f467895629d2 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models Retrieval-vla: Training-free in-context adaptation for vision-language-action models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0f32b0c2-d391-4553-90e7-afdd546692c2 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1be27be-a1ef-4583-917c-99e785ed12d8 · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models ACoT-VLA: Action Chain-of-Thought for Vision-Language-Action Models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0f1c1164-5e90-4490-af96-a4da7be0e23c · outbound
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.