Pith. sign in

Paper Citation Record · LEDGER

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models

As of 23 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 4 inbound Pith citation observations for arXiv:2606.03127.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.03127 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T10:09:08.056968Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:39:40.574154Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-11T17:19:04.425879Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact15
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1de303b2-12d0-4f05-a03f-0cdcc793b7fc · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:28.504075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:9c870829d0284b7a6b860d16830936d2687502e1c3cacd9cdace82bf305631e1

Observation 39fd5f3a-dfd1-4e14-9ffd-87163d50aceb · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models RT-1: Robotics Transformer for Real-World Control at Scale

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:28.484803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:48867a431934fbda3de8f41857ad09b41d60b041ae5702842975bba0b03da62e

Observation 666d686c-896a-4b1a-8b00-888265f36037 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:28.511102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:64041ff55e3953830a71df39519570d9533ede8542ca0abdaca69801ca55000d

Observation 865d9098-46fb-490b-b1ef-4f0526efb8f2 · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models RT-1: Robotics Transformer for Real-World Control at Scale

Reference 4

Resolution
verified exact
doi, observed 2026-06-28T10:12:01.793394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:c31651f61fd9b1ed74e469be0202010129a998f4e86cd016d46772791a173e4b

Observation d5fc3e18-2e05-4757-8d2a-0616854e0e03 · outbound

This paper cites Univla: Learning to act anywhere with task-centric latent actions, 2025.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Univla: Learning to act anywhere with task-centric latent actions, 2025

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:b8786909bad9b0d8941991286bb4ad0bb762d0d21441b3ccc2442f02be6847bc

Observation 3356b53c-acbe-49cf-94af-3bf54ab462a6 · outbound

This paper cites an unresolved cited work.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:9dd2f5456afc172af58b01217d1a1e62da1186e984a1e5bde0c27d4e6647d2b2

Observation 21141447-d0a3-46dd-9454-6dd94c81277f · outbound

This paper cites Self-Supervised Policy Adaptation during Deployment.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Self-Supervised Policy Adaptation during Deployment

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:26:28.505481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:6adc356fbd4f869479276aa0cd8fa7536d71fc0834828cf196c5955d7aa8e89d

Observation 5d202c18-6295-4673-a57d-bcd0c1536cab · outbound

This paper cites ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:28.477933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:90b552af218703f5c89a8f735f979a6608c6dcfe227d36af510313dd97c1e430

Observation 7b656867-0b6f-46dd-bca3-a01aacd4eb35 · outbound

This paper cites ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:28.487849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:51491078aab8b266b350bb82eab941c76779edfc2eb8cf8bc936d3a29e0714cd

Observation a1931005-88d3-4563-bae9-059244b2a299 · outbound

This paper cites Oxe-auge: A large-scale robot augmentation of oxe for scaling cross-embodiment policy learning, 2025.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Oxe-auge: A large-scale robot augmentation of oxe for scaling cross-embodiment policy learning, 2025

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:f2d640d9ee1788d60ecabe306442ef2d76b8922f60a4991eabadc25313a06ddc

Observation 1aefb048-e508-4f54-b95a-556153d7da5f · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model, June 2024.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model, June 2024

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:cfc4c3dd930773fd481b4c87ca1182f30f40c548260eca25c7ce4c642fdaa65d

Observation 0d7ecec6-fd65-40f5-b07b-18aa401cfd08 · outbound

This paper cites Sanketi, Quan Vuong, Thomas Kollar, Benjamin Burchfiel, Russ Tedrake, Dorsa Sadigh, Sergey Levine, Percy Liang, and Chelsea Finn.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Sanketi, Quan Vuong, Thomas Kollar, Benjamin Burchfiel, Russ Tedrake, Dorsa Sadigh, Sergey Levine, Percy Liang, and Chelsea Finn

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:20c824de0584069c96c5b6cabf28912ffdc29ff522de50dbf49d51a76fe688a9

Observation 9b75d92f-66d6-43fc-b5c3-08720e32ce60 · outbound

This paper cites Cosmos policy: Fine-tuning video models for visuomotor control and planning, 2026.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Cosmos policy: Fine-tuning video models for visuomotor control and planning, 2026

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:2547b65d29b83bbb8312eef0a29d7863163d07bf951cd25637dd6ff3ada0f875

Observation b8e4a40c-2ddb-478a-9992-246952570912 · outbound

This paper cites Molmoact: Action reasoning models that can reason in space, 2025.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Molmoact: Action reasoning models that can reason in space, 2025

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:447ff4eab8a281c6bfdfbf70f9d1ef80b1ebb620ff43ca3a5ae9ab977cf5bcd5

Observation ad587a74-b92d-462d-88ef-6bc108a52241 · outbound

This paper cites Sanketi Li, Hao Wang, Ajay Mandlekar, Sichen M.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Sanketi Li, Hao Wang, Ajay Mandlekar, Sichen M

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:54261063a8688c5e6f867a8f8dab26c25a49da12f45742cd6a2dc156138bcea0

Observation ad9f474b-de8f-4fbc-b260-ab51989d2257 · outbound

This paper cites CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:28.494101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:6ea000f0ba2914972b70b1eca9d726e9a9ca731028ec8a7bb1de04be1e11b9d3

Observation 7c69d778-3eff-407a-95fb-8fce7d7dd0aa · outbound

This paper cites Emma-X: An Embodied Multimodal Action Model with Grounded Chain of Thought and Look-ahead Spatial Reasoning.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Emma-X: An Embodied Multimodal Action Model with Grounded Chain of Thought and Look-ahead Spatial Reasoning

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:26:28.475024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:f9bbccd91c88f6cf45e2bface78c0af3352b10b3d3b414d0330fc8915e419f8e

Observation ed99c49c-fc2a-4b33-86ad-b7653c49ee90 · outbound

This paper cites What Matters in Building Vision-Language-Action Models for Generalist Robots.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models What Matters in Building Vision-Language-Action Models for Generalist Robots

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:28.502942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:944ff9fc2cae6c9a265749c1d52ecbf45f8ffdc139238abc7081a20ee124f7f1

Observation f1cc4b5a-f056-48a9-8832-336b76fc8966 · outbound

This paper cites Gr00t n1: An open foundation model for generalist humanoid robots, 2025.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Gr00t n1: An open foundation model for generalist humanoid robots, 2025

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:48a17dbd1c36464316cca766ef2c60a7d50cb27da2e37cb44c4a64581bfa73aa

Observation 27a69fa1-e6e6-4317-bfae-f16eedac0f50 · outbound

This paper cites Octo: An open-source generalist robot policy.https://octo-models.github.io, 2023.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Octo: An open-source generalist robot policy.https://octo-models.github.io, 2023

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:8178e5675fa9729fb8b2d0d7393cf7b06edc9c937b3b14dcfb14f71f116e479a

Observation c9977ce6-79ad-42cc-9f9b-072d186fb25e · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:28.508225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:8fef0215158cee3cc5908c16cb1892941572397594ae106a193104ccb38cc1ab

Observation 04f432b7-2fac-4a4c-af88-a387d17d58b8 · outbound

This paper cites Gr-rl: Going dexterous and precise for long-horizon robotic manipulation, 2025.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Gr-rl: Going dexterous and precise for long-horizon robotic manipulation, 2025

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:5514782765f2a12fcb909c2acd2418dce9ac0b3205e8138834dec9281d52d9a4

Observation fdfcfe03-4583-4963-970f-fcca4cfc334e · outbound

This paper cites FAST: Efficient Action Tokenization for Vision-Language-Action Models.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:28.510094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:8b07a5406a53cbe05addd8aacd2ab855f09347959c297c770a351ab6ee43dc0a

Observation 85df7934-ae6a-4e04-8057-b352a87e264d · outbound

This paper cites an unresolved cited work.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:4e5328306473a741e87f96401915ec36108945f950ba792d7bb4ba1c64b70baf

Observation f969c281-3fc6-4981-bfdf-6d4136f55104 · outbound

This paper cites an unresolved cited work.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:a92b978764e5941b320da2a92988008bc06f4c61d9683bde5ab1cda29b347135

Observation a3d5d798-2b76-44d6-913f-03822ec69bee · outbound

This paper cites an unresolved cited work.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:e6794ac7697d4af88897fc34c5b447cc3ead9f48aad9dd898691499c08e4d6c3

Observation c22ee1b2-ca61-4a0a-b44f-8438c63953bf · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:28.489151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:086b521316f033885dce94ebf4b4636d1d9d2d41f341b65fe59249d7b06c8693

Observation 1e0829d0-a4f2-42f7-9140-c6a9a8d0f1d5 · outbound

This paper cites Zhao, Archit Sharma, Karl Pertsch, Jianlan Luo, Sergey Levine, and Chelsea Finn.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Zhao, Archit Sharma, Karl Pertsch, Jianlan Luo, Sergey Levine, and Chelsea Finn

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:9eb4e4bc44707a7575fda0deb53f1a78ee965ac714538f4276294845f1026ed8

Observation d81a32c0-e5b7-4c1d-b12e-dc5db14936a1 · outbound

This paper cites Efros, and Moritz Hardt.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Efros, and Moritz Hardt

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:6c853e07ecf9598b157e63cf51cd85684f1b60a0f8c267b9c3320e5e8e3d59a2

Observation a6fcac0e-0239-4be0-92cc-41ebd23c42df · outbound

This paper cites Bridgedata v2: A dataset for robot learning at scale.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Bridgedata v2: A dataset for robot learning at scale

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:58c2268543646369bd6a06882c18fd1c182e159430e4dcf8a9ca234cd24adadf

Observation c6a6e66a-9908-4b32-ac3a-1c8c6e1c45af · outbound

This paper cites Worldagen: Unified state-action prediction with test-time world model training.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Worldagen: Unified state-action prediction with test-time world model training

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:2d0ce8770e11f965a63537fda410308098c84203792477b8820b9c9c9671b084

Observation d8637e81-6adf-4441-8dde-d8c7a95f7279 · outbound

This paper cites Tent: Fully test-time adaptation by entropy minimization.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Tent: Fully test-time adaptation by entropy minimization

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:01363a153e5404328e9c5219515edfb2b3ae953469202d640d6a0ca3337591b7

Observation 378a51da-1c51-41dd-a59a-aacd70665916 · outbound

This paper cites Scaling proprioceptive-visual learning with heterogeneous pre-trained transformers.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Scaling proprioceptive-visual learning with heterogeneous pre-trained transformers

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:635c1560913709489f397da7dae3227e500dfd22c0c8fa610d9590ca06a870d8

Observation 6175a318-6396-46d6-94ee-8780b88c2f5a · outbound

This paper cites Continual test-time domain adaptation.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Continual test-time domain adaptation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:54b13a3f636a2e02c8e5973118380f4234adcb813a4836d6600412bf51d99d30

Observation 8cc4cf91-12ec-456a-8f76-a55a25695a78 · outbound

This paper cites Rosa: Harnessing robot states for vision-language and action alignment, 2025.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Rosa: Harnessing robot states for vision-language and action alignment, 2025

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:d5d5a4b58637e6ec670ef311f970901facb4d262bc8cf3981aeac4f170cef592

Observation 3ac2ccc1-7c9a-4d5f-b580-4bb4d454f11b · outbound

This paper cites Magma: A Foundation Model for Multimodal AI Agents.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Magma: A Foundation Model for Multimodal AI Agents

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:26:28.500048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:425fa363ac75948642c3c159424fbd50a96dd41608bdd8325271587f1f6cb973

Observation 0e6242a0-668c-4130-be75-caf3054b3322 · outbound

This paper cites Instructvla: Vision-language-action instruction tuning from understanding to manipulation.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Instructvla: Vision-language-action instruction tuning from understanding to manipulation

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:26:28.507147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:b55cf537f267b3732bf7d196e9418980d2ee043a75aeba6b41bb4c2ba277dfed

Observation 9e037bf2-3845-465c-ba55-97b2f04e73b9 · outbound

This paper cites Latent action pretraining from videos, 2025.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models Latent action pretraining from videos, 2025

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:556d6721f7b12edba9b5048114b463d19a97eee7eb24c40955933ebe5cd56127

Observation aa95a2a4-5fc3-4e4e-a844-3b8a79281f9e · outbound

This paper cites World action models are zero-shot policies, 2026.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models World action models are zero-shot policies, 2026

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-28T10:09:08.056968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:9997077ff2bd8812e53d85d27d3f66f724329bdeda62c262b3720001680d28be

Observation 270e5845-5332-48f4-97f2-f4c7013d2efa · outbound

This paper cites CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:28.515355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:446fa29856602a8416272c4a7d78485c96040a2606c1f1117c16f10c7cc2ff57

Observation 372b9453-ca3d-4b3f-b237-fab6ffa9ec55 · outbound

This paper cites TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies.

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:26:28.512910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T10:09:08.056968Z digest=sha256:9d43f366331f84e92311054c2bd1d388ee0f7491eb3fb5d34aec2587e98bbbaa

Pith citing papers

Observation cc23585c-deac-4580-b185-afe683077b5d · inbound

VANE: Reliable Test-Time Training for Vision-Language-Action Models via Future Visual Representation Prediction cites this paper.

VANE: Reliable Test-Time Training for Vision-Language-Action Models via Future Visual Representation Prediction TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-11T17:19:04.439639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-11T17:19:03.262124Z digest=sha256:766b1fe3ea29d1cc5be61f51d0f2009efc8948277147f01a5b95cced2f318520

Observation c96c18bf-6874-429e-9c42-f13a36a879de · inbound

VANE: Reliable Test-Time Training for Vision-Language-Action Models via Future Visual Representation Prediction cites this paper.

VANE: Reliable Test-Time Training for Vision-Language-Action Models via Future Visual Representation Prediction TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:10.083589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:20:10.083589Z digest=sha256:d66ceb9bb2a1d0e22b685b9c47fbd9237f05faff81763a51e15f08341b036a34

Observation 9b0edc57-01e3-4df2-93af-3482d0c51932 · inbound

Self-Evolving Embodied Agents via Skill-Harness Evolution cites this paper.

Self-Evolving Embodied Agents via Skill-Harness Evolution TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T14:16:49.908677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:16:49.908677Z digest=sha256:a14c8eacf359a3878ecfcb9603f46b83475aa47820532420cc78dda739cf7c6b

Observation c9c63866-d224-4bf7-840f-dddfeabfb829 · inbound

StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models cites this paper.

StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-16T00:39:40.574154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:39:40.574154Z digest=sha256:b8a76f97345b1a28258ae55bbaac32b453694006ec2352293054d2b4cb4c217c