Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T12:55:09.062349Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 7 inbound Pith citation observations for arXiv:2510.01711.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T12:55:09.062349Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T04:48:33.666343Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-09T20:16:29.394437Z
27 of 27 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5bce5ec1-0743-4d7c-8548-6566d4ba1f11 · outbound
Contrastive Representation Regularization for Vision-Language-Action Models Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7286cbf-518e-4ace-96dd-24a59d2fd82b · outbound
Contrastive Representation Regularization for Vision-Language-Action Models GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6bea193-1965-4835-8961-8034b8ffadd5 · outbound
Contrastive Representation Regularization for Vision-Language-Action Models π0.5: A vision-language-action model with open-world generalization
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6aa2bef-5345-46b2-bf97-04435c4764aa · outbound
Contrastive Representation Regularization for Vision-Language-Action Models Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e592b8ee-39dd-44e2-8d92-2ef177885807 · outbound
Contrastive Representation Regularization for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c93a65da-8c90-463e-ae33-75d26468fd64 · outbound
Contrastive Representation Regularization for Vision-Language-Action Models Rectified Flow: A Marginal Preserving Approach to Optimal Transport
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37d05fd4-08c9-4391-b1e0-188528c9210e · outbound
Contrastive Representation Regularization for Vision-Language-Action Models Representation Learning with Contrastive Predictive Coding
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebc4c4b4-72dd-4335-bef4-e9d2bf973cff · outbound
Contrastive Representation Regularization for Vision-Language-Action Models FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 721b8a4e-f81f-432b-8093-86e6492b33a4 · outbound
Contrastive Representation Regularization for Vision-Language-Action Models RoboBrain 2.0 Technical Report
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66231df5-5a33-4d02-bc12-78ed3707f785 · outbound
Contrastive Representation Regularization for Vision-Language-Action Models SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f739e7e-dcb6-4c8a-88cf-b8d7f6607c8d · outbound
Contrastive Representation Regularization for Vision-Language-Action Models Instructvla: Vision-language-action instruction tuning from understanding to manipulation.arXiv preprint arXiv:2507.17520,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81a5e571-eef7-4a7d-8bfe-cc1b12e59f34 · outbound
Contrastive Representation Regularization for Vision-Language-Action Models ChatVLA: Unified Multimodal Understanding and Robot Control with Vision-Language-Action Model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 141c18cf-8d78-4b88-9872-44c05f7541eb · outbound
Contrastive Representation Regularization for Vision-Language-Action Models Under review
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2303cfd-1190-4784-bf84-05fd52a4410f · outbound
Contrastive Representation Regularization for Vision-Language-Action Models Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8de92d7-c11c-49e5-9e85-06084d552415 · outbound
Contrastive Representation Regularization for Vision-Language-Action Models We omit the use of future tokens (Zheng et al., 2025), as they are beyond the scope of this work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d0239ec-050a-4031-b824-e62640eeb580 · outbound
Contrastive Representation Regularization for Vision-Language-Action Models We randomly sample 10 trajectories per task in RoboCasa-Kitchen, totaling 240 trajectories
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b036cb1-3a6b-42d5-be7d-10a50ea3bbea · outbound
Contrastive Representation Regularization for Vision-Language-Action Models This result indicates the effectiveness of our proposed training framework, together with the augmen- tation strategyview cutoff
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3485f25-d6a6-4468-9a50-ec903283213a · outbound
Contrastive Representation Regularization for Vision-Language-Action Models Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffc5700d-2524-4e39-9e9d-3313b1a8dde0 · outbound
Contrastive Representation Regularization for Vision-Language-Action Models At inference, we use an action horizonH= 16and execute all actions without re-planning
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aacef3bc-9217-40d9-8880-a3d3eea812f0 · outbound
Contrastive Representation Regularization for Vision-Language-Action Models A Simple but Tough-to-Beat Data Augmentation Approach for Natural Language Understanding and Generation
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7ed0226-df83-4acc-85f7-de7d50c3feaf · outbound
Contrastive Representation Regularization for Vision-Language-Action Models SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf66936f-26c3-413e-b8b7-75197fa61bb3 · outbound
Contrastive Representation Regularization for Vision-Language-Action Models Contrastive Language, Action, and State Pre-training for Robot Learning
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8684ed19-c732-44f0-b64a-6561ec1fe91b · outbound
Contrastive Representation Regularization for Vision-Language-Action Models Visual Embodied Brain: Let Multimodal Large Language Models See, Think, and Control in Spaces
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b27409a-ab77-4523-ae8d-7674f4300620 · outbound
Contrastive Representation Regularization for Vision-Language-Action Models Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4efb3bb9-9d4e-4e97-bfc3-f181a9694406 · outbound
Contrastive Representation Regularization for Vision-Language-Action Models NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83b67c86-3d92-4be5-8a20-d3e8ea4aec3c · outbound
Contrastive Representation Regularization for Vision-Language-Action Models Qwen2.5-VL Technical Report
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f32c699b-454d-47a1-bcb8-988ae78d1607 · outbound
Contrastive Representation Regularization for Vision-Language-Action Models Layer Spatial Object Goal Long Avg
Reference 2048
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f656816d-568a-4ff6-9255-f2bae2644b94 · inbound
Understanding the Impact of Geometric Foundation Models on Vision-Language-Action Models Contrastive Representation Regularization for Vision-Language-Action Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 284162fa-9d21-4585-b7a7-56e799c22527 · inbound
QuoVLA: Quotient Space for Vision-Language-Action Models Contrastive Representation Regularization for Vision-Language-Action Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9e6b5330-20e8-4f91-ad74-8b56879c6d8c · inbound
Mitigating State Aliasing in Vision-Language-Action Models via Inverse Dynamics Learning Contrastive Representation Regularization for Vision-Language-Action Models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 49449ac5-7ac0-458d-9085-1d1aff007f73 · inbound
FiberTune: Preserving Action-Fiber Visual Residuals in Vision-Language-Action Fine-Tuning Contrastive Representation Regularization for Vision-Language-Action Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 24f9ca34-e0ed-414f-9595-1f9cc1826962 · inbound
Contrastive Action-Image Pre-training for Visuomotor Control Contrastive Representation Regularization for Vision-Language-Action Models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 891ec8fd-1fe0-4561-b352-77ea4a9bb78a · inbound
GeoProp: Grounding Robot State in Vision for Generalist Manipulation Contrastive Representation Regularization for Vision-Language-Action Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8aca349e-d11e-48ff-a233-a54667a6fffb · inbound
Semantic Anchoring for Robotic Action Representations Contrastive Representation Regularization for Vision-Language-Action Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.