Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-12T07:33:25.188358Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 100 inbound Pith citation observations for arXiv:2411.19650.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-12T07:33:25.188358Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T20:10:13.556241Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T23:07:48.078477Z
78 of 78 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation bdad2685-5d3e-4b25-8538-2b25b201222c · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a10b3bb1-ed70-422d-9a01-09135e8e5cb1 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation GPT-4 Technical Report
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a328a0f5-29de-43d7-8740-83e84761cac0 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a9405621-2096-41ca-8f50-d507567ba0ee · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Hydra: Hybrid robot actions for imitation learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7ca60efb-2d6a-426f-a05c-562b94407f37 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c8ecfdb4-8137-4178-babc-53f856d25a4e · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation RT-1: Robotics Transformer for Real-World Control at Scale
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 929fd749-adcc-45c6-a506-7b00056469fe · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d935b006-cfff-42b3-9c10-4f7454382896 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Language Models are Few-Shot Learners
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d68ad38f-b0e2-44b4-94c2-a57b93c02d11 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation The ycb object and model set: Towards common benchmarks for manipula- tion research
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 939f2fcf-40a1-4df2-8026-e1512e961037 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7d36be1d-c2e3-4e2b-94d2-20ce17f303d7 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Berkeley UR5 demonstration dataset
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f145f177-8fdd-43b4-83eb-23cbd2598287 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation PaLI-X: On Scaling up a Multilingual Vision and Language Model
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 52d6193b-d590-47fd-8026-a2f839625bfb · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bf677c2b-a41d-4af2-b26a-a3ad4a2591ac · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Diffusion policy: Visuomotor policy learning via action dif- fusion
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d25037d1-e163-4827-a927-102195fe502c · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Analysis and observations from the first amazon picking challenge
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f008091f-c63f-422f-bba6-f6574dd46ef4 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation From Play to Policy: Conditional Behavior Generation from Uncurated Robot Data
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 91303da1-9952-483f-9594-c14c5fa0bc11 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b84486d2-215e-4103-a69d-85fbae207ef6 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation High Fidelity Neural Audio Compression
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 24be5b98-0f26-4938-8929-e01405908ce7 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Bridge Data: Boosting Generalization of Robotic Skills with Cross-Domain Datasets
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 51aa9ab1-8662-4175-af38-c302287a4b89 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation The brain basis of language process- ing: from structure to function
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7566d190-7abc-410d-bcba-fb440f316824 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c9459b40-ec60-4360-8a12-f0addd15cbc6 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Classifier-Free Diffusion Guidance
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 67c35d05-8266-46cc-9d15-80efad97d79f · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Diffusion Transformer Policy
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation eb23a595-7923-41c4-96d8-2829cc545675 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Neu- roanatomy, visual cortex
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 73bee261-1e61-44d5-bcca-91b394ea002f · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Bc-z: Zero-shot task generalization with robotic imitation learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7efdcb50-c867-4e03-94a2-6a2acbba6ce9 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2bb8ec2b-6c06-4786-92cf-32f81a4804e2 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Prismatic VLMs: Investigating the Design Space of Visually-Conditioned Language Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation da1562b6-609c-42bd-a9c5-36f0fcc0088a · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 21a8ddec-88b3-4905-af0f-0d39e0f187e7 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation OpenVLA: An Open-Source Vision-Language-Action Model
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9c3211c9-ef97-45d2-bef3-fdee0e271d90 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation The darpa robotics challenge finals: Results and perspectives
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0350bb56-cf30-4593-a6f5-ce0ff128479c · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Vision-Language Foundation Models as Effective Robot Imitators
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bda642d4-e433-434d-8fc2-2fd4ca3039a5 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Evaluating Real-World Robot Manipulation Policies in Simulation
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e5b52590-9bb6-4238-80e3-8cadf796d462 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Vision-language foundation models as effective robot imitators
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4276f456-64a7-4f79-864d-4ff7c93e2cdd · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Visual instruction tuning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a2d5a6fc-f9ee-4d0b-934f-8567f38a26c4 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Robot learning on the job: Human-in-the- loop autonomy and learning during deployment
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3c17883e-c440-48ed-82a5-cf8c02d54682 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Visual instruction tuning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ba9e594a-06b1-4aec-a0d6-fd4607716c92 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 61928e4d-253a-4c78-9541-ded92475db90 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Multi- stage cable routing through hierarchical imitation learning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bad407ba-9c8b-428b-9f5b-b932c445796f · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation FMB: a Functional Manipulation Benchmark for Generalizable Robotic Learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 84700920-7c13-4375-84bf-c96dc51433f1 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Interactive language: Talking to robots in real time
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fef41e66-8054-4bf3-88ed-952b03dd747b · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Scaling robot supervision to hundreds of hours with roboturk: Robotic manipulation dataset through human reasoning and dexterity
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1c81b962-d6e3-43db-8602-303796cda353 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Grounding language with visual affordances over unstruc- tured data
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c537171e-d526-47bb-9218-5647c024c325 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Struc- tured world models from human videos
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4afa4b1f-f088-4740-8347-e3560d33114c · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation R3m: A universal visual repre- sentation for robot manipulation
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 64ead727-bb12-401c-8a28-758af0ea720b · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Learning and retrieval from prior data for skill- based imitation learning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9155c0ba-a5e2-4f5b-adaa-7a1a74d07f25 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Improved denoising diffusion probabilistic models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b6a80201-e81c-4c84-ae3a-a8b2b45b3f0a · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Open X-Embodiment: Robotic Learning Datasets and RT-X Models
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f8364b6c-e6c2-4abc-9337-3486a98f7904 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c1108305-3206-42f0-a79a-96c10a057f9b · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Imitating Human Behaviour with Diffusion Models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7ed1e77b-176f-474f-9ded-f7db409e4ed9 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Scalable diffusion models with transformers
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6504951c-d04e-4fe5-a51d-074c8032b53a · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Shared Control Templates for Assistive Robotics
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e6350ad4-357d-4968-87f0-1e63627388fe · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Goal-Conditioned Imitation Learning using Score-based Diffusion Policies
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a3555736-4d3f-4625-869c-82cd74ec6237 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Latent plans for task ag- nostic offline reinforcement learning
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d9feb6a8-1fdd-403d-a0e2-1ea6b04206e3 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Multi- resolution sensing for real-time control with vision-language models
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dad853bb-44d7-453a-b037-1740021f637c · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation On bringing robots home
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c600602f-da84-40da-8297-fed8414fc894 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation MU- TEX: Learning unified policies from multimodal task spec- ifications
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a598f94d-4bec-4f05-93b3-70ec06e1ffba · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Perceiver- actor: A multi-task transformer for robotic manipulation
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9c889824-e32a-4ebd-818b-9447e343f0e0 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Denoising Diffusion Implicit Models
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 136df55e-9bbd-4501-9422-3109b549eb85 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Open-world object manipulation using pre-trained vision-language models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a4de158d-d32c-4e43-ac4a-cd2d6f88bc37 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Gemini: A Family of Highly Capable Multimodal Models
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dc12dba8-a312-40fe-ae79-7c9701e7d606 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Octo: An Open-Source Generalist Robot Policy
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2488e747-09f6-4f1d-8a75-befb7e54c0e3 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation LLaMA: Open and Efficient Foundation Language Models
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a036d4c4-3bcf-4f27-bac1-aae2db5efc80 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation df40284e-3bf2-40e4-8dc8-320c428a655b · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Neural discrete representation learning
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f8065990-ddcf-4541-86c1-3e97d334e754 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Bridgedata v2: A dataset for robot learning at scale
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f8ff1e47-d93b-41d2-a82e-d5c3acbcd3d0 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 620e002c-2485-4775-a40d-c18df9282237 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5f02ae89-c0be-4da8-919e-6d74475607d5 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Unleashing large-scale video generative pre-training for visual robot manipulation
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 18e3cea8-af13-42ab-8011-4ed01be40baa · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation ucsd kitchens Dataset
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4c528fd5-a176-4975-9df2-3467b44afdcf · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Physi- ology, motor cortical
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b73c6851-bd68-4f63-9cf0-aff0883c8baa · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 795d8f4c-b9be-48f3-a807-57c729e3ecfa · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Soundstream: An end- 11 to-end neural audio codec
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 89d48fc8-5ef0-4550-a0f9-8b71e5aa984c · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Sigmoid loss for language image pre-training
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1abb2c93-fe40-4382-8c29-4fa1061f9842 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0cad2161-d4b7-4ae9-8041-726127041dbe · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Train of- fline, test online: A real robot learning benchmark
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7d9f5cc2-d716-40c1-9604-9baefe8cbb82 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Fanuc manipulation: A dataset for learning-based manipulation with fanuc mate 200id robot
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 95a1be72-d529-4d67-8bc3-5d5986b4263d · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation Vi- ola: Imitation learning for vision-based manipulation with object proposal priors
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4a1b8fc8-4a49-4769-890b-247060848ab6 · outbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation pick Coke can
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a27b1738-825c-45dd-9294-872b061cb96f · inbound
A Survey on Vision-Language-Action Models for Embodied AI CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 124
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ec5f6f61-d0e8-4a6b-9991-89f108183ab6 · inbound
RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c0784892-5359-4eb4-a0a5-9822fc377a53 · inbound
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6d7b5753-3a88-4cf9-ab63-0a5e52d7584a · inbound
Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 70a55dd5-a052-4f79-9548-f2001a1cfd4c · inbound
HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9dfaf2b2-2e11-45fb-91d2-6a8693942608 · inbound
$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f1fb35b8-f6fa-433d-9e14-8971a8802d87 · inbound
GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 02548de6-f486-422f-8613-0afab74b1cbc · inbound
RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e9881423-aec0-44e2-8a90-2c4b36631448 · inbound
A Survey on Vision-Language-Action Models: An Action Tokenization Perspective CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 258
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 64afa7fd-d989-4831-813c-e537b3a8d5b4 · inbound
DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d08b26d3-4b3e-40d5-8c62-95b628d49435 · inbound
AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d4222ffe-a0d0-430f-b66f-a968ee9b1f6e · inbound
GR-3 Technical Report CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 94de6eee-62c7-4672-954c-2ffe4c28fdd8 · inbound
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 379ce469-9b70-47ae-b86b-f13114d76289 · inbound
Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 128
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5499412b-8750-4524-8f5c-eb674b7e6d91 · inbound
RoboChemist: Long-Horizon and Safety-Compliant Robotic Chemical Experimentation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d935dbba-20ba-4735-800f-079d712a1010 · inbound
SQAP-VLA: A Synergistic Quantization-Aware Pruning Framework for High-Performance Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80494ba4-6565-40e0-a562-7a1910025732 · inbound
Ask-to-Clarify: Resolving Instruction Ambiguity through Multi-turn Dialogue CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bf2ccd7-86ba-40a2-9672-9a0f50fc6633 · inbound
RobustVLA: On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ec6e6c9-6307-4f76-b434-01caf8d000ca · inbound
FailSafe: Reasoning and Recovery from Failures in Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e592b8ee-39dd-44e2-8d92-2ef177885807 · inbound
Contrastive Representation Regularization for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89caafa7-a060-4bfc-9977-7146770f40de · inbound
R2RGEN: Real-to-Real 3D Data Generation for Spatially Generalized Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 23db38a5-6796-4753-a499-10603157dce8 · inbound
AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dd7dfac0-5aee-4ef5-85e6-b94a5a7173eb · inbound
Bridging the Semantic-Action Gap in Visual Token Pruning for Efficient VLA Inference CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25861ae6-a327-4ebe-a1c4-e0075c521b12 · inbound
SPEAR-1: Scaling Beyond Robot Demonstrations via 3D Understanding CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8f219e04-3574-4462-b777-6eb5ff28b57e · inbound
ActDistill: General Action-Guided Self-Derived Distillation for Efficient Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7c19f252-ddaa-487b-baa0-9c347290fcd0 · inbound
AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0d09a04f-9252-4e29-b2c9-4f203e45552c · inbound
Mixture of Horizons in Action Chunking CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfa15989-a43f-40c4-9fd3-1a001a401664 · inbound
HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 97dccf5f-d256-472f-b89a-48a53538c6ff · inbound
StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 357484ab-3dc4-4dbf-87b0-9427706542bd · inbound
Learning to Feel the Future: DreamTacVLA for Contact-Rich Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1d9dc54b-437c-4ba3-bebc-ea0e974b9d6e · inbound
Learning to Feel the Future: DreamTacVLA for Contact-Rich Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d38c2a55-30ef-4e5f-a63e-cb167be56822 · inbound
MobileManiBench: Simplifying Model Verification for Mobile Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6214642d-950d-46ac-a6d1-7c53307f2485 · inbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 204061ce-907d-43ba-9c62-c9c7c6148957 · inbound
TwinRL: Digital Twin-Driven Reinforcement Learning for Real-World Robotic Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a5f896ce-bf66-40c6-96e1-0b15b85664ef · inbound
ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 69740d5b-f5af-4cf8-a247-527255dc14e6 · inbound
Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Efficient Robotic Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 605b2227-62e3-416a-9137-7f91d5a6a167 · inbound
UniLACT: Depth-Aware RGB Latent Action Learning for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 269df38e-9a8f-4dc3-a11a-03bba2072aaa · inbound
RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2e0f352-eaa6-44a3-8345-87c49b874441 · inbound
RoboMME: Benchmarking and Understanding Memory for Robotic Generalist Policies CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d71669b-a873-44c3-8a8d-b3e10713ce77 · inbound
Choose What to Observe: Task-Aware Semantic-Geometric Representations for Visuomotor Policy CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2687cd36-760a-4f77-b6a1-36810ebfd7ea · inbound
AR-VLA: True Autoregressive Action Expert for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 621e5b19-016e-47cd-bb3c-41389e7c8d86 · inbound
vla-eval: A Unified Evaluation Harness for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6a77913b-c5f6-409c-86a6-2271fc550a74 · inbound
OxyGen: Unified KV Cache Management for VLA Inference under Multi-Task Parallelism CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c2f83144-a14d-40da-a357-b1b76f4569e3 · inbound
RoboECC: Multi-Factor-Aware Edge-Cloud Collaborative Deployment for VLA Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 93d23a39-77c5-4eef-8406-535997a96b5c · inbound
VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 27da5bd5-abe3-4225-984f-1a0e694a3056 · inbound
DIAL: Decoupling Intent and Action via Latent World Modeling for End-to-End VLA CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4bdba29e-5207-4985-a5d2-61fc395d6b0d · inbound
BiCoord: A Bimanual Manipulation Benchmark towards Long-Horizon Spatial-Temporal Coordination CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4a664327-97fa-4bc3-bc67-5ee2a948409f · inbound
ViVa: A Video-Generative Value Model for Robot Reinforcement Learning CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dfe11aac-4bb4-4ccf-921c-573478b4986d · inbound
Activation Steering for Aligned Open-ended Generation without Sacrificing Coherence CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6707b904-7c09-4902-8169-9c4919e93b82 · inbound
ComSim: Building Scalable Real-World Robot Data Generation via Compositional Simulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 07d74926-32b0-45a4-a490-6635cdf4a33b · inbound
DA-PTQ: Drift-Aware Post-Training Quantization for Efficient Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8f585254-030f-4ae2-bd7e-fda12d0e6807 · inbound
${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c45aa43b-8f0e-4a04-bd90-f9fcf59b1b88 · inbound
ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f31d60cc-5fa1-4afb-b8bd-b7b680dd5f3c · inbound
ST-$\pi$: Structured SpatioTemporal VLA for Robotic Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 948a5e35-2362-45ce-a839-a7ea9b88b9b3 · inbound
Mask World Model: Predicting What Matters for Robust Robot Policy Learning CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5c08c2c7-fee6-49ca-b403-41a23b5a2ded · inbound
PRTS: A Primitive Reasoning and Tasking System via Contrastive Representations CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 07f9ad9f-4ea7-449c-8bfb-f0928ed58e89 · inbound
LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 93557888-4f12-4824-8f50-9563626760ef · inbound
LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9edd534a-2cce-41dc-9b8f-c36e3dac82c1 · inbound
Being-H0.7: A Latent World-Action Model from Egocentric Videos CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 42599b07-c3eb-4c0b-9b31-3d02dfb79048 · inbound
VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ae6ce2d8-bf45-4619-9bc9-e132bbeba89c · inbound
Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0f171601-d7fd-40cc-b4a2-c69f1a79ddda · inbound
RLDX-1 Technical Report CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6724b243-19e6-4503-b68b-bb428aa4cde3 · inbound
RLDX-1 Technical Report CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 25cfe5ad-7f94-4a79-8368-92dbbbe4039d · inbound
ConsisVLA-4D: Advancing Spatiotemporal Consistency in Efficient 3D-Perception and 4D-Reasoning for Robotic Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d1b4786e-4d04-43e0-a9c1-6e4a93580a56 · inbound
TriRelVLA: Triadic Relational Structure for Generalizable Embodied Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c83e848b-98ea-4814-a5a1-782200fac122 · inbound
OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 64e7d716-a1ac-4fed-8b1b-dc5b30601863 · inbound
AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 357629e7-9cd4-4b9a-9ae0-b57766986edb · inbound
AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b1253d25-9c34-4b1e-93c3-7625cbea20d6 · inbound
Escaping the Diversity Trap in Robotic Manipulation via Anchor-Centric Adaptation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6719f627-0f5b-4ec2-aecd-3486be808042 · inbound
ForgeVLA: Federated Vision-Language-Action Learning without Language Annotations CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 691951da-9af1-4738-bcad-72d9c9b8fdb5 · inbound
One Token Per Frame: Reconsidering Visual Bandwidth in World Models for VLA Policy CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 37ddafba-0a1a-4bb8-92be-315466bcab0b · inbound
One Token Per Frame: Reconsidering Visual Bandwidth in World Models for VLA Policy CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 73b32afc-3b72-4fed-bef6-2d8b957e73b0 · inbound
One Token Per Frame: Reconsidering Visual Bandwidth in World Models for VLA Policy CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 88ff709f-1491-4697-b1fe-c9f88588626c · inbound
Attention Itself Could Retrieve.RetrieveVGGT: Training-Free Long Context Streaming 3D Reconstruction via Query-Key Similarity Retrieval CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 94a70600-bc5e-4ef3-87c4-0944b1245702 · inbound
Retrieve-then-Steer: Online Success Memory for Test-Time Adaptation of Generative VLAs CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bc576819-4022-4039-9e16-42084e59754c · inbound
Retrieve-then-Steer: Online Success Memory for Test-Time Adaptation of Generative VLAs CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4259af01-e8ad-43a1-b06e-dbfb31c362e1 · inbound
ALAM: Algebraically Consistent Latent Action Model for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ef9acaa9-a358-4a16-86d5-af225a980fe8 · inbound
ALAM: Algebraically Consistent Latent Action Model for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 12b7695b-a460-4cef-aacd-ad59a1cffc09 · inbound
RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 82a1eb27-0d2e-41b7-a33a-a8f5af80db32 · inbound
PriorVLA: Prior-Preserving Adaptation for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0f748f54-4a10-4a34-b3cb-b7e1a97ddc5e · inbound
HarmoWAM: Harmonizing Generalizable and Precise Manipulation via Adaptive World Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ebcf0aee-58ff-4df0-8f2b-43fb44824646 · inbound
Nautilus: From One Prompt to Plug-and-Play Robot Learning CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 95b53f66-8db5-4743-9071-0ab1dc07222a · inbound
Nautilus: From One Prompt to Plug-and-Play Robot Learning CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad5bd448-e4c7-4b4c-b9e7-b48bb1a7632c · inbound
See What Matters: Differentiable Grid Sample Pruning for Generalizable Vision-Language-Action Model CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fef89edf-faa5-4a28-ae26-b5dc647fc156 · inbound
Learning Action Manifold with Multi-view Latent Priors for Robotic Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 10ba2a1d-cdfb-46c9-9ca9-12c50e46d99c · inbound
BlockVLA: Accelerating Autoregressive VLA via Block Diffusion Finetuning CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a2a914b1-053b-489d-b470-07912a04480e · inbound
AttenA+: Rectifying Action Inequality in Robotic Foundation Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 498dc31b-ca37-4578-a8f7-db801df05129 · inbound
AttenA+: Rectifying Action Inequality in Robotic Foundation Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6fbdd620-dbe4-49b4-b666-a77330246e70 · inbound
Guide, Think, Act: Interactive Embodied Reasoning in Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ab9a7db4-067e-4973-8e44-01b03e2fbb69 · inbound
Guide, Think, Act: Interactive Embodied Reasoning in Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d33b3de2-4ea3-4822-851e-42c751257825 · inbound
FrameSkip: Learning from Fewer but More Informative Frames in VLA Training CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 15a093e0-c38f-4079-aeec-b38dd886283c · inbound
IntentVLA: Short-Horizon Intent Modeling for Aliased Robot Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e6bae885-bf40-4605-b46e-d91184b38a36 · inbound
IntentVLA: Short-Horizon Intent Modeling for Aliased Robot Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cab310b9-d243-4fd0-a165-0b0d9a99ac7f · inbound
PhysBrain 1.0 Technical Report CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 61ae4a8d-066b-45b2-93e2-eef339ae3071 · inbound
DyGRO-VLA: Cross-Task Scaling of Vision-Language-Action Models via Dynamic Grouped Residual Optimization CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4ad8c2ca-eaff-4a4b-a800-2b715f7c9aa3 · inbound
ManiSoft: Towards Vision-Language Manipulation for Soft Continuum Robotics CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fa106154-94f9-4d79-9c29-53b48a0b9986 · inbound
Rethinking Muon Beyond Pretraining: Spectral Failures and High-Pass Remedies for VLA and RLVR CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6eca272d-c2f6-4805-89ee-bb29627eaa98 · inbound
PointACT: Vision-Language-Action Models with Multi-Scale Point-Action Interaction CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2084485e-40bc-4efa-bde5-20ce469339f3 · inbound
ActQuant: Sub-4-bit Action-Guided Quantization for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1a03a248-e5b8-43b6-91ce-19f4c831ff50 · inbound
Afford-VLA: Action-Aligned Visual Planning via Internalized Affordance CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.