Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-30T22:11:21.596611Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 100 of 118 outbound references and 1 inbound Pith citation observation for arXiv:2605.12369.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-30T22:11:21.596611Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-29T06:57:41.245418Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-06-29T07:23:13.398570Z
100 of 118 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation dc2a593c-80c5-40ad-b0bc-599497373c1e · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 907794c5-0769-4fb1-8f13-fee97b8c8c09 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Qwen3-VL Technical Report
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f65455d5-38c7-49b0-99e0-6171b4ea0117 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization PaliGemma: A versatile 3B VLM for transfer
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation abda6cf9-8281-44b4-936c-7d0965dcaf16 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization 3d cavla: Leveraging depth and 3d context to generalize vision language action models for unseen tasks
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c382d61c-c90a-406c-877b-012fe65853ff · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 36a9286f-4481-4d43-a7cf-b7108fd59050 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization In9th Annual Conference on Robot Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5aca3279-be2c-4130-8f72-277c3af3d875 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization InRSS
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation bf6cb1ad-f015-40b1-9f32-f64204dfd787 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Real-Time Execution of Action Chunking Flow Policies
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 66cd4033-5063-4a39-8746-44f5c99b5beb · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization RT-1: Robotics Transformer for Real-World Control at Scale
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0205eb4e-a35b-4fbc-ab84-9bf1061acbd5 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization UniVLA: Learning to Act Anywhere with Task-centric Latent Actions
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4fcdae89-cf79-434d-91d1-eb06871ca981 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization WorldVLA: Towards Autoregressive Action World Model
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation fe976080-1848-4161-8805-2ac4693fa229 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization STORM: Slot-based Task-aware Object-centric Representation for robotic Manipulation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 3e0a94e2-c7fd-49de-b8d3-70363e7a1287 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Unified diffusion vla: Vision-language-action model via joint discrete denoising diffusion process
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c1c9e977-2617-446a-9622-4c579efbbdff · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f7ac345a-fff2-4351-9276-ad895e2992b7 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0eab96a9-07cf-4109-9690-748843fe3d67 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Moe-dp: An moe-enhanced diffusion policy for robust long-horizon robotic manipulation with skill decomposition and failure recovery
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 20fa5fa8-2a7d-4e44-9604-17af9a37af51 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Diffusion policy: Visuomotor policy learning via action diffusion.The International Journal of Robotics Research, 44(10-11):1684–1704
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 80bc0e30-54ac-474b-a87c-e3285ae2f55d · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization AimBot: A Simple Auxiliary Visual Cue to Enhance Spatial Awareness of Visuomotor Policies
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation daa4452f-bd03-4453-a140-ff5f26d2f238 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization RoboNet: Large-Scale Multi-Robot Learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1f3b432a-a490-4b19-9950-73688370d0e8 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Causal confusion in imitation learning.Advances in neural information processing systems, 32
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e57183a7-40e2-4b37-981a-ad9601eefce0 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6a89c519-2d95-49e3-8222-22cb4b3e4e34 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Palm-e: An embodied multimodal language model
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 62b3164c-147c-4f20-b68d-4fb83836f7eb · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization BridgeData V2: A Dataset for Robot Learning at Scale
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8874bb89-180b-43d9-a346-e09e726c0e52 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Interleave-vla: Enhancing robot manipulation with interleaved image- text instructions
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 98673013-3de8-477c-aed5-d633874f3ff1 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Peafowl: Perception-enhanced multi-view vision-language-action for bimanual manip- ulation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 01391be9-7701-4855-9166-e2133a39f239 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Learning skills from action-free videos
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5e721c90-c723-4252-bba0-12935dd8fd43 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization LIBERO-Plus: In-depth Robustness Analysis of Vision-Language-Action Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e45e44cd-981e-4a39-85c6-90e86a0ccfd3 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Imagenet-trained cnns are biased towards texture; increasing shape bias improves accuracy and robustness
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5685341c-247a-4456-8774-86bd645646a9 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Shortcut learning in deep neural networks.Nature Machine Intelligence, 2 (11):665–673
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2a215a6f-e48c-4746-9891-8fc8fee02c31 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Octo: An open- source generalist robot policy
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ffddb4f1-8409-4a8a-8282-ef865607fb61 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Point Policy: Unifying Observations and Actions with Key Points for Robot Manipulation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8a2be12c-a90e-4163-b932-3ace62d835b1 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization In: 2025 IEEE International Conference on Robotics and Automation (ICRA), pp
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9793bf6b-f90b-4a80-8771-821d4e32338f · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Lora: Low-rank adaptation of large lan- guage models.ICLR, 1(2):3
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 41cabbca-8318-45c0-9878-268868b632ad · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization arXiv preprint arXiv:2601.11266 (2026)
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 385c36c4-b23e-4695-929c-3085d1fbd09c · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Rekep: Spatio-temporal rea- soning of relational keypoint constraints for robotic manipulation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 41c0716e-5d90-487b-b39c-932016bf342f · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization arXiv preprint arXiv:2601.03782 (2026)
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4afbb126-945a-4ee8-afd7-f2e474b63c54 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f27f4d53-0503-475b-9945-4b8f4c8ed231 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Rlbench: The robot learning bench- mark & learning environment.IEEE Robotics and Automation Letters, 5(2):3019–3026
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 91c22306-6ef1-4d04-aabb-822e51183444 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Galaxea Open-World Dataset and G0 Dual-System VLA Model
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 813fccc6-3acb-4a32-8fe8-ece4f6c704e5 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 23151c02-dbba-4a7e-a20a-c68fe51985fb · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization VIMA: General Robot Manipulation with Multimodal Prompts
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 964b6e31-659f-406c-b6a4-6dce8e440a36 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1194d431-83ae-4828-8a93-0d3773786eca · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation cb9c0ab5-33aa-4577-ad49-d5bfd8952f85 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Openvla: An open-source vision-language- action model
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9a4e08dd-7356-4480-9b19-d54ee8b7d561 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Trace- gen: World modeling in 3d trace space enables learn- ing from cross-embodiment videos
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f7e82ceb-08de-4048-895e-5aa71e945c92 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Spatial forcing: Implicit spatial repre- sentation alignment for vision-language-action model
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8fc40c16-c0f9-4902-b206-f2cfbc94713c · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization H2r: A human-to-robot data augmentation for robot pre- training from videos
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c588b211-cab8-4135-aaa8-63d180fc080a · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Language-guided object-centric diffusion policy for generalizable and collision-aware manipulation
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f2e55135-e957-4b1b-8334-404550414357 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Coa-vla: Improving vision-language-action models via visual-text chain-of- affordance
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e5a625ed-7125-469c-b74b-6676af214999 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation dc46ab99-7999-47bc-8701-123f54108435 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Bridgevla: Input-output alignment for efficient 3d manipulation learning with vision-language models
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 11ae13b6-56c3-4c1a-871a-68da35c742d4 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Posa-vla: Enhancing action generation via pose-conditioned anchor attention
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 3b04815d-efec-4187-a60d-6ead98fb6da2 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Skilldiffuser: Interpretable skill planning for latent diffusion-based manipulation
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ed43a84d-712b-4023-b276-d850142d72b1 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ca914e65-27ce-432c-bed6-778cf7540a44 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization ManiCM: Real-time 3D Diffusion Policy via Consistency Model for Robotic Manipulation
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ded24e10-4128-4875-80d8-7abad477e59d · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Depth Anything 3: Recovering the Visual Space from Any Views
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8a29bcc9-0e99-4b03-914a-a10e5f48866c · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Constraint-preserving data generation for one- shot visuomotor policy generalization
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0259c7f2-4635-43f3-bae9-0feeff25ac83 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Libero: Benchmarking knowledge transfer for lifelong robot learning.Ad- vances in Neural Information Processing Systems, 36: 44776–44791
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4cae5174-5271-46f1-a013-9ff83c5e5f5e · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Rdt-1b: a diffusion foundation model for bimanual manipulation
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 49a110b2-f129-42e4-ba87-ffc0bdd773d9 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Hierarchical diffu- sion policy for kinematics-aware multi-task robotic ma- nipulation
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d2fe6d8f-68c0-4e08-82b5-60bc7443f24e · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization arXiv preprint arXiv:2510.26742 (2025)
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8b13eaf6-778b-4801-a7ad-cfce6ebfcbe0 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Roboturk: A crowdsourcing platform for robotic skill learning through imitation
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 339471e0-9a20-4f9a-960c-50f5a63e568a · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Calvin: A benchmark for language- conditioned policy learning for long-horizon robot ma- nipulation tasks.IEEE Robotics and Automation Let- ters, 7(3):7327–7334
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9be1d31f-8128-44da-9b4f-b5941e5a0b82 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization R3M: A Universal Visual Representation for Robot Manipulation
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b9cae7b2-75fc-445a-9f48-faeea0ba6504 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Vo-dp: Semantic-geometric adaptive diffusion policy for vision- only robotic manipulation
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 08ccf904-cfac-4c24-872c-2f44a7dbd42f · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Open x-embodiment: Robotic learning datasets and rt-x models: Open x-embodiment collabo- ration 0
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5ea24653-d07f-4a96-a24e-b9b3554ebe16 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Omnimanip: Towards general robotic manipulation via object-centric interac- tion primitives as spatial constraints
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 34755786-5e2d-44f1-a6e3-5db2611a8ba6 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 3e860926-4124-47f6-9d5f-98c7f60f4118 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization GeoPredict: Leveraging Predictive Kinematics and 3D Gaussian Geometry for Precise VLA Manipulation
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6bea0502-d20c-421f-b337-877abb3cecf2 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Spatialvla: Exploring spatial representations for visual-language-action model
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 94d9e78a-f3b4-4ea6-ac23-7ca1baafa546 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization SAM 2: Segment anything in images and videos
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 51dbcf3e-eac4-4d55-99a9-fbd9bc24c6c9 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Grounded sam: Assembling open-world models for di- verse visual tasks
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f0261b67-3f12-40b1-95f3-0cd89980cdf9 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization CLARE: Continual Learning for Vision-Language-Action Models via Autonomous Adapter Routing and Expansion
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation afb30d02-6fdf-4379-bac1-0ad5836eb951 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Expertise need not monopolize: Action-specialized mixture of experts for vision-language-action learning
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 02fd77d4-1ee8-499e-9500-af46074635e6 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f744ab31-d85d-42e9-b602-e1a92a3b1570 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Interactive Post-Training for Vision-Language-Action Models
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 424f01bf-4196-4d42-9636-179e6ebaa185 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Gemini Robotics: Bringing AI into the Physical World
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 856ce5a5-da23-4465-af5c-050e10ea4607 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8f495e66-b840-4698-ae55-51a91a413d0e · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Attention is all you need
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 59576cda-2c6b-49f8-8301-01283c491aef · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Bridgedata v2: A dataset for robot learning at scale
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e56ff89c-c3b8-4864-9d4c-a5edd6a83c77 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Aerial Tensile Perching and Disentangling Mechanism for Long-Term Environmental Monitoring
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0869ef57-6046-4d22-b3fd-92bac388203d · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization 3D Diffuser Actor: Policy Diffusion with 3D Scene Representations
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0e35531d-9595-42b1-88bd-421b28ba7fb2 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Vla-adapter: An effective paradigm for tiny-scale vision-language-action model
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6a8e3e34-85de-437e-998b-533c131989ae · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b4de4511-e01f-4789-b99e-83ddd771bac2 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization dvla: Diffusion vision-language-action model with multimodal chain-of-thought.arXiv preprint arXiv:2509.25681
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b66ab7d4-3561-4fae-a941-c696d553740e · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1dbd3003-1750-4132-a549-57ddf7480848 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Tinyvla: Towards fast, data-efficient vision-language-action models for robotic manipulation.IEEE Robotics and Automation Letters
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation cde63f61-20ea-48e5-a2de-24e30f251151 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Af- forddp: Generalizable diffusion policy with transferable affordance
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2a0facf7-99d6-4d0d-bc8a-f0bc42d61b7b · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9d7b9c21-80f2-43ea-be59-0829da473950 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Point what you mean: Visually grounded instruction policy
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 326f7b51-42f0-4f96-a899-adb782c0fcc2 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Meta-world: A benchmark and evaluation for multi- task and meta reinforcement learning
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d6a3858a-c0a8-44ca-b3e0-d605277de5a6 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Depthvla: Enhancing vision-language-action models with depth-aware spatial reasoning
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a81d98b1-e7d0-40c8-92b7-7a122807adc8 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization 3d diffusion policy: Generalizable visuomotor policy learning via simple 3d representations
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d050c9eb-8992-4aea-ae9b-536123d17ea7 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Adding conditional control to text-to-image diffusion models
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c3714327-8f08-43f1-aaf2-60a7276a7e08 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Dreamvla: a vision- language-action model dreamed with comprehensive world knowledge
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4265e98b-0429-462b-aca4-56206e8184c2 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Mos-vla: A vision-language-action model with one-shot skill adaptation
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2464386f-1ff1-42d0-83e5-a098f9a40105 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5cb8089d-27e2-4331-8fe5-61276a61e83e · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization arXiv preprint arXiv:2512.24673 (2025)
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ea800739-84e0-446d-bba8-b640f159be34 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization 3d-vla: A 3d vision-language-action generative world model
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a4a46732-8e89-4f04-a48b-0583ab9006b2 · outbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies
Reference 101
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ca51edbe-f8d4-4f05-ba83-c7afb2962e2e · inbound
ELAN4D: Embodiment-Centric 4D Supervision for Vision-Language-Action Models via Plug-and-Play Adaptation GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.