Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-20T12:49:02.418663Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 77 of 77 outbound references and 1 inbound Pith citation observation for arXiv:2605.17517.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-20T12:49:02.418663Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T00:35:34.154516Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-11T00:35:34.361224Z
77 of 77 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a80f3fb3-6906-453b-962a-2a7991877373 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Flexible robotic hand harnesses large deformations for full-coverage human-like multimodal haptic perception
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e42408a2-f256-4ec3-943d-0c76a2e70be3 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Language-conditioned affordance-pose detection in 3d point clouds
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0a1a5556-9931-449e-aa6f-36b0dca8deb9 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Uad: Unsupervised affordance distillation for generalization in robotic manipulation
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a4e492e0-bb0c-481c-b44b-27dc38e0c981 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Affordancenet: An end-to-end deep learning approach for object affordance detection
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation dba22695-6937-4e36-a271-9c7681c59e48 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Omnimanip: Towards general robotic manipulation via object-centric interaction primitives as spatial constraints
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e00a138e-3384-481c-a54d-b12839579ef9 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment A0: An affordance-aware hierarchical model for general robotic manipulation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ae510d2a-34a0-4984-94a0-27c099ba9fbf · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Manipvqa: Injecting robotic affordance and physically grounded information into multi-modal large language models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3462b2f8-4cd0-4b26-abb0-76a06ad58603 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Robots pre-train robots: Manipulation- centric robotic representation from large-scale robot datasets
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a161325b-5166-4c75-9223-9bde3d11b005 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Tars: Tactile affor- dance in robot synesthesia for dexterous manipulation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0ab42483-c3c2-48fa-aa66-0363f8c7992e · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Sa-dem: Dexterous ex- trinsic robotic manipulation of non-graspable objects via stiffness-aware dual-stage reinforcement learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 75f4f201-3e0b-41c0-a71e-84d5f40f79e4 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Rt-1: Robotics transformer for real-world control at scale
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1d4a99c3-927b-4551-9858-c4bfcd7bd42e · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Rt-2: Vision-language-action models transfer web knowledge to robotic control
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d674d6fd-e408-4bb7-be12-5ca75bf7f1b0 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 51a453dd-e221-4464-96f1-2f1c90c42aff · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment OpenVLA: An open- source vision-language-action model
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9cc0998f-da3c-4191-a37a-803b40ec33a8 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8b8f72f0-26ba-44ad-b8d8-8c8d7b61e4ff · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Vla-jepa: Enhancing vision-language-action model with latent world model
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 59133584-fb53-4ed5-aca1-1e0ad34bd7db · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Reconvla: Reconstructive vision- language-action model as effective robot perceiver
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 432e511c-7e91-4b79-ab64-d558f7ca87dd · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Spatial forcing: Implicit spatial representation alignment for vision-language-action model
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6ab08b49-dd82-4e33-bc51-c421305ab70a · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Rt-affordance: Affordances are versatile intermediate representations for robot manipulation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0145e820-13d5-4025-9184-55029ff32c71 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Moka: Open-world robotic manipulation through mark-based visual prompting
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8e10241c-45ca-45e2-bd69-8b179736809d · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Knowledge enhanced bottom-up affordance grounding for robotic interaction
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 815ad580-1760-4fd9-83fc-2e4cd411505c · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Uncertainty-aware state space transformer for egocentric 3d hand trajectory forecasting
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a79f06d0-d910-4c5b-a070-bbb73373db87 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment arXiv preprint arXiv:2507.10672 , year=
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f2c43b1b-1d17-4e03-ab93-cb4bbc6077ac · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Coa-vla: Improving vision-language- action models via visual-text chain-of-affordance
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3982747b-e186-4606-bb30-e8756490d2de · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 127fac3d-d8a6-4695-bb87-d8394e089365 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment GPT-4 Technical Report
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 26c62ed2-df55-4b25-a3e6-da110d70e524 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Qwen3-VL Technical Report
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 20807dcc-7c92-4e2d-88d1-eff5233743a0 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 213ba8ba-50af-4292-a7bf-99076b11954b · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Open x-embodiment: Robotic learning datasets and rt-x models: Open x-embodiment collab- oration 0
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a56228ed-2ed6-4908-a8be-bd4e9c9ad57b · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Cot-vla: Visual chain-of-thought reasoning for vision-language-action models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2c17658d-5a34-4d27-9ddc-7639f6b5566d · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e9335216-3146-4cdf-b019-fae644691dcc · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bbf4e9fb-371b-497c-bc86-212e66694060 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment $\pi^{*}_{0.6}$: a VLA That Learns From Experience
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d5bc91fc-6de5-4798-8560-472502e8c796 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Ig-rft: An interaction-guided rl framework for vla models in long-horizon robotic manipulation
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d7845171-13e2-4158-8031-77b30695902b · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fa1698af-058c-4e4e-b48d-d132b2766042 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment The ecological approach to visual perception
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2c3ef289-13ec-48d2-8d79-ba01854eafb5 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Affpose: An integrated rgb-based framework for simultaneous pose estimation and affordance detection in robotic tool manipulation
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3da84687-8238-4a0c-8335-f4f2978a242b · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Affordancellm: Grounding affordance from vision language models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 85a8b5fd-b186-4dc2-8111-4528b4494520 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Object affordance detection with relationship-aware network
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9f478216-4223-4a9d-a180-b49e0cfbf81f · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Learning from 10 Demos: Generalisable and Sample-Efficient Policy Learning with Oriented Affordance Frames
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2b143792-06fc-4c29-8e2b-d5ebde73bc9a · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Closed-loop visuomotor control with generative expectation for robotic manipulation
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ffcaae19-44c8-42cb-a209-fc4839c0b3fb · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Manipgpt: Is affordance segmentation by large vision models enough for articulated object manipulation?
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a9d2855f-7075-44ec-a710-dec7d8234f71 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment R3M: A Universal Visual Representation for Robot Manipulation
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a8dec64e-d821-4735-8a8f-b8f966d90d00 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Representation alignment for generation: Training diffusion transformers is easier than you think
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d1ac1003-d649-41a2-9ecc-b93ef8b26e84 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment 3drs: Mllms need 3d-aware representation supervision for scene understanding
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 59b041a6-3c7c-4314-a72a-cfb099456f91 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Genhancer: Imperfect generative models are secretly strong vision-centric enhancers
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fb5616a3-2328-4a7a-acd4-b85ec585c9b6 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Reconstructive visual instruction tuning
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f977aa6c-3148-43c6-97a9-1613c5ab54f6 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Cross-modality alignment perception and multi- head self-attention mechanism for vision-language-action of humanoid robot
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8e927c4a-c817-4a2b-b8bc-c4ac2c54e454 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Spatialvla: Exploring spatial repre- sentations for visual-language-action model
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 55e7d085-89c9-49ce-9ebb-491aa79a0eb2 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment FLARE: Robot Learning with Implicit World Modeling
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 46538472-d2c3-4216-bdca-8c6c7ef2b353 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Flow matching for generative modeling
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b153f260-a7f7-4d0a-869a-ac83214b19bb · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment SAM 3: Segment anything with concepts
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7367755f-ca11-4dd2-bc85-63e443071924 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Deciphering cross-modal alignment in large vision- language models via modality integration rate
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fedddaaf-862b-4a52-a075-5ab7b42bc6b9 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Learning affordance grounding from exocentric images
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 288c51a2-fa5f-4550-85fd-fd26ce3b5db5 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Locate: Localize and transfer object parts for weakly supervised affordance grounding
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5dc70278-2283-42ba-a76a-6d4c015f3619 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment What do different evaluation metrics tell us about saliency models?
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2d48d88a-0444-4c07-a334-145e7b4af50a · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Color indexing
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bb2a4a6d-75c0-469f-a7b4-980e48f40ccc · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Components of bottom-up gaze allocation in natural images
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9ad4f51e-eddb-41d4-99c8-d6813c24b5cc · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Understanding 3d object interaction from a single image
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 10841c49-fa8b-461a-8d98-cba3d08ef548 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment One-shot open affordance learning with foundation models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 03e23903-42be-4165-b5d9-b49c85e3ec21 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment AffordanceSAM: Segment Anything Once More in Affordance Grounding
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 95f9dfb1-24a1-40d7-b1d6-7928772f05ab · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Grounded human- object interaction hotspots from video
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation da60d217-8234-45df-a374-5b2e104f4b55 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Intra: Interaction relationship-aware weakly supervised affordance grounding
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 14467bd2-8949-4af2-87a9-f7d9651dc935 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Resource-efficient affordance grounding with com- plementary depth and semantic prompts
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 444c3b59-7f1c-4494-a816-434a279fd112 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Lisa: Reasoning segmentation via large language model
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d2900695-e64b-4ed3-be2a-fd3b5ecda3bc · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment MMR: A large-scale benchmark dataset for multi-target and multi-granularity reasoning segmentation
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a12663a9-6238-44a9-b0ad-96e20f6dc542 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Affogato: Open-Vocabulary Affordance Grounding with Automated Data Generation at Scale
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d2ee9b20-a786-49d4-b477-149a7b6c18e4 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 281fbaa8-e280-4702-9c90-90d952b320c6 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Learning fine-grained bimanual manipulation with low-cost hardware
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation cfbbd6f9-3a5d-4464-a8a2-b9d83e5a26c4 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Diffusion policy: Visuomotor policy learning via action diffusion
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 397de45f-204b-4998-8c52-20277e796062 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment 3d diffusion policy: Generalizable visuomotor policy learning via simple 3d representations
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 45ab4b1e-79b5-4b88-8693-366b49c3da5d · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment RDT-1b: a diffusion foundation model for bimanual manipulation
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 93351783-2e91-4a60-81d5-39cb3fb18aea · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment One-shot transfer of affordance regions? affcorrs!
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f73198b2-c764-435c-ba7a-3a9773eff58b · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Weakly supervised multimodal affordance grounding for egocentric images
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 18133da4-1426-44f0-8c0b-3e2c8ed6f92e · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Weakly-supervised affordance grounding guided by part-level semantic priors
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 309184ed-5d07-4737-9b7b-7cfd51332b42 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Reasoning mamba: Hypergraph-guided region relation calculating for weakly su- pervised affordance grounding
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7051b13f-ea93-4d3f-80a3-dbeba47f1486 · outbound
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment Visualizing data using t-sne
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 26ed6c4c-190f-40b2-acc4-f190ccbacca1 · inbound
LIRA: Local Cross-Layer Information Routing for Vision-Language-Action Decoding AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.