Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:43:04.627490Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 1 inbound Pith citation observation for arXiv:2602.02533.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:43:04.627490Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:43:04.520409Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-15T15:43:04.746675Z
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b4e81357-9ff2-4195-97c7-84368156384e · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 374d0762-650f-4ea5-904f-4e4928c967a9 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Hyperbolic Semantic Alignment Our HMVLA framework is based on the Lorentz model in hyperbolic geometry, as shown in Figure 2
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3b3459f3-0e28-4d79-8c9d-f5fd089ef3dc · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Specifically, we utilized four datasets: Spatial, Object, Goal, and LONG
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7e2b28c3-31f1-4c66-a8e1-e64a33f8b7b4 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models By embedding multimodal features into a hyperbolic space, our model effectively cap- tures the inherent hierarchical relationships within image-text data
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d351ea94-7afd-49cf-9086-d23d26538873 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models 62277011), National Key Research and Development Program of China (Grant No
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0adcf0ee-9299-4181-839e-83cb85499538 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models GPT-4 Technical Report
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec1c130b-751d-47a7-8162-50061ddafd40 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models The llama 3 herd of models,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation cecb79b9-8a35-4979-b63f-de2fe9d14aef · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Llapa: A vision-language model framework for counterfactual-aware procedural planning,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7d577ea6-d172-4e20-b863-7435f629a6bc · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Sage: A visual language model for anomaly detection via fact enhancement and entropy-aware alignment,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2aa2faff-d727-46f1-92bf-0a2b1a174921 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models RT-1: Robotics Transformer for Real-World Control at Scale
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4232ef47-45e8-483f-b5ba-713a3ae03824 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models RT-H: Action Hierarchies Using Language
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9324eb2-2f2d-4205-b7c6-30da002c49fb · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3a935ef-f69e-49ff-9339-8997b4400539 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models 3d-vla: a 3d vision-language-action generative world model,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 55c8244d-1f2e-41fc-88ac-3851e635161c · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d2e091f-1111-4eab-ab2d-adb679afec2d · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c019518b-7f15-43c3-b7ec-ffe281ea1d04 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models A Survey on Vision-Language-Action Models for Embodied AI
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed3b8eff-bb83-406f-a334-799a3b209fc8 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Vision- language models for vision tasks: A survey,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a349a12c-ef24-414c-8aea-1a20004be9cb · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Rt-2: Vision- language-action models transfer web knowledge to robotic control,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 25bab865-bac0-4005-9ab4-1f1a06fc8f48 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b47d593f-eb4c-453f-9fc5-d3cfe3cccebe · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Learn- ing transferable visual models from natural language super- vision,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0e8122c4-7dcd-42c5-b4ee-676f65a607ee · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models An ex- tensive study on pre-trained models for program understanding and generation,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d262a5c5-f323-48e6-8588-3f9b9e7b8006 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Hyperbolic spaces,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f16ec933-ddd6-428d-8459-7595aac87099 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Hyperbolic image-text representations,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 266bde50-db5b-4ec9-bfb9-1f57b3fb541a · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Libero: Bench- marking knowledge transfer for lifelong robot learning,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 14f30393-63f6-4484-be68-53fb55ef1699 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Zur elektrodynamik bewegter k ¨orper,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 85e28ff1-6b6d-4fdf-9da2-a7318ab2f29a · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Dita: Scal- ing diffusion transformer for generalist vision-language-action policy,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation fd6e3668-c86d-4c90-adab-f88e938af3f1 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Diffusion policy: Visuomotor policy learning via action diffusion,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b7d81815-f5f3-4ad1-a968-98e8a21057cc · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Octo: An Open-Source Generalist Robot Policy
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 895c2bbf-5bee-4c93-9995-09c5864946e0 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Tra-moe: Learning trajectory prediction model from multiple domains for adaptive policy conditioning,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0baec4dd-c9d3-4142-830b-0b86a2662837 · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models Cot-vla: Visual chain-of-thought reasoning for vision-language-action models,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e2e7f64e-bc1c-4d6c-8fc8-398d7eb6489d · outbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models OTTER: A vision-language-action model with text-aware visual feature extraction,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b4e81357-9ff2-4195-97c7-84368156384e · inbound
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.