Pith. sign in

Paper Citation Record · LEDGER

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation

As of 5 August 2026, this Paper Citation Record lists 83 of 83 outbound references and 0 inbound Pith citation observations for arXiv:2607.06564.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.06564 v1

Coverage vector

measured 83 of 83 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-08T01:41:34.889711Z

measured 83 of 83 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

83 of 83 outbound references displayed

  • verified exact30
  • verified fuzzy51
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3a61427c-8886-43b2-9cd0-792118284461 · outbound

This paper cites Lift3D policy: Lifting 2D foundation models for robust 3D robotic manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Lift3D policy: Lifting 2D foundation models for robust 3D robotic manipulation,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.503719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:cf2cbdd35894c14a2be1c030c017c222d1ef2a973a6ab6bc22b15f1cad2e9d2d

Observation df7d3c09-c8ea-43dd-9a6e-5594ad9e540b · outbound

This paper cites Rt-2: Vision-language-action models transfer web knowledge to robotic control,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Rt-2: Vision-language-action models transfer web knowledge to robotic control,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.499182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:fdf2ba0e24dc6f9a48496cadb1515274bea864ef887da362660dc20d54003770

Observation ffa7db6d-d8bd-433b-bd1a-14f8b1779b25 · outbound

This paper cites Hybridvla: Collaborative diffusion and autoregression in a unified vision-language-action model,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Hybridvla: Collaborative diffusion and autoregression in a unified vision-language-action model,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.496299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:3339558fd5f3ca8dc005b450832d6c1b5bb8883692eab0fb16dc491dfffffad1

Observation aaeb6b21-6405-479c-ba3c-d504e391ca62 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.137272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:e1f3b72c10cf3304e227329d571e5f05d4e7ce665a5089cbe9c39ac5d7807566

Observation bc2cbe71-0298-49f8-984c-fe54590b2892 · outbound

This paper cites Cliport: What and where path- ways for robotic manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Cliport: What and where path- ways for robotic manipulation,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.488497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:cc302a2c1fc8e459da933a085ed15bfe9a3977903220a4a3b96a1299da78d074

Observation 2164c557-af3d-49bb-807e-051ba1500b3b · outbound

This paper cites Point cloud matters: Rethinking the impact of different observation spaces on robot learning,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Point cloud matters: Rethinking the impact of different observation spaces on robot learning,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.497057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:0359188cfeb10a2fd9460afd9a19df6c0fe0e341e752aead03c67f4ce379ab38

Observation 30208e19-aeb8-42e2-99fb-165a3c2bc61f · outbound

This paper cites Flowbot3d: Learning 3d articula- tion flow to manipulate articulated objects,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Flowbot3d: Learning 3d articula- tion flow to manipulate articulated objects,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.484292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:a73dd43a7dcb075102cfcb5c7f67fb43056e1077886be07cc8ded666f1c63572

Observation 4c96cea2-b331-4f17-98f1-fe195220f3d7 · outbound

This paper cites Anygrasp: Robust and efficient grasp perception in spatial and temporal domains,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Anygrasp: Robust and efficient grasp perception in spatial and temporal domains,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.488311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:f0058a9b7c460d1c72d7266079d2a6da95ae2763bc92ee4600df2b218663df59

Observation 6fb29a7c-be9c-4f64-8d6b-ba284c1c16e6 · outbound

This paper cites Perceiver-actor: A multi-task transformer for robotic manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Perceiver-actor: A multi-task transformer for robotic manipulation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.490232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:01f02b9cca814484706297260561f3f6ad15eb27573b32dd90c2f5fa26824c84

Observation 51075247-e37d-49a8-996e-9d9f5ecb4cfd · outbound

This paper cites Polarnet: 3d point clouds for language-guided robotic manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Polarnet: 3d point clouds for language-guided robotic manipulation,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.508595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:0f88a11ec3b411faaf7a1009b390e7c402f2115388d45cd91c9e13aa62af68eb

Observation 92dc993f-8ea2-47cb-ae07-e764fdbf54a4 · outbound

This paper cites Frame mining: a free lunch for learning robotic manipulation from 3d point clouds,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Frame mining: a free lunch for learning robotic manipulation from 3d point clouds,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.501251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:78cba45bed791b9c3a64f6be199b630b46456c52aebe9f556f613a5a9c046603

Observation b66272ae-725c-4261-8e4d-48d1cb92f283 · outbound

This paper cites Leveraging locality to boost sample efficiency in robotic manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Leveraging locality to boost sample efficiency in robotic manipulation,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.506030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:24d67086ad3d607cc46012f4a3aa0c2479aebf855ecb4e8ff80ddd657d7b5fc0

Observation 3b74bc51-473b-490a-8282-9f2f7ea5309d · outbound

This paper cites Rise: 3d perception makes real-world robot imitation simple and effective,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Rise: 3d perception makes real-world robot imitation simple and effective,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.492396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:68d242ce453ae9aa98e670487c05c3fa3c4ba474507367e16f46801435e47f20

Observation 25234aae-51e9-4106-82a3-d4ef6fb1fd99 · outbound

This paper cites 3d diffusion policy: Generalizable visuomotor policy learning via simple 3d representations,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation 3d diffusion policy: Generalizable visuomotor policy learning via simple 3d representations,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.494274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:b29d1b099017443371c87d467e8084da164e13bbd1ee2096bc50d69d48748a60

Observation b914d012-7751-46cd-b05a-b9735b851a43 · outbound

This paper cites Coarse-to-Fine Q-attention with Learned Path Ranking.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Coarse-to-Fine Q-attention with Learned Path Ranking

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.133775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:4c2cd314407fd4731abf04eba552d0097f991b494f0f435fc12d9fa7a7014f73

Observation e6d56247-7f24-421b-9be5-b373902e2616 · outbound

This paper cites Pointvla: Injecting the 3d world into vision-language-action models,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Pointvla: Injecting the 3d world into vision-language-action models,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.494685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:35d8ec769801cb78f03a231cd0e447a81eb57531af9b55934a1651a8754737f2

Observation 35b10ed1-9d6e-485e-9c22-e29f22dcb5cc · outbound

This paper cites Act3d: 3d feature field transformers for multi-task robotic manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Act3d: 3d feature field transformers for multi-task robotic manipulation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.419357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:49a46193e8d7824cd76709c4d3a31db70b8cecb084f54aff50c247feb5e078ee

Observation 97ed145f-8cc0-41b5-bba0-2523d84c8722 · outbound

This paper cites 3D Diffuser Actor: Policy Diffusion with 3D Scene Representations.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation 3D Diffuser Actor: Policy Diffusion with 3D Scene Representations

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.129598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:a0672a52a5a226b350ba95589a609a247a3551cba36007c077568bd9b365018e

Observation 9ec3c7f5-b2bd-4c98-ae2d-c9d9aee43984 · outbound

This paper cites Chaineddiffuser: Unifying trajectory diffusion and keypose prediction for robotic manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Chaineddiffuser: Unifying trajectory diffusion and keypose prediction for robotic manipulation,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.466160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:80145126a7c2d2d5b6c712c27088fee1b142e2677f47405a74f4362b4300cd38

Observation 5c3c8caa-8493-455d-80fe-475bba1e53d4 · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.113782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:4252e1e88cb9e4c99f2e61a0856debf0ec41cca466eb32867056b3f2b52a6c53

Observation 0c1c7bad-1f02-4ed1-b645-e3e299bdccca · outbound

This paper cites RVT: Robotic view transformer for 3D object manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation RVT: Robotic view transformer for 3D object manipulation,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.424869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:026924d3e179fa0fbdadd76c5926ec35975842e480aa45cbce13cd0e7a94a3ee

Observation 9ec907ae-5793-4026-8bdf-f5028c4cfc70 · outbound

This paper cites VIHE: Virtual In-Hand Eye Transformer for 3D Robotic Manipulation.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation VIHE: Virtual In-Hand Eye Transformer for 3D Robotic Manipulation

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.116391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:06a231397380ff77920e1201ab5e06329b5c3b27bbdea7c34f9f9a2329e63fb4

Observation 50171bd3-b02f-4796-b624-2df6f8b1f482 · outbound

This paper cites Sam- e: Leveraging visual foundation model with sequence imitation for embodied manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Sam- e: Leveraging visual foundation model with sequence imitation for embodied manipulation,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.442689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:b45165ae5ee96b8c8a1a71182f9b3a028f0510a197ef43a86c8bcd0f95967103

Observation c67dc770-abfa-4483-b00b-677ee33c0302 · outbound

This paper cites 3ds-vla: A 3d spatial-aware vision language action model for robust multi-task manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation 3ds-vla: A 3d spatial-aware vision language action model for robust multi-task manipulation,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.478042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:ca53f9ebd4d0b94e077daec527907d7635a1086d396bfe33d188fc9f366a20ba

Observation 8b29c835-98c5-497d-ab97-c8de68e869d2 · outbound

This paper cites Ac-dit: Adaptive coordination diffusion transformer for mobile manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Ac-dit: Adaptive coordination diffusion transformer for mobile manipulation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.462714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:2a9f486f67366e48c00652ebf4469660655b7e9b1b52c607c857ebe836822076

Observation 55f99781-d6f2-40de-8722-58c0f7fdd8fc · outbound

This paper cites Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.107388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:0dd27fab228a3ffcdf77a123a0b4770c1aee87014ffa6a10f7573d414564b4a6

Observation 279dddb5-cfbb-4c1d-b8b0-db0374ffeb2c · outbound

This paper cites ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.061080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:4345db353e6fa2794f708e311f1cac6c136600972b11264307292f65bf0e1f9b

Observation e6bf9257-e101-4579-9ee1-11f270853ae2 · outbound

This paper cites Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.114231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:b84fbbc9f6643bb24828505f9ceeef9c1967dabe59133e6ff34f94beea11a1af

Observation c5553070-d855-4cca-82b1-977725db42bb · outbound

This paper cites Motus: A Unified Latent Action World Model.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Motus: A Unified Latent Action World Model

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.124568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:d5ea4ca8c6603c2b1894721ffb106dc6bf412b363b9d29d25a31628510942cf5

Observation 537b6a63-eed3-4164-9d91-c82112b04db8 · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.122025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:c30587a1ea5136a1c2cd00986e4ea87333094df5565c5274bc39b49b9451960c

Observation 9dc24a6e-430f-4ad9-ae66-d8ba0d93773e · outbound

This paper cites Droid: A large-scale in-the-wild robot manipulation dataset,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Droid: A large-scale in-the-wild robot manipulation dataset,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.445067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:db83b3516116ffb224a0f3489ffbd60979cbe537e346e2139e159ddd710ae767

Observation 00592834-e3f2-4501-8547-be31c3ba5015 · outbound

This paper cites Robomind: Benchmark on multi-embodiment intelligence normative data for robot manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Robomind: Benchmark on multi-embodiment intelligence normative data for robot manipulation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.447124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:b6047b3d9bfbaa7e19cc0086be610e47e93e672c687d3e7f82008c647f80e101

Observation 947a3ecf-16a8-4169-80a3-b1bd4754edc7 · outbound

This paper cites Meta-world: A benchmark and evaluation for multi-task and meta 12 reinforcement learning,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Meta-world: A benchmark and evaluation for multi-task and meta 12 reinforcement learning,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.456874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:6550a4ac75eb7104877970475e042089948f87737a4aee69999533f3035f38c7

Observation cb76c6c8-7d12-460c-bfb0-9987521209d5 · outbound

This paper cites Rlbench: The robot learning benchmark & learning environment,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Rlbench: The robot learning benchmark & learning environment,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.443239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:32a35206785cc6eaa11a03ddf5647e064ae897ebbf36d73bb4742c1ea2e70844

Observation 60b390b2-ce2c-4043-9678-efac8fdd2822 · outbound

This paper cites Pointcontrast: Unsupervised pre-training for 3d point cloud understanding,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Pointcontrast: Unsupervised pre-training for 3d point cloud understanding,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.426590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:13cc883d990ede58669a1cd59f7f78473b96a178a6b61c7879489141829542c9

Observation 1ec05467-921f-4057-b1ea-800f5972cb08 · outbound

This paper cites Unsupervised learning of visual features by contrasting cluster assign- ments,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Unsupervised learning of visual features by contrasting cluster assign- ments,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.455939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:ba17f897231de1e969eb587eff956a63ffa41b28c04b1003c31e64e4ce3ee424

Observation d2beec50-01e8-4c16-8f9b-ec979e2ac5ef · outbound

This paper cites Masked autoencoders are scalable vision learners,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Masked autoencoders are scalable vision learners,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.467168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:12e9bb77e314e905d020a689603efd686c78deee40e66cf55619aea481262c8a

Observation 41f22684-170b-425d-a7b3-65b5082be875 · outbound

This paper cites R3M: A Universal Visual Representation for Robot Manipulation.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation R3M: A Universal Visual Representation for Robot Manipulation

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.116804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:a4fae59b17265f40a8104211707a4a6980a1bc725b691a9defd9afddf7a67ce4

Observation 062a0bd8-3d25-43a6-a460-ea13d9c89059 · outbound

This paper cites VIP: Vision Instructed Pre-training for Robotic Manipulation.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation VIP: Vision Instructed Pre-training for Robotic Manipulation

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.074047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:d029fef1540a40fd523b3e71a60486bb627174b1de7a92e8511f8010cd09ad1f

Observation 499d873b-8dce-4348-a906-6a67ceaba8da · outbound

This paper cites Masked Visual Pre-training for Motor Control.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Masked Visual Pre-training for Motor Control

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.091125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:948594458682392e57334228bfe1590a2ead0115d633dad05128c0ce8bf58ced

Observation 8e7393ca-bc1d-43de-b992-d494366b3670 · outbound

This paper cites Where are we in the search for an artificial visual cortex for embodied intelligence?.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Where are we in the search for an artificial visual cortex for embodied intelligence?

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.460621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:c71dd1608af19808eaedde9b8c7bcfc1cd67d2ccfb1e5103d19eb1e1d3643dd3

Observation ba03ce90-d236-4bd3-8b9b-e3f9bde7272b · outbound

This paper cites Language-Driven Representation Learning for Robotics.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Language-Driven Representation Learning for Robotics

Reference 42

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T01:44:26.099688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:540fc75eaefc93f99b47eda7148cdb57f398be41f9f7bb19e99348ba09b61dca

Observation 78ec78ba-3305-439c-b0c9-7a957547f24d · outbound

This paper cites Multi-view masked world models for visual robotic manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Multi-view masked world models for visual robotic manipulation,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.407505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:83ebb1845a12bbb82242dd636769bbf9df7e4a7b7ab8993007e66a6a897f308d

Observation af032d8d-b1f9-48c0-b8f4-ba6a09b1d1d0 · outbound

This paper cites 3d- mvp: 3d multiview pretraining for manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation 3d- mvp: 3d multiview pretraining for manipulation,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.464410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:ad3f2e1e43311de05eb39bbcb91fd6e81f4635bcc67f5d1003bc367cbf43f386

Observation dead4535-6678-4c71-8296-f629af013947 · outbound

This paper cites SPA: 3D Spatial-Awareness Enables Effective Embodied Representation.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation SPA: 3D Spatial-Awareness Enables Effective Embodied Representation

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.093945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:bc88a6675552af9669c418948daf2798ea292c1e99e1c952bf06e5a87cc793de

Observation 59db35e6-24a5-4f4a-9513-c2eda190a162 · outbound

This paper cites CLAR: Learning 3D Representations for Robotic Manipulation by Fusing Masked Reconstruction with Multi-Level Contrastive Alignment.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation CLAR: Learning 3D Representations for Robotic Manipulation by Fusing Masked Reconstruction with Multi-Level Contrastive Alignment

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.080033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:b3ef47ac896642397d6557a3a74b75ad28047e9b89c73a89391bdf9f692a5765

Observation 3816a99f-5a93-4fc0-89ef-a869e900d24e · outbound

This paper cites Hyperbolic multiview pretraining for robotic manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Hyperbolic multiview pretraining for robotic manipulation,

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-08T01:44:26.111505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:521110cac74f5b5decbf934620ffb93c005ababe37242a2de217fdd7c1c662e0

Observation f3f8c7ed-bdb4-49ba-98fc-4277d0144a55 · outbound

This paper cites Robo3R: Enhancing Robotic Manipulation with Accurate Feed-Forward 3D Reconstruction.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Robo3R: Enhancing Robotic Manipulation with Accurate Feed-Forward 3D Reconstruction

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.082852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:0dc1d5ec4cf1df3b71349e316810d5aec0c57caff62251b33402ecf3e2e5035d

Observation eadf99a7-119a-45b9-87f5-522a97b815b8 · outbound

This paper cites Canonical policy: Learning canonical 3d representation for se (3)-equivariant policy,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Canonical policy: Learning canonical 3d representation for se (3)-equivariant policy,

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-07-08T01:44:26.127649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:1e8d39f8fc113ca66ee57623b9d5183981d254e2453f092ce2851ebe6f774f14

Observation a597a7be-2ca3-45d0-97fe-ffe29112f2a8 · outbound

This paper cites Prismatic vlms: Investigating the design space of visually- conditioned language models,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Prismatic vlms: Investigating the design space of visually- conditioned language models,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.467960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:b401b6bc933da3d154740e096790342e115b764e256e48f3df8633d46eefeeb3

Observation 67c8ec45-ec17-4b57-9ae9-cb61a6757ee5 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.065347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:df15737de581a52b87cc84e2ae7aa54354ee21301c863c34908d7784383d360a

Observation acaa1267-462f-40a5-9303-97d560c8893c · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.104619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:e0f30ba1d35bc1f63344511476a1a7b81b7ea36f32b906a08e7be99e137d5e4d

Observation 033cd1a7-6eec-4a1d-bbd6-9bc48f8dacf1 · outbound

This paper cites Cot-vla: Visual chain-of-thought reasoning for vision-language-action models,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Cot-vla: Visual chain-of-thought reasoning for vision-language-action models,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.473892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:48b98f424205c7fe1df0f2534cb37ca6e35b3486033c4900653939e42e986ad7

Observation 64056262-b835-4a91-b468-da5ca487d580 · outbound

This paper cites Up-vla: A unified understanding and prediction model for embodied agent,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Up-vla: A unified understanding and prediction model for embodied agent,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.475984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:1e86e01481d357fdce3fe287dfe67a48c203bd208259a890bd548c5da82b6c5f

Observation 45a37dab-e35f-4736-9e24-4e6783129abe · outbound

This paper cites Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.085774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:b54822d8379801d0e57be9bcce8ca301d12fa8d33a26dee9a59e4f3d641a9f59

Observation 9c986b73-d9ab-4617-be13-1542a5493b26 · outbound

This paper cites Manualvla: A unified vla model for chain-of- thought manual generation and robotic manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Manualvla: A unified vla model for chain-of- thought manual generation and robotic manipulation,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.451268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:a61e072142eadc390af7ff1009438da808ab381fcbdd5ff076ee9be8e189d4b1

Observation 93c8e412-fc67-438b-bb9e-7ea27066bd84 · outbound

This paper cites Last {0}: Latent spatio-temporal chain-of- thought for robotic vision-language-action model.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Last {0}: Latent spatio-temporal chain-of- thought for robotic vision-language-action model

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-07-08T01:44:26.124722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:2c328cc34536e74cfd1e5c2d549c78144a6f34c9a216075a8e35b8ce1a2553b7

Observation 01b0a1a4-acf2-4197-b9ff-c9271706f0ae · outbound

This paper cites MV-WAM: Manifold-Aware World Action Model with Value Augmentation.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation MV-WAM: Manifold-Aware World Action Model with Value Augmentation

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.130746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:d5f2066a410f8a2cec35656bf853b5a1a3d111a563a67bc49d64cf0ae0a4926d

Observation 212a7f7a-816e-4353-8795-351d36a2b1ac · outbound

This paper cites LaST-HD: Learning Latent Physical Reasoning from Scalable Human Data for Robot Manipulation.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation LaST-HD: Learning Latent Physical Reasoning from Scalable Human Data for Robot Manipulation

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.108197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:fc34118589be81e9aa5d38e60f51fc154b5efc661bc85bd034a971a4737969a1

Observation 0dbfbe39-d6cd-407f-a9ad-0bd44f5ee82a · outbound

This paper cites Mla: A multisen- sory language-action model for multimodal understanding and forecasting in robotic manipulation.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Mla: A multisen- sory language-action model for multimodal understanding and forecasting in robotic manipulation

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-07-08T01:44:26.099387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:20c2a8b5c69a4ac5d86a833d68ecac00291b873405d7fce0dd0fed4bd45e6acb

Observation e3c5b95c-3e9d-4c21-8d0c-cae193053c8a · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.080970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:a633b0f2bebf1e35c406944b9ddce0ef27c00b468c9b40b9d6f3edbf3528ffe6

Observation d001f877-23cc-43c3-8733-2ca9043d8331 · outbound

This paper cites Sigmoid loss for language image pre-training,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Sigmoid loss for language image pre-training,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.439660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:9e4b2c97fc84d344ee7dc7133349d0d68122115fbbac180759af37ad24ac5904

Observation b2f25930-2f4b-4ab2-9ba3-107cb7b74c51 · outbound

This paper cites DINOv2: Learning robust visual features without supervision,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation DINOv2: Learning robust visual features without supervision,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.437506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:da6be89f20fbcf0b4096c2de49a529e7c1000fa44c0d9be6f2d1121b6b986c8f

Observation f9a41e6a-6d0c-4e11-a41b-4e9abb027bbe · outbound

This paper cites Pointnet: Deep learning on point sets for 3d classification and segmentation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Pointnet: Deep learning on point sets for 3d classification and segmentation,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.454985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:563ae701ee3f7a13548a5f83edf5a104fa5b347a37899e2775080cdf91e40419

Observation 99a1473c-4ad1-4b13-a892-90342a8ddc52 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 66

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.126973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:c5fe8e3ccc8323378f9cc5299a138fe3d782b72f3cef1ed7efd41663eecf144b

Observation 16c5df7e-0a7e-4cca-9150-533de23edb4f · outbound

This paper cites Open x-embodiment: Robotic learning datasets and rt-x models,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Open x-embodiment: Robotic learning datasets and rt-x models,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.448916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:cc3d0a1d7b29c2825e5b2cd49409050ce14dc69418e0d572926538cbabe577e6

Observation 95d8601c-04b7-47fa-9e92-a3220503a81f · outbound

This paper cites Robomind 2.0: A multimodal, bimanual mobile manipulation dataset for generalizable embodied intelligence.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Robomind 2.0: A multimodal, bimanual mobile manipulation dataset for generalizable embodied intelligence

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-08T01:44:26.119324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:32af4cf4c17eb32c9555cc0210c7b792a30551b5ec9dce3fa99e77e171beb6cf

Observation c7dd7034-60f9-4d40-838e-5eb0a6e23c61 · outbound

This paper cites Vggt: Visual geometry grounded transformer,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Vggt: Visual geometry grounded transformer,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.411232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:cd7fdc73e7bf027bcc73c357417fed59bef3e0bc4b29cefc8febc0141cc96727

Observation b7571889-2941-4f7a-a9eb-634ed3ff3d5d · outbound

This paper cites Masked autoencoders for 3d point cloud self-supervised learning,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Masked autoencoders for 3d point cloud self-supervised learning,

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.430460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:d7e2a3e715aed75d209f754ca4630e8e5de5c4f6385b1a23e0913ec0e97dfef1

Observation 917dbb0d-444f-4a5d-97c4-047eb420db1e · outbound

This paper cites Lora: Low-rank adaptation of large language models,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Lora: Low-rank adaptation of large language models,

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.471721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:ab176e0f6223132b32f227d4bf4dce8fd6e730369f187b1c094c89f347521f20

Observation 2b0d0b1d-ad22-4f3b-b403-77684b71ff1c · outbound

This paper cites Decision transformer: Reinforcement learning via sequence modeling,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Decision transformer: Reinforcement learning via sequence modeling,

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.452969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:dbf51abe4eefee9e91abd81384eba8aa2b97d46bcc901ca27d6aba49decf0e57

Observation acbaec18-87b7-4d5f-864e-e8f0aa648cb2 · outbound

This paper cites Large sequence models for sequential decision-making: a survey,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Large sequence models for sequential decision-making: a survey,

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.458889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:441841246b2b14d8dbc85b7f4b126282ba35715f131e0b3d4e82bce02c938253

Observation ef2f2ca4-1990-471b-92cf-11bdb5cc107d · outbound

This paper cites DeeR-VLA: Dynamic Inference of Multimodal Large Language Models for Efficient Robot Execution.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation DeeR-VLA: Dynamic Inference of Multimodal Large Language Models for Efficient Robot Execution

Reference 74

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.119689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:1fe4a8923079c193a7825d9ae642d88526eb02a95967bdb03e3b9957513083af

Observation edb5fbfa-7f84-42ff-b93c-14adde45dec0 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 75

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.121773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:61e957c79e5e923b9bdaa2437b1dcc09c6c706014477ff105cfbacc98bf6103d

Observation f3ec79bd-df9c-469d-987c-b0c5b8d68dfe · outbound

This paper cites Denoising diffusion probabilistic models,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Denoising diffusion probabilistic models,

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.469833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:fd34650cc760ec8b316e46874faec7b0832cdb8548edb7bc85feaf5bf58079ae

Observation efad29e5-3d65-4927-b017-d68fecc894a3 · outbound

This paper cites Denoising diffusion implicit models,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Denoising diffusion implicit models,

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.453678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:3e2f7742bd94f9c1324978df6a5f85a153efbf53f2060fb5471f9b86dff1ed18

Observation 020f5720-3f47-4ee8-b7da-7c3c125b1b8b · outbound

This paper cites Rdt-1b: a diffusion foundation model for bimanual manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Rdt-1b: a diffusion foundation model for bimanual manipulation,

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.430898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:fd84eff6d7a1c0f912101da46a21f61426a6f0be3393041236b3a502d79a0c5d

Observation 1cde3025-4293-4b57-94ad-f2b153f477b1 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Learning transferable visual models from natural language supervision,

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.458106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:77a2412d8db1a25d780c98bdb5b4e7888b6bd7164eaac3910c1ef6e34e46e3ee

Observation 95e194ee-e8e6-442c-b5fe-bf183f9fe893 · outbound

This paper cites Pointnet++: Deep hierarchical feature learning on point sets in a metric space,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Pointnet++: Deep hierarchical feature learning on point sets in a metric space,

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.440395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:d5304c18a5fb56845949543af4452a42752cdf14c7f878559257164d8e0f2e1f

Observation c0d5d91f-bd33-4f69-ab99-b6afe09c1297 · outbound

This paper cites Pointnext: Revisiting pointnet++ with improved training and scaling strategies,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Pointnext: Revisiting pointnet++ with improved training and scaling strategies,

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.423155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:01cb10bae6ce030a18dd12b9c30b1f12a94fe4d8c398d523f5cc35e19b6784eb

Observation 2964f16a-97c7-48b9-86ac-a3dea06e67de · outbound

This paper cites The open motion planning library,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation The open motion planning library,

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.434358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:b06e929cd3d93bd29a376c9e265174578c2a6cb33762585cb7dbeec14f98f605

Observation c38e1777-dcef-4ef1-9f7d-abf6ebc5fff4 · outbound

This paper cites Perceiver-actor: A multi-task transformer for robotic manipulation,.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Perceiver-actor: A multi-task transformer for robotic manipulation,

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T01:44:26.465104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:df5bf9b6f90b845cc3c38358599a81c887e7e9cb6aef79add0238da011d9402c

Observation 48bff3fd-d91a-4147-bd5b-839c2206a1fc · outbound

This paper cites Fast-in-Slow: A Dual-System Foundation Model Unifying Fast Manipulation within Slow Reasoning.

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation Fast-in-Slow: A Dual-System Foundation Model Unifying Fast Manipulation within Slow Reasoning

Reference 84

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T01:44:26.102092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:41:34.889711Z digest=sha256:bca4f3e8d07ae5bbe6b6a710b83bac5555ea92a1f55a50784c81ea14a7200675

Pith citing papers

No inbound Pith citation observations are available.