Pith. sign in

Paper Citation Record · LEDGER

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation

As of 8 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 0 inbound Pith citation observations for arXiv:2608.05042.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05042 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:44:18.247570Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

65 of 65 outbound references displayed

  • verified exact2
  • verified fuzzy35
  • unresolved27
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation da635c3e-14ba-454c-ad3a-f0147571b3d3 · outbound

This paper cites OpenVLA: An open-source vision-language-action model,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation OpenVLA: An open-source vision-language-action model,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:11.836856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:11.836856Z digest=sha256:e2bf0bc74a4e92dd8d863e656e58ff90272045050b7cbdf427e9b2aee1000001

Observation 30ef66ac-b196-4113-845a-b1c1b5b72135 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:11.920374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:11.920374Z digest=sha256:c60f56e2677fb711da186bb32888b2c7a66aafd74aea257cadc7ff22149cd609

Observation 4c4286e5-ac21-4ebc-becb-0504efc4af64 · outbound

This paper cites Wall-OSS-0.5 Technical Report.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Wall-OSS-0.5 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:12.106434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:12.106434Z digest=sha256:a3ee770122392cd524c910452229982b398970252dcfa074cc7586b105210eb6

Observation 9ae2648e-564a-4774-8ef4-35d5a4bf726b · outbound

This paper cites Vision-language foundation models as effective robot imitators,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Vision-language foundation models as effective robot imitators,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.055685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:12.235379Z digest=sha256:7cbef33500c4807d6f17063416e758e7615611feb158a54af66d13b4fb61b65d

Observation 72f95a25-9bc3-4002-aaa5-25cae51bd56a · outbound

This paper cites RT-2: Vision-language-action models transfer web knowledge to robotic control,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RT-2: Vision-language-action models transfer web knowledge to robotic control,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.046255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:12.390934Z digest=sha256:be8952245f61994dd52973cc1fed3405ba245ddc6197a499eefeedad2142ea68

Observation 36ec393a-1107-4976-8e73-5b110a128bc7 · outbound

This paper cites Perceiver-Actor: A multi-task transformer for robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Perceiver-Actor: A multi-task transformer for robotic manipulation,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.036180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:12.523105Z digest=sha256:916f1ff03c4e51b60b5b44ab0a5d12758ff9a8235b2e9bfdd7d79febd1e19559

Observation 62fa746e-6e30-4d58-bd9a-fbd27cc63775 · outbound

This paper cites 3D Diffuser Actor: Policy diffusion with 3D scene representations,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation 3D Diffuser Actor: Policy diffusion with 3D scene representations,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.027748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:12.644589Z digest=sha256:c4d44744f63e16ae0c7ac675f2c98995800093ba9c242009630759feb9d8182f

Observation 8b5524d3-a092-490e-a37f-9b0b724fa9be · outbound

This paper cites Act3D: 3D feature field transformers for multi-task robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Act3D: 3D feature field transformers for multi-task robotic manipulation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.019261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:12.852518Z digest=sha256:27e47db1508237cde176c35964641cd36b933de8ac1c757758235e69661db842

Observation c172a669-c510-4cd3-84cc-a552d268c8b2 · outbound

This paper cites RVT: Robotic view transformer for 3D object manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RVT: Robotic view transformer for 3D object manipulation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.010416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:12.979457Z digest=sha256:8508b7d5467f07fa50b55a1833c725061749ad3b808d971ac4b35920b21cef79

Observation 132b4569-ec0d-4f4b-9f54-2f83a980c37f · outbound

This paper cites RVT-2: Learning precise manipulation from few demonstrations,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RVT-2: Learning precise manipulation from few demonstrations,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.001863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:13.078472Z digest=sha256:090780332bf6dcd8041e8d458ac2ec405e5fc83244fa54a9465ec33ea6012750

Observation 111db71a-e83c-4bf4-ae07-e402e17c8359 · outbound

This paper cites 3D-VLA: A 3D vision-language-action generative world model,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation 3D-VLA: A 3D vision-language-action generative world model,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.993333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:13.210283Z digest=sha256:5ff45b551d6bd8b5e9b4643e0ab8af6642d2c534a71f9cf1826f7708a234245d

Observation abec4001-6626-4f0a-aa60-e1d5f1aae51a · outbound

This paper cites SpatialVLA: Exploring spatial representations for visual- language-action models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation SpatialVLA: Exploring spatial representations for visual- language-action models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.984740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:13.350018Z digest=sha256:5e57318158afd8053adaa1e0330e459178706a80e32539f27287910ade7bb519

Observation e7176949-b0a9-44bd-8dc7-6660c0d0d8c7 · outbound

This paper cites RLBench: The robot learning benchmark & learning environment,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RLBench: The robot learning benchmark & learning environment,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.976335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:13.499499Z digest=sha256:80c3f10c8c3cc6475dd7c0a974cb4c33707cd777eaf3413ad9a4a984c6215dc3

Observation edbb9f7e-c70c-4f75-96bf-10ee976757d3 · outbound

This paper cites THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:13.672028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:13.672028Z digest=sha256:9bcaf0a9e28a9905f584441ed2029b4fd1fd783c7d3dccedbc6604915b1267f6

Observation 6973dddc-827f-46de-a672-d9483518cd54 · outbound

This paper cites Towards generalizable vision- language robotic manipulation: A benchmark and LLM-guided 3D policy,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Towards generalizable vision- language robotic manipulation: A benchmark and LLM-guided 3D policy,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.966940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:13.754442Z digest=sha256:b79bbd9fbe8e91ae4abcf7b87cebaaab48a24707c696eea48cfa74c5c96222c3

Observation fe57631e-5ac7-41f1-a37f-d84c9630ec0d · outbound

This paper cites RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:13.849049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:13.849049Z digest=sha256:5949aefaaf0a9302d55ab542387c836904a8b0385e8e49f94d3db690bc15005b

Observation 10fe969f-77ba-4af8-8513-e8fb104d71cf · outbound

This paper cites SAM2Act: Integrating visual foundation model with a memory architecture for robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation SAM2Act: Integrating visual foundation model with a memory architecture for robotic manipulation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.956456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:13.970866Z digest=sha256:9f0ea2ed3001afc9fdccc580ae5b1a54f279cd3fede14f0025febc4c8866de5d

Observation 496c8dc8-a445-49a6-8865-e1fe4a147bc7 · outbound

This paper cites BridgeVLA: Input-output alignment for efficient 3D ma- nipulation learning with vision-language models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation BridgeVLA: Input-output alignment for efficient 3D ma- nipulation learning with vision-language models,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.945963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:14.064287Z digest=sha256:b4c40f532c373bf07e3638c53dd58b86142568ffd8c038e66fc75aa4f34f9555

Observation b4bcc8d6-2459-47ba-b90b-71f01c45a501 · outbound

This paper cites RT-1: Robotics transformer for real-world control at scale,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RT-1: Robotics transformer for real-world control at scale,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.935088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:14.135586Z digest=sha256:b30835ac5ed8a9d9d71b6f8eea8407aaed5fcf4139522b109a60626be5fb0e93

Observation a9af96c5-bcb5-4b3d-a666-3695f9e631c1 · outbound

This paper cites π 0: A vision-language-action flow model for general robot control,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation π 0: A vision-language-action flow model for general robot control,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.924717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:14.302562Z digest=sha256:b53af7216c146b1211959d9b25df71806a8a062d24de59801ded3eec3a046eca

Observation 371af7de-dfa1-40a4-806f-f70e409abec5 · outbound

This paper cites FAST: Efficient action tokenization for vision- language-action models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation FAST: Efficient action tokenization for vision- language-action models,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.914178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:14.470124Z digest=sha256:179e4242a84b3d99dbc171c70335544c66b77b150e32eaf4ef3448afb6499bb9

Observation 2c2b5ade-1432-48d0-830c-4575e227a1bf · outbound

This paper cites $\pi^{*}_{0.6}$: a VLA That Learns From Experience.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation $\pi^{*}_{0.6}$: a VLA That Learns From Experience

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:14.652245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:14.652245Z digest=sha256:79cca3e49092e7ec629a8b169fef46980f5ad990629a2e0dd7fdd4dcb4235a15

Observation 371df10c-c4af-4cec-bc5e-716115ee111e · outbound

This paper cites ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:14.797218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:14.797218Z digest=sha256:feb6b174b268f71ca54bb0e971fc6509bcec9621d8d168ed86bd558f2022a026

Observation 8cf5754c-5969-4399-95a6-b3c9f23c9da4 · outbound

This paper cites GEN-0: Embodied foundation models that scale with physical interaction,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation GEN-0: Embodied foundation models that scale with physical interaction,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.902689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:14.955267Z digest=sha256:7fde00aed41925f017cff55702a129108ab62176565a1235fdda7c3c9876f4ac

Observation 5ddf7257-1253-40f0-a59d-5e17719dbaa1 · outbound

This paper cites GEN-1: Scaling embodied foundation models to mastery,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation GEN-1: Scaling embodied foundation models to mastery,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.892286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:15.131468Z digest=sha256:6e3ada6eeeafd896800c218ecd4a6780af0b763116e47d63ad3d73d151521fa9

Observation ec1cd722-0cfb-4adb-babe-5a00236107cd · outbound

This paper cites GENE-26.5: Advancing robotic manipulation to human level,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation GENE-26.5: Advancing robotic manipulation to human level,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.881643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:15.233329Z digest=sha256:4e8ab596de33bfb9b0a639cf68653637b60c2ac8894c473ef3baf0d2b2ecd3c1

Observation 821c082a-2eb8-4751-a1ee-4bec48253783 · outbound

This paper cites ACT-2 preview: Generalizing reliability,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation ACT-2 preview: Generalizing reliability,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.871047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:15.384527Z digest=sha256:0db50c69a9ae32a0859961f5642622888117059c9ed167d3219ebb5255af1144

Observation e1484038-fc1c-4daa-839f-7bfc9bf6ca4f · outbound

This paper cites Open X-Embodiment: Robotic learning datasets and RT-X models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Open X-Embodiment: Robotic learning datasets and RT-X models,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.860168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:15.548311Z digest=sha256:89ccd3cf3a7dfa3016ef920a223de2971e035eda6ad59e71e63e623908fb3b8a

Observation cd1f516d-f3ef-42f3-9164-897156e98ac9 · outbound

This paper cites PolarNet: 3D point clouds for language-guided robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation PolarNet: 3D point clouds for language-guided robotic manipulation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.848674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:15.653628Z digest=sha256:3106873adf2fd61b595c206700a172d28e977cdcdffadf2fe798915714e53145

Observation eb140a8d-ce8b-4c83-9b44-15e09803f790 · outbound

This paper cites M2T2: Multi-task masked transformer for object-centric pick and place,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation M2T2: Multi-task masked transformer for object-centric pick and place,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.837170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:15.829628Z digest=sha256:f7a7f7288a62f918add0c745fb532f0470644b634127e6ff0d6ad84f4195501d

Observation c4caa6bb-3519-40db-92e1-62c5f229d026 · outbound

This paper cites FP3: A 3D Foundation Policy for Robotic Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation FP3: A 3D Foundation Policy for Robotic Manipulation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:16.040209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:16.040209Z digest=sha256:8f318e611adc4caef6a418a19a878af3ce1aaafeb945369f9259b3245ab4aac4

Observation 12c3079f-f9fd-4bd7-b466-94c5828442f2 · outbound

This paper cites Coarse-to-fine Q- attention: Efficient learning for visual robotic manipulation via discreti- sation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Coarse-to-fine Q- attention: Efficient learning for visual robotic manipulation via discreti- sation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.826462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:16.225713Z digest=sha256:412645e98d5c5a5bb2d39ec9ee8caba4f8c03b3113fa6c75887277ed30f72573

Observation deeacec7-9962-42cf-ab9b-33007842f9c6 · outbound

This paper cites PointVLA: Injecting the 3D world into vision-language-action models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation PointVLA: Injecting the 3D world into vision-language-action models,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.814607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:16.363898Z digest=sha256:d8906b8ed6f1527378a82dcb60eb0cdf3ec83bbf0dceede8c4c85bd82acff657

Observation 518fea34-2208-4803-80f6-e2a854ba2e3a · outbound

This paper cites Lift3D policy: Lifting 2D foundation models for robust 3D robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Lift3D policy: Lifting 2D foundation models for robust 3D robotic manipulation,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.803293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:16.530929Z digest=sha256:51ebe56e698e4de636621b2f6fa0740e6d20632fc30d3e644654c1f6d4d3625c

Observation deb3e36c-995d-4e7d-9f68-269b8fdf977a · outbound

This paper cites DINOv2: Learning robust visual features without supervision,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation DINOv2: Learning robust visual features without supervision,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.791935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:16.777096Z digest=sha256:3c374c83dc69bf6741f5711aa4c0a711e2e3bfd05b33b7ac164826744aec6208

Observation dd8f4b74-cad6-4f31-b450-c15e922ec900 · outbound

This paper cites OG- VLA: Orthographic image generation for 3D-aware vision-language action model,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation OG- VLA: Orthographic image generation for 3D-aware vision-language action model,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:16.921827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:16.921827Z digest=sha256:0c526d2ec1c6ed870534d8bc505fb5bb0628f8f76ed1282e93454be82b55d481

Observation 0c222d1c-b006-4f1c-a73e-1d935006267a · outbound

This paper cites Instruction-driven history-aware policies for robotic manipulations,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Instruction-driven history-aware policies for robotic manipulations,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.011327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.011327Z digest=sha256:b533aa8e54e827a8ec7aaf5582ba24bac3c591a22fd40c029a914f1db6e62db0

Observation e8352d45-348c-455d-9e25-5fba4dab6404 · outbound

This paper cites Causal World Modeling for Robot Control.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Causal World Modeling for Robot Control

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.098344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.098344Z digest=sha256:8bbfe535c6b94f98224948419288ec8612d8869c1151b5acaa19be8d5b222b8a

Observation 19f9ba14-2da4-423f-9532-901e793b4ed8 · outbound

This paper cites World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:44:18.473850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:17.241212Z digest=sha256:571a0f4617eb5fa63e1d5e2c1405be17ead214e247ca2cef70d82ae0801e8ff7

Observation f0b9dd4f-6e94-459e-9229-cbfa958934ba · outbound

This paper cites MemoryWAM: Efficient World Action Modeling with Persistent Memory.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation MemoryWAM: Efficient World Action Modeling with Persistent Memory

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.373966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.373966Z digest=sha256:b3b502c5ac10f51d8cdff9c123b10593528009afc2cdaf3114f8c72530f0909e

Observation 0e8265e4-ae1c-47cd-b236-f55b4fc3419c · outbound

This paper cites TraceVLA: Visual trace prompting enhances spatial- temporal awareness for generalist robotic policies,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation TraceVLA: Visual trace prompting enhances spatial- temporal awareness for generalist robotic policies,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.773326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:17.477048Z digest=sha256:cf6d4150bef31f885a5eac075120b6b04fe4c8c3bb868e68febabdc6de0f2078

Observation f3038fbd-6cc5-4a77-a2eb-dfc5a65702df · outbound

This paper cites RoboMemory: A brain-inspired multi-memory agentic framework for interactive environmental learning in physical embodied systems,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RoboMemory: A brain-inspired multi-memory agentic framework for interactive environmental learning in physical embodied systems,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.604914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.604914Z digest=sha256:afa6c38fa28bb3fee9d8b82a8ff116a3e1467c9281b804bf4ecd7c0194dc6cb7

Observation 066b05a7-3096-4d78-a0b3-e3387bb86b98 · outbound

This paper cites MemoryVLA: Perceptual-cognitive memory in vision- language-action models for robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation MemoryVLA: Perceptual-cognitive memory in vision- language-action models for robotic manipulation,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.761545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:17.711839Z digest=sha256:e4d78cb0f3041dfe3589e1cfbd32ab21e5190ab852b19e9ab3add2486c4c6307

Observation f8a91fcd-be30-4561-8774-439ae3e83a76 · outbound

This paper cites Gated Memory Policy: In-Context Memorization and Adaptation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Gated Memory Policy: In-Context Memorization and Adaptation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.858828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.858828Z digest=sha256:61f0d70fec54a95ad4428b27fcfdd1586ad477f7bb89cd62d50302e9eacb8cb3

Observation b26ee186-b1d0-4ab4-8c02-0d0e5df61fff · outbound

This paper cites You only scan once: A dynamic scene reconstruction pipeline for 6-DoF robotic grasping of novel objects,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation You only scan once: A dynamic scene reconstruction pipeline for 6-DoF robotic grasping of novel objects,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.750220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:17.940226Z digest=sha256:7e150daa97b9f138e069ab0508fb35088595c5ebdd5ea3027fa31d553c291004

Observation 3619bfb8-af15-4b3b-91be-2ab580a5fb28 · outbound

This paper cites Mem-World: Memory-Augmented Action-Conditioned World Models for Persistent Robot Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Mem-World: Memory-Augmented Action-Conditioned World Models for Persistent Robot Manipulation

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:44:18.373945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:18.000401Z digest=sha256:7d13c0cc8e9a2cc87765d75c2240f9671d75d806099b0c4c7e745105b5b63e0d

Observation 62d4a375-70d3-4ef6-b231-9adb322ae235 · outbound

This paper cites Coarse-to-fine imitation learning: Robot manipulation from a single demonstration,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Coarse-to-fine imitation learning: Robot manipulation from a single demonstration,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.103744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.103744Z digest=sha256:2b318f776307ef8b04503d8b4126f0e35d0d98fa16790bc772bb8d474b212a15

Observation c7340192-801b-4d8f-bdc4-a939dbc5ee48 · outbound

This paper cites RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.160988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.160988Z digest=sha256:a73c5e6e8e1760db14417ee1ebc7f9a14366ee5deed254819f64c0fd5b005c20

Observation 501a8a7a-ba8b-4f0f-b121-2bf46dfbbda3 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation PaliGemma: A versatile 3B VLM for transfer

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.188176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.188176Z digest=sha256:632905c21f368fd17a2637e4f967e045f44fa1b64abb69649cd4662a63b177c7

Observation 1bcbf714-17a6-4890-b530-0d816e21e272 · outbound

This paper cites Sigmoid loss for language image pre-training,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Sigmoid loss for language image pre-training,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.200108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.200108Z digest=sha256:d711112ad6d8e2e839ba415587b80a0b5891a636982ffee5ace0dc568f215c2c

Observation 2bea7c9e-b7be-4b08-bd06-822a2d3a4ac7 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Gemma: Open Models Based on Gemini Research and Technology

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.203663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.203663Z digest=sha256:ed25ce99b07fa7d43b5391274b4f9ca6c3a05cf3a3a4a1a5e85bc8020feea013

Observation 3db468e6-720e-4258-be7d-34058747e6b7 · outbound

This paper cites RAFT: Recurrent all-pairs field transforms for optical flow,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RAFT: Recurrent all-pairs field transforms for optical flow,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.722936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:18.207095Z digest=sha256:8ad1a94bd137ee4ee907472af3037340fcf8c212fdee5ed477b0c9022ab035b5

Observation 6b10100f-42fc-4a01-b8d3-4adf5a1fce6c · outbound

This paper cites an unresolved cited work.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.210377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.210377Z digest=sha256:676eae8c7c413227f5655d44309dbf9eabf5da97b3664caa8048a4756f42c4e7

Observation 2ffc9fc0-45bf-4d03-b590-220ec82989dc · outbound

This paper cites On the continuity of rotation representations in neural networks,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation On the continuity of rotation representations in neural networks,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.213450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.213450Z digest=sha256:c3270f86273ae46751b40609f551708aff6a7d9b15981a03c79f042c566a02b9

Observation 06e3c500-0a7f-4650-bbef-286b2a71ad30 · outbound

This paper cites V-REP: A versatile and scalable robot simulation framework,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation V-REP: A versatile and scalable robot simulation framework,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.694734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:18.216594Z digest=sha256:aa0d91403ee81ad4b7a302fbe4202beb168c998b1ceacf02f8287db7938dd9f6

Observation ad897fab-5612-4722-991c-6aaf338a20a4 · outbound

This paper cites Perceiver IO: A general architecture for structured inputs & outputs,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Perceiver IO: A general architecture for structured inputs & outputs,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.682780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:18.220226Z digest=sha256:fe45a27419b4a5a92d3bfe63167ff46435195001c9887074f712b1e745188fff

Observation ec093242-9c4c-44f6-a66a-7587bc53cf7c · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Diffusion policy: Visuomotor policy learning via action diffusion,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.670900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:18.223196Z digest=sha256:7bb285cb5985750e6a0330bd95df02867040e13f61734453d72a416ad49deb77

Observation e2fd6278-b836-44b6-b6a4-eb6503c2a138 · outbound

This paper cites Learning fine-grained bimanual manipulation with low-cost hardware,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Learning fine-grained bimanual manipulation with low-cost hardware,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.226213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.226213Z digest=sha256:f4ec1f5159a1761c737c9885eb55501a5a23fcbc5c988fed6b2c9f5eb42080d6

Observation 78055c9f-cb59-46fe-bac4-338dfdddb5e7 · outbound

This paper cites X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.229394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.229394Z digest=sha256:caaf4b468ee4d699b4643bdb9cbfd769cf2e31b1eab33c3649eb667921d86deb

Observation 748933a2-7d86-4b24-82ee-7c8ab8cb267d · outbound

This paper cites Fast-WAM: Do World Action Models Need Test-time Future Imagination?.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Fast-WAM: Do World Action Models Need Test-time Future Imagination?

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.232308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.232308Z digest=sha256:5a1bc97f2bdd151f25bb71dbbb1322d16978c5543d2d00a0f8296a6558f29507

Observation 5477a936-baab-46d6-8daf-b8bae834b0e5 · outbound

This paper cites R3M: A Universal Visual Representation for Robot Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation R3M: A Universal Visual Representation for Robot Manipulation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.235552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.235552Z digest=sha256:f06ef6ba4bcafd89fc3028063994f3e35cd5237662b0686748ae25f41dd8f563

Observation 222af3a0-3646-46b8-a56f-b3f57fb18cd0 · outbound

This paper cites Masked Visual Pre-training for Motor Control.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Masked Visual Pre-training for Motor Control

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.238620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.238620Z digest=sha256:f5bab5a0c162d62dbd857d2c9d43f3b99729201a4ead5644c1d2e5fb0b41583d

Observation e3908d42-3017-427a-bc38-7ad0c4f92bfd · outbound

This paper cites RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.241421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.241421Z digest=sha256:46a1a2baab2dc491cc35f47e4c65b41dc8a97e77e9435b15118bf947fec2ed84

Observation 09676efc-b973-4a1e-b88d-813e430e0d34 · outbound

This paper cites SAPIEN: A simulated part-based interactive environ- ment,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation SAPIEN: A simulated part-based interactive environ- ment,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.650643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:18.244590Z digest=sha256:59223aa5a75c7f6d8262f80002ae0b5d4680959a3654fd1dc18e1fa1e084c5b9

Observation 884d5e23-93a0-4550-b4cc-d1e6055735fa · outbound

This paper cites Point transformer V3: Simpler faster stronger,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Point transformer V3: Simpler faster stronger,

Reference 65

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T10:44:18.638849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:18.247570Z digest=sha256:bf35f845a28ad00a5cd1bfa5a35e5382aac378ddedbb8ed7e6f2aada752fb96a

Pith citing papers

No inbound Pith citation observations are available.