Pith. sign in

Paper Citation Record · LEDGER

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation

As of 7 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 0 inbound Pith citation observations for arXiv:2608.05042.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05042 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:44:18.247570Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

65 of 65 outbound references displayed

  • verified exact2
  • verified fuzzy35
  • unresolved27
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation da635c3e-14ba-454c-ad3a-f0147571b3d3 · outbound

This paper cites OpenVLA: An open-source vision-language-action model,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation OpenVLA: An open-source vision-language-action model,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:11.836856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:11.836856Z digest=sha256:a4c70ddb307b096a4a514eb282855b4cbb21aaede8a13127809555a4206914db

Observation 30ef66ac-b196-4113-845a-b1c1b5b72135 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:11.920374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:11.920374Z digest=sha256:e4e75608a08e43d78c9b760fa9439fbf0e05136becf09d37081833d4abe6ece0

Observation 4c4286e5-ac21-4ebc-becb-0504efc4af64 · outbound

This paper cites Wall-OSS-0.5 Technical Report.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Wall-OSS-0.5 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:12.106434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:12.106434Z digest=sha256:c1076766a03b419a17b060d3898612c3984a69f174add05745b1462e6f5beaec

Observation 9ae2648e-564a-4774-8ef4-35d5a4bf726b · outbound

This paper cites Vision-language foundation models as effective robot imitators,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Vision-language foundation models as effective robot imitators,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.055685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:12.235379Z digest=sha256:68e2b208dfd507d5b154a99fce8ce71a8ef53708ba95bd29298d4d2731bcbb6d

Observation 72f95a25-9bc3-4002-aaa5-25cae51bd56a · outbound

This paper cites RT-2: Vision-language-action models transfer web knowledge to robotic control,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RT-2: Vision-language-action models transfer web knowledge to robotic control,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.046255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:12.390934Z digest=sha256:96ca97150db832c1147f9cd47187a1d6b75c18a24ad5986c0a379b84f1300f1f

Observation 36ec393a-1107-4976-8e73-5b110a128bc7 · outbound

This paper cites Perceiver-Actor: A multi-task transformer for robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Perceiver-Actor: A multi-task transformer for robotic manipulation,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.036180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:12.523105Z digest=sha256:5673cf143c6ac16bb6a647f436aa3b3352acd0c5fef7e959a03c1fcaea6320fb

Observation 62fa746e-6e30-4d58-bd9a-fbd27cc63775 · outbound

This paper cites 3D Diffuser Actor: Policy diffusion with 3D scene representations,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation 3D Diffuser Actor: Policy diffusion with 3D scene representations,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.027748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:12.644589Z digest=sha256:922a7d5662dfa4780e551674a9b1af37ebfb7983cb9c53e2dcca3f8c34b0e51f

Observation 8b5524d3-a092-490e-a37f-9b0b724fa9be · outbound

This paper cites Act3D: 3D feature field transformers for multi-task robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Act3D: 3D feature field transformers for multi-task robotic manipulation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.019261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:12.852518Z digest=sha256:95177ee6fae89a1a32ab64f5c8e83a0325c8301418a9f7a06c367c1e8d364901

Observation c172a669-c510-4cd3-84cc-a552d268c8b2 · outbound

This paper cites RVT: Robotic view transformer for 3D object manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RVT: Robotic view transformer for 3D object manipulation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.010416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:12.979457Z digest=sha256:98c0a03f2d9e1c5ba7ee8a91d3924c48525be158a81f72f52335a631ac59df04

Observation 132b4569-ec0d-4f4b-9f54-2f83a980c37f · outbound

This paper cites RVT-2: Learning precise manipulation from few demonstrations,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RVT-2: Learning precise manipulation from few demonstrations,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:19.001863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:13.078472Z digest=sha256:ea1ef20d637e19543f11e91b35fd61689ea7839367dd1ce8bc60180254d15c4c

Observation 111db71a-e83c-4bf4-ae07-e402e17c8359 · outbound

This paper cites 3D-VLA: A 3D vision-language-action generative world model,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation 3D-VLA: A 3D vision-language-action generative world model,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.993333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:13.210283Z digest=sha256:f13632068543559ab62e5890321507801f5141bb4ab880713c0a668f63ee0af6

Observation abec4001-6626-4f0a-aa60-e1d5f1aae51a · outbound

This paper cites SpatialVLA: Exploring spatial representations for visual- language-action models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation SpatialVLA: Exploring spatial representations for visual- language-action models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.984740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:13.350018Z digest=sha256:9bd6f1fc6992792a4c930a047cb47ee023a9dd2f8c1a9657aa945112b70d4a7e

Observation e7176949-b0a9-44bd-8dc7-6660c0d0d8c7 · outbound

This paper cites RLBench: The robot learning benchmark & learning environment,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RLBench: The robot learning benchmark & learning environment,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.976335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:13.499499Z digest=sha256:6aea862d3e434fe3a66c98997f2813accbd98a8eef312310c70b66584912ccb1

Observation edbb9f7e-c70c-4f75-96bf-10ee976757d3 · outbound

This paper cites THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:13.672028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:13.672028Z digest=sha256:a5199ce2df99355f290db748f14013ba812c6bbf8e8db26a642a7cc37c012087

Observation 6973dddc-827f-46de-a672-d9483518cd54 · outbound

This paper cites Towards generalizable vision- language robotic manipulation: A benchmark and LLM-guided 3D policy,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Towards generalizable vision- language robotic manipulation: A benchmark and LLM-guided 3D policy,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.966940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:13.754442Z digest=sha256:cc517e3593fbb82f68ea921b14999d5ba9347ff392647745cfea1c2755218656

Observation fe57631e-5ac7-41f1-a37f-d84c9630ec0d · outbound

This paper cites RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:13.849049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:13.849049Z digest=sha256:9084339592aa79be20f942d5c10bd20e72f03960ad278404b9655f489e28ddfc

Observation 10fe969f-77ba-4af8-8513-e8fb104d71cf · outbound

This paper cites SAM2Act: Integrating visual foundation model with a memory architecture for robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation SAM2Act: Integrating visual foundation model with a memory architecture for robotic manipulation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.956456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:13.970866Z digest=sha256:572956e5c056a6ab76f7abb4074144a18aa28f5b80a742af2cb1fe459589b14e

Observation 496c8dc8-a445-49a6-8865-e1fe4a147bc7 · outbound

This paper cites BridgeVLA: Input-output alignment for efficient 3D ma- nipulation learning with vision-language models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation BridgeVLA: Input-output alignment for efficient 3D ma- nipulation learning with vision-language models,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.945963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:14.064287Z digest=sha256:3cc754f6eeb691684ccf288b804f9d6f1193541e73792fbf2db4191e6338af0a

Observation b4bcc8d6-2459-47ba-b90b-71f01c45a501 · outbound

This paper cites RT-1: Robotics transformer for real-world control at scale,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RT-1: Robotics transformer for real-world control at scale,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.935088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:14.135586Z digest=sha256:d966d2612145978e4082f3828e825915963048a5ede94a76bcd36796b0a6b980

Observation a9af96c5-bcb5-4b3d-a666-3695f9e631c1 · outbound

This paper cites π 0: A vision-language-action flow model for general robot control,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation π 0: A vision-language-action flow model for general robot control,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.924717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:14.302562Z digest=sha256:79b2383cd1bb30229834820bd60fdcbceafbf920d696a91423affe44f360d3cf

Observation 371af7de-dfa1-40a4-806f-f70e409abec5 · outbound

This paper cites FAST: Efficient action tokenization for vision- language-action models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation FAST: Efficient action tokenization for vision- language-action models,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.914178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:14.470124Z digest=sha256:28cc6ec085590988b24c8aa542203654ae9a47a40ddff752a0581f5b8daffdb6

Observation 2c2b5ade-1432-48d0-830c-4575e227a1bf · outbound

This paper cites $\pi^{*}_{0.6}$: a VLA That Learns From Experience.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation $\pi^{*}_{0.6}$: a VLA That Learns From Experience

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:14.652245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:14.652245Z digest=sha256:5e5f65d6d20968fc42fcd517c80789c0af3888f1ea99edc6e35f4bfc9d270681

Observation 371df10c-c4af-4cec-bc5e-716115ee111e · outbound

This paper cites ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:14.797218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:14.797218Z digest=sha256:264251178dbd9ef506515b7dd77e6f91ce9701723e7e0bc0f45d18d77ecd6ee5

Observation 8cf5754c-5969-4399-95a6-b3c9f23c9da4 · outbound

This paper cites GEN-0: Embodied foundation models that scale with physical interaction,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation GEN-0: Embodied foundation models that scale with physical interaction,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.902689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:14.955267Z digest=sha256:f7ae10543b66e4cf964c57bddbb4ae18004a0486c9fcbce539f91e5f801971d6

Observation 5ddf7257-1253-40f0-a59d-5e17719dbaa1 · outbound

This paper cites GEN-1: Scaling embodied foundation models to mastery,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation GEN-1: Scaling embodied foundation models to mastery,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.892286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:15.131468Z digest=sha256:a4854d4ce1c999cf34438d018efdfcabb13a19ce86abe9465bfc8fc5fb1f9872

Observation ec1cd722-0cfb-4adb-babe-5a00236107cd · outbound

This paper cites GENE-26.5: Advancing robotic manipulation to human level,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation GENE-26.5: Advancing robotic manipulation to human level,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.881643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:15.233329Z digest=sha256:747c979e61c7fc35fa20b9535839b9b6e63f55df70ff64b62dd77a5a61c733c7

Observation 821c082a-2eb8-4751-a1ee-4bec48253783 · outbound

This paper cites ACT-2 preview: Generalizing reliability,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation ACT-2 preview: Generalizing reliability,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.871047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:15.384527Z digest=sha256:200de1ffc9b83f85ff23476971811416cd118e54b269284b001efa943f5cbfed

Observation e1484038-fc1c-4daa-839f-7bfc9bf6ca4f · outbound

This paper cites Open X-Embodiment: Robotic learning datasets and RT-X models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Open X-Embodiment: Robotic learning datasets and RT-X models,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.860168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:15.548311Z digest=sha256:73c677c21672ed5776c4817a9ce68c7c28c3d38f7e8f5753b6f0444f13cc2851

Observation cd1f516d-f3ef-42f3-9164-897156e98ac9 · outbound

This paper cites PolarNet: 3D point clouds for language-guided robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation PolarNet: 3D point clouds for language-guided robotic manipulation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.848674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:15.653628Z digest=sha256:7fd316677a8e85a2c2077beb4c2520bdce5357fd30523be1a41d89c1c1952cd6

Observation eb140a8d-ce8b-4c83-9b44-15e09803f790 · outbound

This paper cites M2T2: Multi-task masked transformer for object-centric pick and place,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation M2T2: Multi-task masked transformer for object-centric pick and place,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.837170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:15.829628Z digest=sha256:14da8d7f90c9e3810835dbca2429c9234e4f0228a899ed925d8490c810d5107b

Observation c4caa6bb-3519-40db-92e1-62c5f229d026 · outbound

This paper cites FP3: A 3D Foundation Policy for Robotic Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation FP3: A 3D Foundation Policy for Robotic Manipulation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:16.040209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:16.040209Z digest=sha256:d2a837468a037f1771b5e904dc040c55e8f113ecfac5a3a0dc64e722d1e3bd7a

Observation 12c3079f-f9fd-4bd7-b466-94c5828442f2 · outbound

This paper cites Coarse-to-fine Q- attention: Efficient learning for visual robotic manipulation via discreti- sation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Coarse-to-fine Q- attention: Efficient learning for visual robotic manipulation via discreti- sation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.826462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:16.225713Z digest=sha256:8f52ac25df20296dfc10ffa43be1d1e73e1428b42b8d91ace9d663f392c2b923

Observation deeacec7-9962-42cf-ab9b-33007842f9c6 · outbound

This paper cites PointVLA: Injecting the 3D world into vision-language-action models,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation PointVLA: Injecting the 3D world into vision-language-action models,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.814607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:16.363898Z digest=sha256:7f48fd3102c23bef331dac8d1a8a0afa662e6cd8b104ab168233f34e5f36df14

Observation 518fea34-2208-4803-80f6-e2a854ba2e3a · outbound

This paper cites Lift3D policy: Lifting 2D foundation models for robust 3D robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Lift3D policy: Lifting 2D foundation models for robust 3D robotic manipulation,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.803293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:16.530929Z digest=sha256:490689e547f427ca2652f78df2b956a4cd61b03528ce77da9583abac2dc2e9b9

Observation deb3e36c-995d-4e7d-9f68-269b8fdf977a · outbound

This paper cites DINOv2: Learning robust visual features without supervision,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation DINOv2: Learning robust visual features without supervision,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.791935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:16.777096Z digest=sha256:c738c65906532cef3d28bd8f5b853c7b578686cd6cd2677714f15e6b3b24770f

Observation dd8f4b74-cad6-4f31-b450-c15e922ec900 · outbound

This paper cites OG- VLA: Orthographic image generation for 3D-aware vision-language action model,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation OG- VLA: Orthographic image generation for 3D-aware vision-language action model,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:16.921827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:16.921827Z digest=sha256:a44fb5ea08687f97c49b0beba9f7b8ac48cea6e74e8b39f67aa51cc07d685f90

Observation 0c222d1c-b006-4f1c-a73e-1d935006267a · outbound

This paper cites Instruction-driven history-aware policies for robotic manipulations,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Instruction-driven history-aware policies for robotic manipulations,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.011327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.011327Z digest=sha256:e0a059bbb2fdcc105183960cd9c27688744934e5b520a7c12a732f4324f5fa1c

Observation e8352d45-348c-455d-9e25-5fba4dab6404 · outbound

This paper cites Causal World Modeling for Robot Control.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Causal World Modeling for Robot Control

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.098344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.098344Z digest=sha256:bb6116fcfb0dee51b03598e3b8792b0844c182fd1568dca4bedd7024870c275e

Observation 19f9ba14-2da4-423f-9532-901e793b4ed8 · outbound

This paper cites World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:44:18.473850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:17.241212Z digest=sha256:e36ba97fd2b45ba28c349f5e9b8d9e6be026c4840a6d8f5b41b66864d4d7039a

Observation f0b9dd4f-6e94-459e-9229-cbfa958934ba · outbound

This paper cites MemoryWAM: Efficient World Action Modeling with Persistent Memory.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation MemoryWAM: Efficient World Action Modeling with Persistent Memory

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.373966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.373966Z digest=sha256:e4f15e7d409e7bf3d7377cb8685e54e47c2cc81bcdb35c82f640cf5c81a8b9af

Observation 0e8265e4-ae1c-47cd-b236-f55b4fc3419c · outbound

This paper cites TraceVLA: Visual trace prompting enhances spatial- temporal awareness for generalist robotic policies,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation TraceVLA: Visual trace prompting enhances spatial- temporal awareness for generalist robotic policies,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.773326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:17.477048Z digest=sha256:6d7bf2b642470ece0c9445443d3d8eb046f67df2637f27c3530f30a306b4064f

Observation f3038fbd-6cc5-4a77-a2eb-dfc5a65702df · outbound

This paper cites RoboMemory: A brain-inspired multi-memory agentic framework for interactive environmental learning in physical embodied systems,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RoboMemory: A brain-inspired multi-memory agentic framework for interactive environmental learning in physical embodied systems,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.604914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.604914Z digest=sha256:e3127aec89f702d27eaba9ce814ac36ca95fba64d6847b2e9daf3f19b10d7b49

Observation 066b05a7-3096-4d78-a0b3-e3387bb86b98 · outbound

This paper cites MemoryVLA: Perceptual-cognitive memory in vision- language-action models for robotic manipulation,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation MemoryVLA: Perceptual-cognitive memory in vision- language-action models for robotic manipulation,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.761545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:17.711839Z digest=sha256:d366cb8a3248b3123e8eeb5503c633eb9d7598bd6eaecd56f348ce29e41eaaf1

Observation f8a91fcd-be30-4561-8774-439ae3e83a76 · outbound

This paper cites Gated Memory Policy: In-Context Memorization and Adaptation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Gated Memory Policy: In-Context Memorization and Adaptation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:17.858828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:17.858828Z digest=sha256:382bddf21545f684051ca83081e14d331dd306a5249ac3a0e7ca0fec457b037e

Observation b26ee186-b1d0-4ab4-8c02-0d0e5df61fff · outbound

This paper cites You only scan once: A dynamic scene reconstruction pipeline for 6-DoF robotic grasping of novel objects,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation You only scan once: A dynamic scene reconstruction pipeline for 6-DoF robotic grasping of novel objects,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.750220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:17.940226Z digest=sha256:c041fa7fb12220f182e84cc3d18ed533f65a514b89a7f991e2db9be0e75deda3

Observation 3619bfb8-af15-4b3b-91be-2ab580a5fb28 · outbound

This paper cites Mem-World: Memory-Augmented Action-Conditioned World Models for Persistent Robot Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Mem-World: Memory-Augmented Action-Conditioned World Models for Persistent Robot Manipulation

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:44:18.373945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:18.000401Z digest=sha256:eaf884e0aa19a2b37c41bc08e0e527528ecdd0ab73bc8076114b78ba089c958d

Observation 62d4a375-70d3-4ef6-b231-9adb322ae235 · outbound

This paper cites Coarse-to-fine imitation learning: Robot manipulation from a single demonstration,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Coarse-to-fine imitation learning: Robot manipulation from a single demonstration,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.103744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.103744Z digest=sha256:3ef9c5b73d897325e33163eaa720f8a9465ce16c64ffdfa8aa7b6a28a3a5768d

Observation c7340192-801b-4d8f-bdc4-a939dbc5ee48 · outbound

This paper cites RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.160988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.160988Z digest=sha256:29dcdcc5190f627cecce7b8c22bc3b745022ad0d38a46e73e2998314e9b9941f

Observation 501a8a7a-ba8b-4f0f-b121-2bf46dfbbda3 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation PaliGemma: A versatile 3B VLM for transfer

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.188176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.188176Z digest=sha256:63ff9f83406a07a34de0196fed409dea40e5b83648e84c6c4661405fe5640b78

Observation 1bcbf714-17a6-4890-b530-0d816e21e272 · outbound

This paper cites Sigmoid loss for language image pre-training,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Sigmoid loss for language image pre-training,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.200108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.200108Z digest=sha256:48b834537a070c0900687a20f4346171d80d892a282e87a54f6f64b1e9a3961d

Observation 2bea7c9e-b7be-4b08-bd06-822a2d3a4ac7 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Gemma: Open Models Based on Gemini Research and Technology

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.203663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.203663Z digest=sha256:dd7f96895c0d9ad31b4075f08721d994b7d51c7d489110ae5ae492351744c0f0

Observation 3db468e6-720e-4258-be7d-34058747e6b7 · outbound

This paper cites RAFT: Recurrent all-pairs field transforms for optical flow,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RAFT: Recurrent all-pairs field transforms for optical flow,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.722936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:18.207095Z digest=sha256:fee22005dab4804312762412ef64ed6df6af873f9e61acc5df73c61fabeabc7e

Observation 6b10100f-42fc-4a01-b8d3-4adf5a1fce6c · outbound

This paper cites an unresolved cited work.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.210377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.210377Z digest=sha256:0247e63f22ac4ea4ad94b62cb620a9b4c8ed632bdaf9ba566b5306ac49f0ae45

Observation 2ffc9fc0-45bf-4d03-b590-220ec82989dc · outbound

This paper cites On the continuity of rotation representations in neural networks,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation On the continuity of rotation representations in neural networks,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.213450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.213450Z digest=sha256:28e9573a4309b95a982606abc7cde489fc823946172f50f78ef9a29a1de6714b

Observation 06e3c500-0a7f-4650-bbef-286b2a71ad30 · outbound

This paper cites V-REP: A versatile and scalable robot simulation framework,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation V-REP: A versatile and scalable robot simulation framework,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.694734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:18.216594Z digest=sha256:f3d03578b8e3384a52cba7dc4afad1beaa5ee413df6c8b6d9f54064d61f3179b

Observation ad897fab-5612-4722-991c-6aaf338a20a4 · outbound

This paper cites Perceiver IO: A general architecture for structured inputs & outputs,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Perceiver IO: A general architecture for structured inputs & outputs,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.682780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:18.220226Z digest=sha256:be286d0df78e93dba3c614968c139e87bfde986a4301ef9ee7b15a57edaa4d25

Observation ec093242-9c4c-44f6-a66a-7587bc53cf7c · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Diffusion policy: Visuomotor policy learning via action diffusion,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.670900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:18.223196Z digest=sha256:95d070ed402a94776071bfac5ed53c131911af1a84022f27f02b515f71e8f117

Observation e2fd6278-b836-44b6-b6a4-eb6503c2a138 · outbound

This paper cites Learning fine-grained bimanual manipulation with low-cost hardware,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Learning fine-grained bimanual manipulation with low-cost hardware,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.226213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.226213Z digest=sha256:ceceda0393fc37af1a20e19b424e9bb2b1ff018861e61fd486a94a721fd349f2

Observation 78055c9f-cb59-46fe-bac4-338dfdddb5e7 · outbound

This paper cites X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.229394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.229394Z digest=sha256:51a72406a88e6508dcdadbd5c3bc74560b2e4a2dfb9a47ea05814328dfce8067

Observation 748933a2-7d86-4b24-82ee-7c8ab8cb267d · outbound

This paper cites Fast-WAM: Do World Action Models Need Test-time Future Imagination?.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Fast-WAM: Do World Action Models Need Test-time Future Imagination?

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.232308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.232308Z digest=sha256:f98a7367e976aeeb55fb0081d1188fc4e8b67c0cf803406e0f0ba2b602f93912

Observation 5477a936-baab-46d6-8daf-b8bae834b0e5 · outbound

This paper cites R3M: A Universal Visual Representation for Robot Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation R3M: A Universal Visual Representation for Robot Manipulation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.235552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.235552Z digest=sha256:5b5153c9bd2c50fae9d6533b680ef6af542e6f97d25cb74bdd33e35acbb48db2

Observation 222af3a0-3646-46b8-a56f-b3f57fb18cd0 · outbound

This paper cites Masked Visual Pre-training for Motor Control.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Masked Visual Pre-training for Motor Control

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.238620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.238620Z digest=sha256:07580e202fbc3fafcb82d6351198e51ad65c2fa2edd4274268f4d982eefc824d

Observation e3908d42-3017-427a-bc38-7ad0c4f92bfd · outbound

This paper cites RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T10:44:18.241421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:44:18.241421Z digest=sha256:ad0db44ab5d3d027d848adc90a6925a7f4df100d5488a4cf31595253b95d5e2c

Observation 09676efc-b973-4a1e-b88d-813e430e0d34 · outbound

This paper cites SAPIEN: A simulated part-based interactive environ- ment,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation SAPIEN: A simulated part-based interactive environ- ment,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:44:18.650643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:18.244590Z digest=sha256:5f40f54771081b7cfefc3a380a52ec30249899a9e5b132033f8dfc04a6ff97f5

Observation 884d5e23-93a0-4550-b4cc-d1e6055735fa · outbound

This paper cites Point transformer V3: Simpler faster stronger,.

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Point transformer V3: Simpler faster stronger,

Reference 65

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T10:44:18.638849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T10:44:18.247570Z digest=sha256:21550fd3cbdcdbc533124cfdcab11c895770c99d24bf593afab9cf86473a146c

Pith citing papers

No inbound Pith citation observations are available.