Pith. sign in

Paper Citation Record · LEDGER

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model

As of 5 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2605.22671.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.22671 v2

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-30T16:52:05.565650Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact25
  • verified fuzzy3
  • unresolved2
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation edf2a703-f21b-4b59-91fc-96c0e1a2061b · outbound

This paper cites Dexterous manipulation through imitation learning: A survey.arXiv preprint arXiv:2504.03515.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Dexterous manipulation through imitation learning: A survey.arXiv preprint arXiv:2504.03515

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.323038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:3f2be97d92af1124ebe9d134172e2ef2e99a2926a573d9cb45172f380afa9a2f

Observation 42a34d34-0ef1-4e4c-a273-f7230f411361 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.216782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:dbd5548ab675ef1243e148ccc58db6c17e4da8b620c868dac0baf634fa7870de

Observation 6bdb2d9b-0ad3-4166-8465-67e835f68bd6 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 4

Resolution
malformed identifier
local_arxiv, observed 2026-06-30T16:54:58.214010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:5565d8ee5cb0c039cffc4ec36ad9127f6318865920cb0a6ecc4470da6d5a61c2

Observation 1f13d2af-bea4-49d6-aba8-249e78b13760 · outbound

This paper cites Towards Synergistic, Generalized, and Efficient Dual-System for Robotic Manipulation.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Towards Synergistic, Generalized, and Efficient Dual-System for Robotic Manipulation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.233817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:feb23051fe389ebaf2db558d1661b9189860acdd1f2ef571720cc07755875d0a

Observation b0569cfb-8cd5-40f0-84b6-150fda208485 · outbound

This paper cites UniVLA: Learning to Act Anywhere with Task-centric Latent Actions.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.236871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:2c36ca4023c6585775380db4312ef47c322b927dcf72d11d9c6dc33ef542659e

Observation f9b97e5e-3184-4795-b1c5-16552b6f11b5 · outbound

This paper cites Fast-in-Slow: A Dual-System Foundation Model Unifying Fast Manipulation within Slow Reasoning.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Fast-in-Slow: A Dual-System Foundation Model Unifying Fast Manipulation within Slow Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.227637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:71b9de521f3ed26bba86d4076f350f7e83435a8af2081ed4e00e0105e224feee

Observation 7abd9c0b-9a9f-42e7-a012-a43ea7361629 · outbound

This paper cites Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.230285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:458e322ccd7efbea977a9746643f701841eedcfc8b8b26181c001923dc676804

Observation 7f88e7fa-b614-4824-ba0f-d15e1f5c9d3e · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.345421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:41c7c2fd30b33eb5b1f18d9ed0137d3f12de8b6c7c7d735e43f24464e2d8b37e

Observation 334b672b-37c1-4fd0-ac2f-21edf41037b5 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model OpenVLA: An Open-Source Vision-Language-Action Model

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T16:54:58.219324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:186a32083d74d1d6772cf5427bfaa0a34c6ccaf62ce2f0c82ed41fa5ca455deb

Observation 9e10d44b-f737-44f2-98d7-233acb10287b · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T16:54:58.274271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:36412ef2ad374849a2b4ec4a2bc55288ffeb29c8b6cc09fe7b2751d78a8a900a

Observation c4825ece-e642-4d71-ad7f-8142bccfefc3 · outbound

This paper cites Behavior Generation with Latent Actions.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Behavior Generation with Latent Actions

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.223988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:7b6fcd310b68485adbc8c46a536195097f4b64b767735df78efbf53f74c2b44f

Observation 134aa164-34ca-4cba-aeb5-4e13ff9a65d3 · outbound

This paper cites PointVLA: Injecting the 3D World into Vision-Language-Action Models.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model PointVLA: Injecting the 3D World into Vision-Language-Action Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.317685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:e4ed90680d7f43993c2a59994fa01ff211c056329bf9e8512db780324f2f2e1b

Observation 618c47b4-d8e8-4042-b9c3-43a2b8ed51f3 · outbound

This paper cites Optimus-3: Towards generalist multi- modal minecraft agents with scalable task experts.arXiv preprint arXiv:2506.10357, 2025f.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Optimus-3: Towards generalist multi- modal minecraft agents with scalable task experts.arXiv preprint arXiv:2506.10357, 2025f

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.239700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:27c5ab6bebe75a80258f1c70db9397db6779e26280f868c6920d490c1922eaee

Observation 1f1d091f-cb2a-45b1-840d-23471d374259 · outbound

This paper cites 10 From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Lin, T., Zhang, Y ., Li, Q., Qi, H., Yi, B., Levine, S., and Malik, J.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model 10 From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Lin, T., Zhang, Y ., Li, Q., Qi, H., Yi, B., Levine, S., and Malik, J

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.262599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:72e351a50768f5ff40083e0ec43d0d701e8fc5f7332c4936cbd92d1a63193472

Observation bf497a9b-e28d-4e57-bb2a-4a0c1a9dc739 · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.339954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:b788ee6094f110e25813d4f9373b93011466ba03ae9a6c360d604aecf09e715e

Observation 56913f0b-eeba-4508-bc15-9c168cd73f55 · outbound

This paper cites FAST: Efficient Action Tokenization for Vision-Language-Action Models.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model FAST: Efficient Action Tokenization for Vision-Language-Action Models

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.310840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:163186abb5930ac430a2add9429e09eac3f64e1001372daa150da92f0edada4e

Observation aabec67e-ffff-4845-b570-e94cc08d353b · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.255600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:bbc85d02556899b30d4d098fe3651f07c573b4b214c80b1009e2c3b9fb3b3c26

Observation b587474c-e797-4f3d-ab87-d06a0091dcf7 · outbound

This paper cites Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.258837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:169733531035c21822483a3ef9a7860282d4241f6c4e556ea5870cc8966def29

Observation 775e8523-e2fb-43b7-b40c-e6264aac3d29 · outbound

This paper cites Hats: Hardness-aware trajectory syn- thesis for gui agents.arXiv preprint arXiv:2603.12138.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Hats: Hardness-aware trajectory syn- thesis for gui agents.arXiv preprint arXiv:2603.12138

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.328683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:80154bf1fd668f004c05f0c308c3bda11048977057083e29447cbeada4d8d011

Observation b3b9ec32-92ec-413f-b0d7-894cd808009d · outbound

This paper cites MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.250065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:71032c4d7c496fe58e6db67d7d88891047bb048fccf9ae9e0ef44ce81036931c

Observation 8b097ac1-bc3f-4f23-9af9-adcdf580fa00 · outbound

This paper cites SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.242954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:5c710dcc4a65c1dcea66237cff587878bdf34fb6b7df462b2335bd68a1298269

Observation b45b1ea2-afc0-4140-a249-687f7ee2fdc2 · outbound

This paper cites ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.253115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:660a58f2f5f5f5d2d6c5ccdb06d9e05f14086a96442de3c36f62c1d05a852bf5

Observation 16baf14c-cdb9-43d9-bd1a-24cc032789b7 · outbound

This paper cites Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.246432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:742e2d117526c538d1b36be46acbec674916e0dc86b850c8ede4f65e0800f16d

Observation d7a5fd1e-bd22-4e94-bdcd-8398898ae269 · outbound

This paper cites VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.265667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:5298d07497fad5a6c3381a319d4d0ff694baaa5fdc803972f7bb6741fa151350

Observation 5532b8f5-0744-4180-bb16-5e52412c8bd7 · outbound

This paper cites Embodied-RAG: General Non-parametric Embodied Memory for Retrieval and Generation.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Embodied-RAG: General Non-parametric Embodied Memory for Retrieval and Generation

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.269682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:e3e9e449ed2ea1250ce249c45880dc75ae6de6dcf21bae37016ef56c57574d7a

Observation 11ad7e9d-a6ed-4f93-acc6-ef0e6cf19d17 · outbound

This paper cites Mirage-1: Augmenting and Updating GUI Agent with Hierarchical Multimodal Skills.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Mirage-1: Augmenting and Updating GUI Agent with Hierarchical Multimodal Skills

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.334615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:ffe434b431e25a4a685daef7521c99a323cb802a1655a8eeb8c76031b4f5332f

Observation 390260fa-235b-4580-8657-fae872950668 · outbound

This paper cites RoboSSM: Scalable In-context Imitation Learning via State-Space Models.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model RoboSSM: Scalable In-context Imitation Learning via State-Space Models

Reference 28

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T16:54:58.298368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:ca326405a22ddc73d7e1baf29300446b7aa32ea098a4a82009fe5586bdea0116

Observation 39d1ad7d-3f52-4139-b1a4-414c8556d667 · outbound

This paper cites 3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model 3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.289934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:0185752a8b836d8277b1b82707e35a17227b1ed50b132122347b22ec72bdb6e9

Observation d45834e8-781a-4585-a3e2-d40fe382e7bb · outbound

This paper cites UP-VLA: A Unified Understanding and Prediction Model for Embodied Agent.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model UP-VLA: A Unified Understanding and Prediction Model for Embodied Agent

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.304885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:57f42716ad9e5d178cd9d2806714d394f6815de2cbc926a7270cad8c838c99a6

Observation 05d4939f-8b83-404d-8912-30cd78ec3e9e · outbound

This paper cites an unresolved cited work.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-07-08T11:54:52.451863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:bb18df62777ced201fdbe018da3e6de3650b08aeecbbbd92eb09f371711c7bf9

Observation 9f810584-b1c2-4893-9e78-57952f622fb8 · outbound

This paper cites an unresolved cited work.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-07-08T11:54:52.455419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:55d9b8e8ee7df9b5a7e62f6557397dd6c340dc6db2d67518cd476bfe0cf742af

Observation 57ef0521-44ea-44c7-90a3-3085e62cc4fc · outbound

This paper cites MAP-VLA(Li et al., 2025c) further reduces fragment inconsistency through stage-wise segmentation and alignment.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model MAP-VLA(Li et al., 2025c) further reduces fragment inconsistency through stage-wise segmentation and alignment

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T11:54:52.448421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:e4d95232c21c58b31e9ed0ae8163d988e7060d141dcac914c3148b8822f44e64

Observation 3218a125-65eb-4a2c-9548-4532cf40b97e · outbound

This paper cites Recent VLAs(Black et al., 2025; Black et al.; Shukor et al.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Recent VLAs(Black et al., 2025; Black et al.; Shukor et al

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T11:54:52.442343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:47b8a8ab94d412e83ce0a6dc6d8542937e4ca8e3a82e790e9d793236c9287094

Observation 6fdffb18-654a-4066-a183-d73e451a2893 · outbound

This paper cites distribution shift.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model distribution shift

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T11:54:52.445113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:d4cc078ca77c9b0397d5392ef5d4a0dba8dec3ccc81039ff80f77e0ea4f18b98

Pith citing papers

No inbound Pith citation observations are available.