Pith. sign in

Paper Citation Record · LEDGER

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model

As of 19 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2605.22671.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.22671 v2

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-30T16:52:05.565650Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact25
  • verified fuzzy3
  • unresolved2
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation edf2a703-f21b-4b59-91fc-96c0e1a2061b · outbound

This paper cites Dexterous manipulation through imitation learning: A survey.arXiv preprint arXiv:2504.03515.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Dexterous manipulation through imitation learning: A survey.arXiv preprint arXiv:2504.03515

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.323038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:4a2618d1bb3f1f6465581b221373862097bc2862046aa58714e7143f10a2c1ba

Observation 42a34d34-0ef1-4e4c-a273-f7230f411361 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.216782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:f347713b1c63fc722f567a719a53b347a28b0f2026dae60ec29d1c487b604d1a

Observation 6bdb2d9b-0ad3-4166-8465-67e835f68bd6 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 4

Resolution
malformed identifier
local_arxiv, observed 2026-06-30T16:54:58.214010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:359d612a905d7f85c534904475e6bf70eefe7f1a14f12ab1527d222b4b9bab51

Observation 1f13d2af-bea4-49d6-aba8-249e78b13760 · outbound

This paper cites Towards Synergistic, Generalized, and Efficient Dual-System for Robotic Manipulation.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Towards Synergistic, Generalized, and Efficient Dual-System for Robotic Manipulation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.233817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:28f4077b0f77d0644fde2cc9cdbe48e0de2d92dcd5b20e37e2a2bb983daa6f70

Observation b0569cfb-8cd5-40f0-84b6-150fda208485 · outbound

This paper cites UniVLA: Learning to Act Anywhere with Task-centric Latent Actions.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.236871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:ef248677886e84c2c0a464aae3bfd9f237749b8aec9a66852fb53b84123c2a6b

Observation f9b97e5e-3184-4795-b1c5-16552b6f11b5 · outbound

This paper cites Fast-in-Slow: A Dual-System Foundation Model Unifying Fast Manipulation within Slow Reasoning.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Fast-in-Slow: A Dual-System Foundation Model Unifying Fast Manipulation within Slow Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.227637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:8fb67a5443e9e9eb6f1e47e9918e72fff2ac7d60b178d28f540e5c4ab3190558

Observation 7abd9c0b-9a9f-42e7-a012-a43ea7361629 · outbound

This paper cites Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.230285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:ce23f01912aebd1d9eb4d267cb4db51a66a07db6824484f9f3f34a0ca520d6f2

Observation 7f88e7fa-b614-4824-ba0f-d15e1f5c9d3e · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.345421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:e2a40751fe5a1d45688d4d88891014f4be1374755706e68dea7dc6652febabc5

Observation 334b672b-37c1-4fd0-ac2f-21edf41037b5 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model OpenVLA: An Open-Source Vision-Language-Action Model

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T16:54:58.219324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:e87abc290271e4a6bf5a2160b99cd3c690587ebe20cd5f33312f3df5eb0f914e

Observation 9e10d44b-f737-44f2-98d7-233acb10287b · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T16:54:58.274271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:f92c8dd8919e14482a342cbca068578d258b1cb19398c94f72891b51c6485342

Observation c4825ece-e642-4d71-ad7f-8142bccfefc3 · outbound

This paper cites Behavior Generation with Latent Actions.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Behavior Generation with Latent Actions

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.223988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:cd9dff27b8bea1a04f494e2dee9ba1b0c5cbc1ba5bcab59a68ddbd1899eeb78c

Observation 134aa164-34ca-4cba-aeb5-4e13ff9a65d3 · outbound

This paper cites PointVLA: Injecting the 3D World into Vision-Language-Action Models.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model PointVLA: Injecting the 3D World into Vision-Language-Action Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.317685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:aa63c8bf90f8bce91b76ca748396efeab6d1c46434423bc561a87a7ac9af01a1

Observation 618c47b4-d8e8-4042-b9c3-43a2b8ed51f3 · outbound

This paper cites Optimus-3: Towards generalist multi- modal minecraft agents with scalable task experts.arXiv preprint arXiv:2506.10357, 2025f.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Optimus-3: Towards generalist multi- modal minecraft agents with scalable task experts.arXiv preprint arXiv:2506.10357, 2025f

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.239700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:47ec2a764d425f7e4c622497766a753bb619bca8f6f3c047fa4b46b25e69c092

Observation 1f1d091f-cb2a-45b1-840d-23471d374259 · outbound

This paper cites EchoVLA: Robotic Vision-Language-Action Model with Synergistic Declarative Memory for Mobile Manipulation.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model EchoVLA: Robotic Vision-Language-Action Model with Synergistic Declarative Memory for Mobile Manipulation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-08-10T02:09:13.785697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:ef57c9bc82245612c7a4548c7810d1ad3490cb8ae6c467a91ecf6c20f8e1e8b8

Observation bf497a9b-e28d-4e57-bb2a-4a0c1a9dc739 · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.339954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:87a831e581c21919b0ab5df3928074c51b93d2e3738ae1414c6d1c79977c1cc2

Observation 56913f0b-eeba-4508-bc15-9c168cd73f55 · outbound

This paper cites FAST: Efficient Action Tokenization for Vision-Language-Action Models.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model FAST: Efficient Action Tokenization for Vision-Language-Action Models

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.310840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:382d1f0fe32494b4b4a7b68862ade89d51ec54badae144f6c21c7151bccc527b

Observation aabec67e-ffff-4845-b570-e94cc08d353b · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.255600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:b13e8a2d5e2b574a56b08ba8017570b31fd7fdbfff954da245bfed9b6026042c

Observation b587474c-e797-4f3d-ab87-d06a0091dcf7 · outbound

This paper cites Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.258837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:976950076f16e2ffc6b011564e2fe13880828989bd533ab47af655c8a28d3c51

Observation 775e8523-e2fb-43b7-b40c-e6264aac3d29 · outbound

This paper cites Hats: Hardness-aware trajectory syn- thesis for gui agents.arXiv preprint arXiv:2603.12138.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Hats: Hardness-aware trajectory syn- thesis for gui agents.arXiv preprint arXiv:2603.12138

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.328683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:6ee2a976e7aa0c8dedd002243c19cd3b97b1e2ef35906ef5c626f97cf5036f64

Observation b3b9ec32-92ec-413f-b0d7-894cd808009d · outbound

This paper cites MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.250065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:fd3382ebef2e5bce0fd808aefa68197fddfd767db5e951f0a0c63ed0cad5506f

Observation 8b097ac1-bc3f-4f23-9af9-adcdf580fa00 · outbound

This paper cites SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.242954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:141eed8c02d970511bd7e34fa81450f8914f5ead2215ebcd899f597cb7b5ce6e

Observation b45b1ea2-afc0-4140-a249-687f7ee2fdc2 · outbound

This paper cites ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.253115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:0c022c754b0c253f29b82b13fcd184f688a165cd5f82c2db6e2b0b95a9140dab

Observation 16baf14c-cdb9-43d9-bd1a-24cc032789b7 · outbound

This paper cites Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.246432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:665f8c322db2055788168c73f6aec5add2d628bdaa496141095b4dcf7cfcb016

Observation d7a5fd1e-bd22-4e94-bdcd-8398898ae269 · outbound

This paper cites VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.265667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:1e8bd5cab0411d1aede64e515300e86a2a1e43124c404b1153680669714e97d8

Observation 5532b8f5-0744-4180-bb16-5e52412c8bd7 · outbound

This paper cites Embodied-RAG: General Non-parametric Embodied Memory for Retrieval and Generation.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Embodied-RAG: General Non-parametric Embodied Memory for Retrieval and Generation

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.269682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:f5be9ce993f46cd4df2ae7ab59d500511a90403de12f063d1bacddad82c826c6

Observation 11ad7e9d-a6ed-4f93-acc6-ef0e6cf19d17 · outbound

This paper cites Mirage-1: Augmenting and Updating GUI Agent with Hierarchical Multimodal Skills.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Mirage-1: Augmenting and Updating GUI Agent with Hierarchical Multimodal Skills

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.334615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:751abf3e888ea29021f4be5e5ced029db3ef482de4f3fc744f7dd435a1a4a8ea

Observation 390260fa-235b-4580-8657-fae872950668 · outbound

This paper cites RoboSSM: Scalable In-context Imitation Learning via State-Space Models.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model RoboSSM: Scalable In-context Imitation Learning via State-Space Models

Reference 28

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T16:54:58.298368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:eb67c9a3e0b9f851c73902a7c3b7b0e4fec23ca12613abb0fb49d7afc2e04380

Observation 39d1ad7d-3f52-4139-b1a4-414c8556d667 · outbound

This paper cites 3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model 3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:58.289934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:ee5e7fc51d6aa9ee948b817dbb084249b5b52be9f0bc76781ed9763b2d0501d5

Observation d45834e8-781a-4585-a3e2-d40fe382e7bb · outbound

This paper cites UP-VLA: A Unified Understanding and Prediction Model for Embodied Agent.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model UP-VLA: A Unified Understanding and Prediction Model for Embodied Agent

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.304885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:ffef7b2ba5ccc9699bdd2b5164fb5c2bd063b2dd92da5272154bdb94cb13a25b

Observation 05d4939f-8b83-404d-8912-30cd78ec3e9e · outbound

This paper cites an unresolved cited work.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-07-08T11:54:52.451863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:04c41e089e8d385a7459db395e95d305da5cfd2bfbc83334e1cdc7653e714e8a

Observation 9f810584-b1c2-4893-9e78-57952f622fb8 · outbound

This paper cites an unresolved cited work.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-07-08T11:54:52.455419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:6ccf4d4048540b5bb4353dfef6e77bc154ee1d07bf198d4e288eb9226eead0a8

Observation 57ef0521-44ea-44c7-90a3-3085e62cc4fc · outbound

This paper cites MAP-VLA(Li et al., 2025c) further reduces fragment inconsistency through stage-wise segmentation and alignment.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model MAP-VLA(Li et al., 2025c) further reduces fragment inconsistency through stage-wise segmentation and alignment

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T11:54:52.448421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:63f4861b64e0bfd76d42f23bbbe54b73eb3310ec604097b97fdfbf6fbc9e2c29

Observation 3218a125-65eb-4a2c-9548-4532cf40b97e · outbound

This paper cites Recent VLAs(Black et al., 2025; Black et al.; Shukor et al.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Recent VLAs(Black et al., 2025; Black et al.; Shukor et al

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T11:54:52.442343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:b7a0a2da367125287b5d8f20f7b8f5cd383612c250744fbeff7e249c517877fe

Observation 6fdffb18-654a-4066-a183-d73e451a2893 · outbound

This paper cites distribution shift.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model distribution shift

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T11:54:52.445113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:5e7f145cf44028282619fae03b2e3bf73839dc8249d219e10a83ebe485f94705

Pith citing papers

No inbound Pith citation observations are available.