Pith. sign in

Paper Citation Record · LEDGER

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment

As of 6 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2606.11221.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.11221 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-29T14:12:23.837961Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact6
  • verified fuzzy0
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch6

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6ac64050-0746-49eb-a0b6-ec61c4efd1ff · outbound

This paper cites Latent Space Oddity: on the Curvature of Deep Generative Models.

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment Latent Space Oddity: on the Curvature of Deep Generative Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-29T14:13:29.950112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T14:12:23.837961Z digest=sha256:fd8209bf1ceb036afa8638c714fe3bff503a2c29fd2723b2fc03d897b87ff1ed

Observation 260ea739-3e6b-4ade-aa8f-c62f23ab969f · outbound

This paper cites Qwen2.5-VL Technical Report.

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment Qwen2.5-VL Technical Report

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-06-29T14:13:29.959519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T14:12:23.837961Z digest=sha256:0ed3278a6ecb27257979cb76b00ce0919e68cef072891be5cd1b1f19be18b49b

Observation 447d351f-623b-41d1-ae64-35139099e568 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-06-29T14:13:29.952309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T14:12:23.837961Z digest=sha256:cc1a81b3b6dba3f7abcbc689ee766d57e6354d6e75c0b17e6d797deb291f1ba3

Observation 344d3232-85f7-4772-996d-8a356656a5f2 · outbound

This paper cites Geometric Deep Learning: Grids, Groups, Graphs, Geodesics, and Gauges.

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment Geometric Deep Learning: Grids, Groups, Graphs, Geodesics, and Gauges

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T14:13:29.947270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T14:12:23.837961Z digest=sha256:e23ebbd1fdb8f7bd7385b11326580f6aed6c6eeef3a82f49b59232b92a1daa94

Observation 96dd8cd7-305c-4366-bb46-f9632770ecae · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T14:13:29.935104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T14:12:23.837961Z digest=sha256:5ea5516514cb87001c8d3b71d99eea9b729a72125f0f93bdc2e69dc94c1d5726

Observation 86b0072c-fd94-4e0e-aaf9-3e1f5e49dc30 · outbound

This paper cites Evaluating Real-World Robot Manipulation Policies in Simulation.

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment Evaluating Real-World Robot Manipulation Policies in Simulation

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T14:13:29.955878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T14:12:23.837961Z digest=sha256:85afb39ecb7034f9fd8efe566b63d77fb8a2e370fe5fcc009f55410d5ca904cc

Observation 963adb28-36cf-4c0d-894e-b1dc113a3e37 · outbound

This paper cites Cross the Gap: Exposing the Intra-modal Misalignment in CLIP via Modality Inversion.

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment Cross the Gap: Exposing the Intra-modal Misalignment in CLIP via Modality Inversion

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T14:13:29.953305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T14:12:23.837961Z digest=sha256:88be664970cf02d13cf342a2dafd97be57ad6c4d77644a4eec4206162b037268

Observation 5f2d0f6c-7f5e-48bf-a122-9684efe35fad · outbound

This paper cites FAST: Efficient Action Tokenization for Vision-Language-Action Models.

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment FAST: Efficient Action Tokenization for Vision-Language-Action Models

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-06-29T14:13:29.954508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T14:12:23.837961Z digest=sha256:40edcc100ff6e182a6c861afeedae31cada2391444790deb0b104a5ecb95cb23

Observation d81c774f-e79e-4ff3-bd37-e214f4a0a2be · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-06-29T14:13:29.937178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T14:12:23.837961Z digest=sha256:3bd6c025de32476abda0a981cb938ea1bfa5bb6ff162a5ef1d2681557418aad9

Observation 3c409798-ae39-4bb0-a740-c3ed94afa7c3 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment Learning Transferable Visual Models From Natural Language Supervision

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T14:13:29.947112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T14:12:23.837961Z digest=sha256:e898f6bbc9f76d27859b0ceee4264179af6663f653e1472d163fcf9e9f882f76

Observation 4ed50052-a74e-42d5-aabc-fcccb1f4efea · outbound

This paper cites A micro Lie theory for state estimation in robotics.

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment A micro Lie theory for state estimation in robotics

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-29T14:13:29.957292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T14:12:23.837961Z digest=sha256:e2eac4d9e1f36d597964f6daedb9e4b64c9fa39ccb791df8f8692acc26817fe5

Observation e2652186-7d4d-46d0-96f8-3a5f023140a5 · outbound

This paper cites arXiv preprint arXiv:2506.06072 , year=.

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment arXiv preprint arXiv:2506.06072 , year=

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T14:13:29.961899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T14:12:23.837961Z digest=sha256:1d9b010b41bae32029953f81b920fc300c66413e4aed86c4a3454998f85aaaa6

Observation a7df753a-d58c-4745-a669-b860b7064652 · outbound

This paper cites an unresolved cited work.

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-29T14:12:23.837961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T14:12:23.837961Z digest=sha256:d155058f20257d2f6fb3f26bd26b528040e8dfbb343eb88c4cfb9d135a153dd3

Observation d04a7e99-708e-4406-bca0-62091d42b977 · outbound

This paper cites an unresolved cited work.

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-29T14:12:23.837961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T14:12:23.837961Z digest=sha256:f383c789e1859c77ba3546ded3b5680bcf6deb87150216daf5a45f18fb1e090d

Pith citing papers

No inbound Pith citation observations are available.