Pith. sign in

Paper Citation Record · LEDGER

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference

As of 5 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 2 inbound Pith citation observations for arXiv:2605.02739.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.02739 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-08T17:49:40.820795Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T03:23:47.033536Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

13 of 13 outbound references displayed

  • verified exact9
  • verified fuzzy1
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0a8761d6-1317-4627-8e54-b283aba21db5 · outbound

This paper cites Revisiting Feature Prediction for Learning Visual Representations from Video.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference Revisiting Feature Prediction for Learning Visual Representations from Video

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-12T12:40:24.452580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:9af357f552bb87cf8d85cfdd23a9e62250e1981bb7bb168269f5f2e0c658f0fd

Observation f703e831-dc06-48ad-8577-37a1f8a4d6af · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-11T17:11:14.715264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:d319f84e9362d568c6d170f99da01f1d8a85b53d69914cf0d77f1c784a9ce067

Observation ae339173-9ef0-4f2e-acc5-2befc3beb167 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-11T17:11:14.464836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:afe1f51d7ab19b816676c16370a2885d828cb7382055d44394f1ecefe6eb0659

Observation b30eee6c-ec9e-47cb-8a6c-b6f96814cf84 · outbound

This paper cites SQAP-VLA: A Synergistic Quantization-Aware Pruning Framework for High-Performance Vision-Language-Action Models.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference SQAP-VLA: A Synergistic Quantization-Aware Pruning Framework for High-Performance Vision-Language-Action Models

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:11:14.246348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:436e1f84b613dfe0733dde102212b34241dce855f50e4af269618c723ec51f20

Observation c8e3087e-b066-408f-8efe-2d371bedc4ab · outbound

This paper cites arXiv preprint arXiv:2511.18950 (2025).

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference arXiv preprint arXiv:2511.18950 (2025)

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:11:14.827432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:c794a49ca4e7890da0418a234f353985133122b0d0b553f0e848535210483856

Observation ae6ce2d8-bf45-4619-9bc9-e132bbeba89c · outbound

This paper cites CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:11:14.495324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:b523f4607bb2281ca9fccc2562cc41dafd24009930ed3a2a51bde39dc453ec6c

Observation 07e8a87a-679e-4656-b5d9-029ff9a003a0 · outbound

This paper cites Bridging the Semantic-Action Gap in Visual Token Pruning for Efficient VLA Inference.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference Bridging the Semantic-Action Gap in Visual Token Pruning for Efficient VLA Inference

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-26T03:04:58.134498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:23a18e90547c540610dac4d77f121102dd6f845126188662e6b85f69d4ea4e54

Observation 75ec6636-16a9-48fb-ab9e-8e62662b3964 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-11T17:11:14.924749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:f7fbbd56883ad821cf2b4b94c105e2d0c161bada31df2842302c8828e8b86d98

Observation 920aff8b-a957-4f57-9785-0153d80c0084 · outbound

This paper cites SpecPrune-VLA: Accelerating Vision-Language-Action Models via Action-Aware Self-Speculative Pruning.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference SpecPrune-VLA: Accelerating Vision-Language-Action Models via Action-Aware Self-Speculative Pruning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-26T02:03:02.651395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:8b8884b627984515e328782d42487f9d9b69a88ca173600d99e08aa49d8b4c8c

Observation 914138b2-21be-4b2a-9d18-0ab2604a01a4 · outbound

This paper cites Scaling Proprioceptive-Visual Learning with Heterogeneous Pre-trained Transformers.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference Scaling Proprioceptive-Visual Learning with Heterogeneous Pre-trained Transformers

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:11:14.816154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:59a84f667c9e398dc3504ccaa7d66c784479fca0c2e5a11d8d266326b5db9955

Observation 63df7d0a-c5e7-4a05-9075-32a9c2218a05 · outbound

This paper cites arXiv preprint arXiv:2502.02175 (2025).

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference arXiv preprint arXiv:2502.02175 (2025)

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:11:15.083137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:0c58af92f1bb4feede9d5b84eeebc1854106c990f8182fd1f1c8d98f5327af88

Observation 557a335d-fd68-4611-9203-0e46311f8909 · outbound

This paper cites Dyq-vla: Temporal-dynamic-aware quanti- zation for embodied vision-language-action models.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference Dyq-vla: Temporal-dynamic-aware quanti- zation for embodied vision-language-action models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:11:14.470207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:27b5e7bbb88958860ceeab9058a44b57055ee7934807edc8b42b3a56ddc2ca7d

Observation 9b803458-10fd-4dab-b82f-00e09e2cba34 · outbound

This paper cites max-autotune.

Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference max-autotune

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T06:36:52.134951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T17:49:40.820795Z digest=sha256:d870bc770b974fd9ca332df02c78750870bca4f41d9abb449a81ef19d363037b

Pith citing papers

Observation 2e45585a-2c9f-4e9f-807e-b52c0d84d0be · inbound

Enfold: Folding World Model Imagination into Predictive Representations for Ultra-Efficient Embodied Control cites this paper.

Enfold: Folding World Model Imagination into Predictive Representations for Ultra-Efficient Embodied Control Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T11:27:05.757435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:27:05.757435Z digest=sha256:ddaadeb4fd4d9008b063f6ff2188ec8b6fc089732fdc23699fe821a9ed88993b

Observation 4730151a-369f-4c89-898a-02df7368b868 · inbound

Enfold: Folding World Model Imagination into Predictive Representations for Ultra-Efficient Embodied Control cites this paper.

Enfold: Folding World Model Imagination into Predictive Representations for Ultra-Efficient Embodied Control Latent Bridge: Feature Delta Prediction for Efficient Dual-System Vision-Language-Action Model Inference

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T03:23:47.033536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T03:23:47.033536Z digest=sha256:8ef52e50dacd05a870679978afe3d6a67c8cbefc4fd2370c76153e304a1659da