Pith. sign in

Paper Citation Record · LEDGER

VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2507.05116.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.05116 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T09:08:15.654858Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-09T05:36:01.255549Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e6df99e5-6f91-4b17-aa77-f34e0db2dec4 · inbound

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey cites this paper.

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

Reference 109

Resolution
verified exact
arxiv_id, observed 2026-07-09T02:20:27.970763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T20:28:15.818016Z digest=sha256:820c39bd060eb3b9ed64193a6191f40d90da385d6716210bc56044c297243817

Observation a1651af6-f46b-4280-9f52-a8537f756317 · inbound

Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey cites this paper.

Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:15.654858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:15.654858Z digest=sha256:a18f006f261b95cbbdbc436d47a551ac38bae95ad79d6468f0cfcb874c1006aa

Observation 6d7b1da5-4b15-4a25-9913-1dc72f49ae4b · inbound

Human Cognition in Machines: A Unified Perspective of World Models cites this paper.

Human Cognition in Machines: A Unified Perspective of World Models VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

Reference 108

Resolution
metadata mismatch
arxiv_id, observed 2026-07-09T02:20:27.970763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T08:12:15.663761Z digest=sha256:b9c050c56b641fb62085e4e8eadec8bee7cf05152a06f3c27095e761b98aad2c

Observation 6a7c1953-01be-49dc-be94-3e098ffe693e · inbound

AnchorRefine: Synergy-Manipulation Based on Trajectory Anchor and Residual Refinement for Vision-Language-Action Models cites this paper.

AnchorRefine: Synergy-Manipulation Based on Trajectory Anchor and Residual Refinement for Vision-Language-Action Models VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-09T02:20:27.970763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T05:11:52.984680Z digest=sha256:37d28a6669423614d0caf0ee10243160cf713955c083367a6d838bd7093a45db

Observation 8e788380-0f35-4f94-8021-b386b0d88566 · inbound

AnchorRefine: Synergy-Manipulation Based on Trajectory Anchor and Residual Refinement for Vision-Language-Action Models cites this paper.

AnchorRefine: Synergy-Manipulation Based on Trajectory Anchor and Residual Refinement for Vision-Language-Action Models VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T15:56:08.326817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:56:08.326817Z digest=sha256:a941adcd41d1f39085785fb06c1f949990964a8a5f8c477c0fd79a0a7a9f499d

Observation d1936401-242f-4e89-94b2-bf03d872615a · inbound

PhyWorld: Physics-Faithful World Model for Video Generation cites this paper.

PhyWorld: Physics-Faithful World Model for Video Generation VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-09T02:20:27.970763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:a84b7c4edb5ccbda1c9811d240ec586f2f10362b14606735dc7ef0016f1fc7d5

Observation 67461939-4baf-44ee-9823-4ca56e207122 · inbound

QuoVLA: Quotient Space for Vision-Language-Action Models cites this paper.

QuoVLA: Quotient Space for Vision-Language-Action Models VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-09T02:20:27.970763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T12:09:12.124995Z digest=sha256:06c9adad37491f0784836372aafa6daeb73a21a1bc62d24dc0fdacdc6d4c17b7

Observation 94f44ce0-77ed-42e4-926c-ddb43c4e056c · inbound

Flash-WAM: Modality-Aware Distillation for World Action Models cites this paper.

Flash-WAM: Modality-Aware Distillation for World Action Models VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-09T02:20:27.970763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T07:00:46.566214Z digest=sha256:20f1c1c96b6ce2ad0ba19a86c4f435bb453f657e9196fce1e1db3470348c38b0

Observation 0816f57c-57e0-4f93-ad67-b58a0d26215a · inbound

Dual Latent Memory in Vision-Language-Action Models for Robotic Manipulation cites this paper.

Dual Latent Memory in Vision-Language-Action Models for Robotic Manipulation VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-07-09T05:36:01.257153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-09T05:29:36.753268Z digest=sha256:07fc07c79dbc87fbaf0221b6ce7182daf55cd0e15f2b02a523edfff2b597f077