Pith. sign in

Paper Citation Record · LEDGER

VideoFlexTok: Flexible-Length Coarse-to-Fine Video Tokenization

As of 5 August 2026, this Paper Citation Record lists 2 of 2 outbound references and 1 inbound Pith citation observation for arXiv:2604.12887.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.12887 v1

Coverage vector

measured 2 of 2 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T15:45:02.848369Z

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T17:02:00.653422Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-03T00:47:30.426505Z

Reference resolution

2 of 2 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2f928ea8-fd36-4197-9698-c4ba073f6a51 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

VideoFlexTok: Flexible-Length Coarse-to-Fine Video Tokenization DINOv2: Learning Robust Visual Features without Supervision

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T09:56:01.919530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T15:45:02.848369Z digest=sha256:de7afed8b5d140e7e6fb160cea20b6e7910cd06479d8a3166e2f94cdee3206be

Observation c3b88d40-c09c-4e34-8a75-229a6a0aa527 · outbound

This paper cites an unresolved cited work.

VideoFlexTok: Flexible-Length Coarse-to-Fine Video Tokenization Unresolved cited work

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:56:01.931585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T15:45:02.848369Z digest=sha256:caaa7ef369fab76a68a20b623c89eb93654150cee99246f3918b58eaa59b1382

Pith citing papers

Observation 074c408d-0a55-4863-96bd-d87011695396 · inbound

MilliVid: Hierarchical Latents for Long-Range Consistency in Video Generation cites this paper.

MilliVid: Hierarchical Latents for Long-Range Consistency in Video Generation VideoFlexTok: Flexible-Length Coarse-to-Fine Video Tokenization

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T00:47:30.428152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T17:02:00.653422Z digest=sha256:3c6f846a6a4430825b176906a1a81e3b03c0ebd8f9d606bb5aa4eba9fa77dcd7