Pith. sign in

Paper Citation Record · LEDGER

OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2406.09399.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.09399 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:52:28.111511Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T20:46:13.353259Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 23112e8d-a45e-42ed-b0cf-ca7b1847bb1b · inbound

Cosmos World Foundation Model Platform for Physical AI cites this paper.

Cosmos World Foundation Model Platform for Physical AI OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation

Reference 203

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.973996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:44d85209efd748cdb107eba633c829475b13f0d4c17ed5275d9b04b70695d25a

Observation b6d397e4-2449-4189-a894-d9f5393264fe · inbound

Long-Context Autoregressive Video Modeling with Next-Frame Prediction cites this paper.

Long-Context Autoregressive Video Modeling with Next-Frame Prediction OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:05:17.352046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T23:05:17.201790Z digest=sha256:6112749fe0ac0c76872ae746a42d75cd401a7145271e77503ae045657d324292

Observation 56516b94-7394-4542-9928-be50ede623da · inbound

MambaVideo for Discrete Video Tokenization with Channel-Split Quantization cites this paper.

MambaVideo for Discrete Video Tokenization with Channel-Split Quantization OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T19:52:28.111511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:52:28.111511Z digest=sha256:fec78b82ef3c520e528099629078bdabfd964024bc03b8769889e43ec99c3cb2

Observation cda5a9f7-cd61-4ce3-9d75-581450429c58 · inbound

OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation cites this paper.

OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:31:09.045343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T08:52:07.545191Z digest=sha256:d420d351c2d4b6d0e81c3074a84adce6b373828567c0ee2f92225a7e3d0c3994

Observation fc68574a-08c1-4ff7-8feb-bf567ef7ff6c · inbound

MergeTok: Unified Continuous and Discrete Visual Tokenization via Token Merging cites this paper.

MergeTok: Unified Continuous and Discrete Visual Tokenization via Token Merging OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:42:46.039598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T22:42:20.280777Z digest=sha256:7879b987f9e9db8d50101ed3526ac11bfe97cf848e41b49e0ca9faa21b7bb72b

Observation 1e2435d6-d802-41cb-806a-337761e6e8af · inbound

Wavelet as Tokenizer: Preliminary Results on a Shared Wavelet Token Schema for Natural Signals cites this paper.

Wavelet as Tokenizer: Preliminary Results on a Shared Wavelet Token Schema for Natural Signals OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:46:13.355163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T18:05:11.392504Z digest=sha256:cedd18997f827aba057487380e8b0246ee73225a7730e82fc5dec68e63f4c9aa