Pith. sign in

Paper Citation Record · LEDGER

OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2406.09399.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.09399 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:52:28.111511Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T20:46:13.353259Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 23112e8d-a45e-42ed-b0cf-ca7b1847bb1b · inbound

Cosmos World Foundation Model Platform for Physical AI cites this paper.

Cosmos World Foundation Model Platform for Physical AI OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation

Reference 203

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.973996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:5b1a5c47823d8c47afba8c454c0beeaa11d7d281db5e552cd7912a1c6a3b27a6

Observation b6d397e4-2449-4189-a894-d9f5393264fe · inbound

Long-Context Autoregressive Video Modeling with Next-Frame Prediction cites this paper.

Long-Context Autoregressive Video Modeling with Next-Frame Prediction OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:05:17.352046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T23:05:17.201790Z digest=sha256:8b04f37a10ce1da21a12d822f7a3cb4d7c4330c9ad5309e2e12f3448b63582dd

Observation 56516b94-7394-4542-9928-be50ede623da · inbound

MambaVideo for Discrete Video Tokenization with Channel-Split Quantization cites this paper.

MambaVideo for Discrete Video Tokenization with Channel-Split Quantization OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T19:52:28.111511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:52:28.111511Z digest=sha256:4f075b253d5682c65cc21a37ef456ef0351c019bc427f4d1e2a99698b3698de1

Observation cda5a9f7-cd61-4ce3-9d75-581450429c58 · inbound

OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation cites this paper.

OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:31:09.045343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T08:52:07.545191Z digest=sha256:171dd9c0040a0467b4d114e014d836b53bbc348fcfb9eacd656b38e87bb83228

Observation fc68574a-08c1-4ff7-8feb-bf567ef7ff6c · inbound

MergeTok: Unified Continuous and Discrete Visual Tokenization via Token Merging cites this paper.

MergeTok: Unified Continuous and Discrete Visual Tokenization via Token Merging OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:42:46.039598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T22:42:20.280777Z digest=sha256:09686a808072d2ca4d98893c5268f0ea126fc9dda8bd30ad823460a98d1c7299

Observation 1e2435d6-d802-41cb-806a-337761e6e8af · inbound

Wavelet as Tokenizer: Preliminary Results on a Shared Wavelet Token Schema for Natural Signals cites this paper.

Wavelet as Tokenizer: Preliminary Results on a Shared Wavelet Token Schema for Natural Signals OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:46:13.355163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T18:05:11.392504Z digest=sha256:472760fff175512ed665ed8afdf58f700fc60b2b4d4a6f40c4520ec5a1ae3e1b