Pith. sign in

Paper Citation Record · LEDGER

MAGVIT: Masked Generative Video Transformer

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2212.05199.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2212.05199 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:10:14.507734Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b8ac0df3-9642-4073-aa10-fa8153dde81c · inbound

FAST: Efficient Action Tokenization for Vision-Language-Action Models cites this paper.

FAST: Efficient Action Tokenization for Vision-Language-Action Models MAGVIT: Masked Generative Video Transformer

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:52:32.112802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T08:52:31.686474Z digest=sha256:4dd2d90cb8e520b6e81b72d2cb3e765d3c4b4d296e04d6398aa4f7d471b9d21e

Observation 9dda5626-d89f-474a-92d8-7620a5c62f59 · inbound

Infinite Video Understanding cites this paper.

Infinite Video Understanding MAGVIT: Masked Generative Video Transformer

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T18:10:14.507734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:10:14.507734Z digest=sha256:b01131c721c0a7fa4cceec59d797678d1d592b89f85d27a56c124ca9466e5458

Observation 144296a1-b1b6-42ab-a870-9fc2f61ec084 · inbound

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models cites this paper.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models MAGVIT: Masked Generative Video Transformer

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:03.235158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:03.235158Z digest=sha256:65dd155862c4b00c0b760cd52750fa9e5824a8505e5c0b27ec905acd13ffc1d3

Observation 6321bf00-d4cd-4399-b8f1-2cfa6976d55c · inbound

World Action Models: The Next Frontier in Embodied AI cites this paper.

World Action Models: The Next Frontier in Embodied AI MAGVIT: Masked Generative Video Transformer

Reference 288

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:07:18.096556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T05:01:16.802019Z digest=sha256:74e101688702bf8ef067f8488ed78e4c3605fb1a037460c151e3e7ad7fbb1391

Observation 8b4f0df9-7aac-41d6-a8f9-6183e02d0ad2 · inbound

Wavelet as Tokenizer: Preliminary Results on a Shared Wavelet Token Schema for Natural Signals cites this paper.

Wavelet as Tokenizer: Preliminary Results on a Shared Wavelet Token Schema for Natural Signals MAGVIT: Masked Generative Video Transformer

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:46:13.363676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T18:05:11.392504Z digest=sha256:8b47d61140b1d665c21ca0b44e5e7f8f31b5858e5ef5e171e47b4b042df45fa2

Observation a2d94116-368c-42cb-8088-809127e1ff68 · inbound

BioVid: Autoregressive Video Generation with Biological Behavior Semantic Comprehension cites this paper.

BioVid: Autoregressive Video Generation with Biological Behavior Semantic Comprehension MAGVIT: Masked Generative Video Transformer

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-06-27T18:41:08.034995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T18:35:14.331659Z digest=sha256:eab05f09e2a24bf15e2f52d8badd5b28fa2fcc83f8cc2bc671f7e3c5bf146d17