Pith. sign in

Paper Citation Record · LEDGER

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer

As of 9 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 1 inbound Pith citation observation for arXiv:2507.04947.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.04947 v1

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:41:23.831463Z

measured 76 of 76 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T04:23:08.137866Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T04:23:09.921374Z

Reference resolution

75 of 75 outbound references displayed

  • verified exact0
  • verified fuzzy30
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 34dd3310-39fe-4f53-a573-f35974eb1cd1 · outbound

This paper cites FlexTok: Resampling Images into 1D Token Sequences of Flexible Length.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer FlexTok: Resampling Images into 1D Token Sequences of Flexible Length

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.597184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.597184Z digest=sha256:c56c4868ecdabfc60367388417c5edfcf0f0194a1daf80b9ef4121e3c6feb33a

Observation 12388c83-15cb-42a7-aad7-f388b1f89336 · outbound

This paper cites Meissonic: Revitalizing Masked Generative Transformers for Efficient High-Resolution Text-to-Image Synthesis.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Meissonic: Revitalizing Masked Generative Transformers for Efficient High-Resolution Text-to-Image Synthesis

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.601433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.601433Z digest=sha256:2e2847c9932c7760302d3125853d9ae1b0e08d3f045a286d94c4b5763f6c7d5f

Observation b34dd59c-ddf7-498e-8ecd-29082de15795 · outbound

This paper cites All are worth words: A vit backbone for diffusion models.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer All are worth words: A vit backbone for diffusion models

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.602385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.605018Z digest=sha256:1c3a6eae7e2dd2f56866b9342834204d9ad1065a2eed4b903249e1354b238f3b

Observation 1b725b72-339a-46c0-8f26-3a378878204c · outbound

This paper cites Flux, 2024.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Flux, 2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.593338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.608086Z digest=sha256:0903fbfed7ff01e7f49ac2bc3a113f8d2c30fe57673315ed1e31e5208d621b87

Observation 4cec7b2a-686c-4369-a2fd-a9e758073b30 · outbound

This paper cites Efficientvit: Lightweight multi-scale attention for high- resolution dense prediction.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Efficientvit: Lightweight multi-scale attention for high- resolution dense prediction

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.583437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.611773Z digest=sha256:e61bb5c17015f534cc3082e2357c61bd1f255d32bff41948781dbac81f03f6c6

Observation 745ad257-39eb-4eba-aade-9af00a02f3e0 · outbound

This paper cites Condition-aware neural network for controlled image generation.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Condition-aware neural network for controlled image generation

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.572142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.614980Z digest=sha256:ba39ff4304172c3f44253e5e7daf00979b9e8aac573178fecbe33855f40efc94

Observation 047dad19-1554-4487-9528-3b8954e7e9d6 · outbound

This paper cites Maskgit: Masked generative image transformer.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Maskgit: Masked generative image transformer

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.560754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.617922Z digest=sha256:07569350a0ac5962b7468b6d8c153695d9a41fbfd2507951632b6f8290173c32

Observation 0757d32b-5882-4aac-a7ab-e83195aef7f6 · outbound

This paper cites Muse: Text-To-Image Generation via Masked Generative Transformers.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Muse: Text-To-Image Generation via Masked Generative Transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.621281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.621281Z digest=sha256:9fb6ab3dffe1f97a6e9f9f123ef7a113d5718a5c4a73c090cf244e5a3415e7b4

Observation 62f40565-28e3-4d95-bbfe-fa5f79d53533 · outbound

This paper cites SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.624757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.624757Z digest=sha256:c62eb97b3974736eb5609c0375646e8e17aa4cc6017ea67576a599e259d504a2

Observation 137e2f05-51d3-4712-84a0-7b96c1511d17 · outbound

This paper cites Masked Autoencoders Are Effective Tokenizers for Diffusion Models.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Masked Autoencoders Are Effective Tokenizers for Diffusion Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.628002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.628002Z digest=sha256:b72c427e481fcd567194d272f3ae160409f8c97e3d52421b13c708a592fc500a

Observation a32b09d4-4e31-4bb0-b5a3-beb141350c54 · outbound

This paper cites Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.631407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.631407Z digest=sha256:896a3d66b72b3fa46f7fef3fdc8b73982e5abdd0469eea66fc751b65bee28b3c

Observation 4ba7edef-3ab1-4b73-a8ab-98746842fe2c · outbound

This paper cites Pixart- σ: Weak-to-strong training of diffusion transformer for 4k text-to-image generation.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Pixart- σ: Weak-to-strong training of diffusion transformer for 4k text-to-image generation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.550348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.634884Z digest=sha256:03f2ce01f38df8529c62ae252b4ea770f71ed48e130b282b78d9d0b5eba27177

Observation 1ccf0a15-739d-47c5-9915-32ae42beb7c9 · outbound

This paper cites PIXART-{\delta}: Fast and Controllable Image Generation with Latent Consistency Models.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer PIXART-{\delta}: Fast and Controllable Image Generation with Latent Consistency Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.637843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.637843Z digest=sha256:f17901d83e4aad59f8a5d14a6d7b4e3b79b02de666c020189b0318e2a627c815

Observation 7f6b40b9-ce03-41c3-9da7-c0df6c9da372 · outbound

This paper cites Pixart- α: Fast training of diffusion transformer for photorealistic text-to-image synthesis.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Pixart- α: Fast training of diffusion transformer for photorealistic text-to-image synthesis

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.539299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.641117Z digest=sha256:a8e5e6f24ddd6a7ab46419687f53b9e9613959772f086c9e7e9e034b2e159b65

Observation e8a436c8-2254-4d9e-92f2-2f1249cd3316 · outbound

This paper cites MaskMamba: A Hybrid Mamba-Transformer Model for Masked Image Generation.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer MaskMamba: A Hybrid Mamba-Transformer Model for Masked Image Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.643833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.643833Z digest=sha256:4e37f910bfd4bc123a7bbb17376e68191faf335579817752a1da9cd1155f8497

Observation db8be56c-dd30-42d4-8ed1-30208cf160cb · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.646952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.646952Z digest=sha256:76c89c2dd3f935e0c0ab1e2ad2e53b8c79018ced65e25c79d8dc1840b9486441

Observation f3f8b3ac-f320-46cd-8758-8e241995f1bb · outbound

This paper cites Collaborative Decoding Makes Visual Auto-Regressive Modeling Efficient.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Collaborative Decoding Makes Visual Auto-Regressive Modeling Efficient

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.650563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.650563Z digest=sha256:36005e25a95106a9424f420f85eddcce0c791a468910c9e74b907acd67d129e5

Observation 1905c30c-1215-4bfb-b4ec-04f85879daa4 · outbound

This paper cites Vqgan-clip: Open domain image generation and editing with natural language guidance.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Vqgan-clip: Open domain image generation and editing with natural language guidance

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.528012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.653858Z digest=sha256:6ea51c1a9038bac437b3122b0cc8dba43c25ce00ef80f9fdbff0b3b439ab35b0

Observation 2e503436-5d51-4628-b25e-9fff171b8223 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Imagenet: A large-scale hierarchical image database

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.517000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.656533Z digest=sha256:fecd3db65ecb4ad200e2510e6966a630e3166536d2c8ce027d992c0ec0710c1a

Observation 3dc5b9e6-e0b5-4149-b72d-e91cbb5a07dd · outbound

This paper cites Bert: Pre-training of deep bidirectional trans- formers for language understanding.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Bert: Pre-training of deep bidirectional trans- formers for language understanding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.659371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.659371Z digest=sha256:1c77a015e024d7fe9fda3b462bb884ada1525c4bd7edc48aebf54faad737af5f

Observation b26dfbd7-02f7-408d-a6c8-daa19cfa3668 · outbound

This paper cites Cogview: Mastering text-to-image generation via transformers.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Cogview: Mastering text-to-image generation via transformers

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.499523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.662520Z digest=sha256:42fd215f14e78797a147e9e2a57ccbb907cb66d7d9900887e6fdfdbc4ba7bf46

Observation ed0c676b-6098-46ad-aa37-1aa6ec6573e9 · outbound

This paper cites Cogview2: Faster and better text-to-image generation via hierarchical transformers.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Cogview2: Faster and better text-to-image generation via hierarchical transformers

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.487740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.665292Z digest=sha256:367f65047be0ffa6832003f84eb38145e700e33b40488f33efc9e16964496106

Observation bf37ec64-e006-42f6-a2c2-d3cb9b640474 · outbound

This paper cites Taming transformers for high-resolution image synthesis.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Taming transformers for high-resolution image synthesis

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.668024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.668024Z digest=sha256:571e6116f5b3f74def023b559a7f7778950b4245e51b68dcba60e2e380cff68c

Observation bd3707a4-1393-4e79-a4a8-870fe9598436 · outbound

This paper cites Scaling recti- fied flow transformers for high-resolution image synthesis.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Scaling recti- fied flow transformers for high-resolution image synthesis

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.670970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.670970Z digest=sha256:8923a6d91a4d5093cc73f56531aa0dfcdd12185407eedb22a90a11fd64bc5fb1

Observation ff003569-80a0-4a66-a324-9c1a05f1c1a9 · outbound

This paper cites Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.673806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.673806Z digest=sha256:a486fe3045dd95f953afb1778452bae0d8ccde8e6806b6744aac8da5a8900184

Observation 5cbe8f3e-1ded-4354-af83-b2672b2a4d1d · outbound

This paper cites Make-a-scene: Scene- based text-to-image generation with human priors.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Make-a-scene: Scene- based text-to-image generation with human priors

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.453151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.677307Z digest=sha256:99e84787b509fe4feadb00816e01cb545ff60060c71d3bc02fd760a91dc01165

Observation 02221b71-5d38-4501-90ea-30a1900a85e3 · outbound

This paper cites Geneval: An object-focused framework for evaluating text- to-image alignment.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Geneval: An object-focused framework for evaluating text- to-image alignment

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.442337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.680293Z digest=sha256:384f6182305bb6936d8f4742ec7bca33b3e3e3c821fabef2c7bea772af36db0c

Observation 4ad5d9f1-7fac-4d07-8739-4632c1a45992 · outbound

This paper cites Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.683029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.683029Z digest=sha256:dc1596b86ba377958fe07ac2272e8d3881924793746f6ce8d7d76fe61af69b53

Observation 188553c7-4b49-4eb3-a5f5-8047a6497060 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilib- rium.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Gans trained by a two time-scale update rule converge to a local nash equilib- rium

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.685861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.685861Z digest=sha256:698446eb018adb6b26749ff8979d897314249d79a2b2f154f3b742cd1d11d22a

Observation 7737570e-471b-4ad9-8965-39b0e94ec915 · outbound

This paper cites LANTERN: Accelerating Visual Autoregressive Models with Relaxed Speculative Decoding.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer LANTERN: Accelerating Visual Autoregressive Models with Relaxed Speculative Decoding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.688829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.688829Z digest=sha256:13a2310f502c3127aabe0e542a2df17c0318fade99947ada529b02fff8f401fa

Observation 60d468d8-6607-4641-8640-b7445fdc2ca0 · outbound

This paper cites Democratizing Text-to-Image Masked Generative Models with Compact Text-Aware One-Dimensional Tokens.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Democratizing Text-to-Image Masked Generative Models with Compact Text-Aware One-Dimensional Tokens

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.691857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.691857Z digest=sha256:32083d797cb21d7008f1b0064cef677ba484bee6f02ec431962114a4bf2c495c

Observation 14c41753-a012-4487-a846-551f7d8cc919 · outbound

This paper cites Videopoet: A large language model for zero-shot video gen- eration.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Videopoet: A large language model for zero-shot video gen- eration

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.425715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.694896Z digest=sha256:664daf3e2809cfa69c48ac3624f0b0018b9828eb5d4dc58bfae1179c83798eb6

Observation ad64d74f-3a96-4457-af1b-4ad3542c923a · outbound

This paper cites Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.697555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.697555Z digest=sha256:0a341c51e27f8114450c412fa2695906853e1c64f051252ab57c10cf9d0f7160

Observation d018cac0-b23c-4277-9dc4-1541957b1376 · outbound

This paper cites Mage: Masked generative encoder to unify representation learning and image synthe- sis.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Mage: Masked generative encoder to unify representation learning and image synthe- sis

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.700734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.700734Z digest=sha256:05c94a43dc3a793e3cc43da35994dc8d78d1081967c35c3620b23f3b75676572

Observation 6b558d4a-4b2c-4a46-acf1-1aa8dd686bf9 · outbound

This paper cites Autoregressive image generation without vec- tor quantization.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Autoregressive image generation without vec- tor quantization

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.406704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.704341Z digest=sha256:9fe6482098793291f0c558ee24681b43b8e4777a0e604c78014285d9b83bc023

Observation da46fd01-6f63-4522-9731-2f592007beca · outbound

This paper cites ControlVAR: Exploring Controllable Visual Autoregressive Modeling.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer ControlVAR: Exploring Controllable Visual Autoregressive Modeling

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.707412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.707412Z digest=sha256:c78eedf0e56d718b66e49af4919aac62dd32984ffb201841b7fda0bac34af532

Observation 61176306-b6d5-4f11-adea-42bbed8557de · outbound

This paper cites Vila: On pre-training for vi- sual language models.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Vila: On pre-training for vi- sual language models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.395848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.710603Z digest=sha256:92a00a94c73a1d932f5fcd3e7ef9de7ce70d56bd7027cccd411aeb3aadf7763e

Observation 7a3d0dcb-6129-494c-8159-78cd2e95eefa · outbound

This paper cites Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.713635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.713635Z digest=sha256:3e0cea5ede077c5d5eb45b136e9cda3810ebace6981c6108c9d69ef284e6f564

Observation d2bd89df-c72d-4e5d-8a5c-2f7b2e9f196a · outbound

This paper cites Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.716761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.716761Z digest=sha256:e3f93bd6c1c9467a9e3c639f64433f3d198b42fdeb967616c32b540aac410f75

Observation 10e17c38-ac4e-4510-a198-8794726844f2 · outbound

This paper cites World Model on Million-Length Video And Language With Blockwise RingAttention.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer World Model on Million-Length Video And Language With Blockwise RingAttention

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.720034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.720034Z digest=sha256:ccc2fa623f0320aa7d0b793bc0a93848d6fdbc22e6a20320474a1873ec6c26c7

Observation 05608fb7-ecfb-41a0-b473-66459712bb50 · outbound

This paper cites Exploring the role of large language models in prompt encoding for diffusion models.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Exploring the role of large language models in prompt encoding for diffusion models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.385166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.723272Z digest=sha256:c7019750580c15396045891ea7f437f461c65478ac993d4a9a0d50a9031ef113

Observation 217bf85a-1191-4a80-8f3b-3e43d2497ea0 · outbound

This paper cites STAR: Scale-wise Text-conditioned AutoRegressive image generation.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer STAR: Scale-wise Text-conditioned AutoRegressive image generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.726623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.726623Z digest=sha256:d63b40557e57c3e274073ca6741d4b9e281f9ac9cc7f4c77a2ebf95683337f8c

Observation 3a459c38-41de-4b25-bb4f-1463c15ff854 · outbound

This paper cites Hello gpt-4o, 2024.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Hello gpt-4o, 2024

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.373082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.729953Z digest=sha256:09da28cf3e63019e515d22eb0092cf578599a69c18ec749c4a8a41be7303b5ac

Observation e70fe837-c7d5-4f4b-b432-cad4b1270b93 · outbound

This paper cites Scalable diffusion models with transformers.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Scalable diffusion models with transformers

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.732818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.732818Z digest=sha256:1f68dd21196beffdf16bb9ed570e0a749f0152c1a07d482cd63854698ce4759c

Observation d32be879-4738-4d53-8d2a-1f2e9962b981 · outbound

This paper cites W ¨urstchen: An ef- ficient architecture for large-scale text-to-image diffusion models.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer W ¨urstchen: An ef- ficient architecture for large-scale text-to-image diffusion models

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.356049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.736414Z digest=sha256:90f81b0c47f674c5d8d3db0a3f64aa83ef592434d02ace2ce262bd37ca857e57

Observation 387bf42c-f125-42aa-942a-126750c839f8 · outbound

This paper cites Sdxl: Improving latent diffusion models for high-resolution image synthesis.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Sdxl: Improving latent diffusion models for high-resolution image synthesis

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.345544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.740004Z digest=sha256:62d1429ccf38128f9f028c2bea3dffd2998abc59c3d4dfb204b70ee0470897e3

Observation e77a85cc-3c45-4774-b6cd-e6648a3feeb8 · outbound

This paper cites TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.743296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.743296Z digest=sha256:9d19d42ec1a5c7968fbaaa731ab08c6deaa9737ab0bc08498d66f5decd817697

Observation 96036b35-5845-4c61-89e6-8ba9814ae24e · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Exploring the limits of transfer learning with a unified text-to-text transformer

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.334820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.746402Z digest=sha256:0fbf50a44b610db41e49e84c90253490989c6b9cf5a64bb1a3cb37cca28c2a4f

Observation 514fabfe-dc50-4fb9-8220-db01381d887b · outbound

This paper cites Zero-shot text-to-image generation.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Zero-shot text-to-image generation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.749368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.749368Z digest=sha256:fe1aea7d994a8f0613654ad76281e6d5956b429d670fda16cf7053bef7b6bae3

Observation b59c94d5-6f14-4b37-a016-3c49affc816d · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer High-resolution image synthesis with latent diffusion models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.752473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.752473Z digest=sha256:7a6442d8a3ea49cd29e1ed3e0a690949a3dbd8c9f096b7a933f15df76485ebd4

Observation 98a0a5e4-7bc7-4417-86af-e1c317274499 · outbound

This paper cites Journeydb: A benchmark for generative im- age understanding.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Journeydb: A benchmark for generative im- age understanding

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.310911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.755319Z digest=sha256:df3bc9b57486efcba30f271131c88470c1805c30adccf305add89cf1c55595e1

Observation b3db0879-fb1b-466e-8256-142c6ec613f1 · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.758914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.758914Z digest=sha256:f3cd15e7d192ea6b16394c9758d1ff87f0331bed77214f6bd7b06a8112678e46

Observation 32ffb138-09c0-44ba-ae2e-ee0063776e3d · outbound

This paper cites HART: Efficient Visual Generation with Hybrid Autoregressive Transformer.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer HART: Efficient Visual Generation with Hybrid Autoregressive Transformer

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.762860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.762860Z digest=sha256:597b272d18bfbb21cbffe1447718b2bde851f910dbc24fc44e3e974b131d6e74

Observation ef40a9a9-7bfc-4dea-91d1-b34525e77634 · outbound

This paper cites Introducing auraflow v0.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Introducing auraflow v0

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.299918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.766316Z digest=sha256:431983fc9ef85a0b2a5f64a524ae0ec9c0067d146dc0f07457208da7fc4ff0e0

Observation ea4877f2-701f-4f21-a926-cce601e32e35 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Gemini: A Family of Highly Capable Multimodal Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.769261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.769261Z digest=sha256:37587f67727130d420b33759113d279d765172742c9b560498d95df5bbfa4629

Observation 05c6b64d-7efb-40a4-b309-aa4736357ae1 · outbound

This paper cites Kolors: Effective training of diffusion model for photorealistic text-to-image synthesis.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Kolors: Effective training of diffusion model for photorealistic text-to-image synthesis

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.289469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.772210Z digest=sha256:4c597579430a4139560844d6e90ebe3cb404cfaa77427498dc89b9f7b1ad8e98

Observation 47186ff9-ba17-4818-a852-85093f69afeb · outbound

This paper cites Visual autoregressive modeling: Scalable image generation via next-scale prediction.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Visual autoregressive modeling: Scalable image generation via next-scale prediction

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.277788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.775229Z digest=sha256:fea78f9daf4936bd6e63e835bbef6f3304a1b0e0cfd4342133eaa15b1df644a9

Observation 24ca1245-6f49-4bc6-8998-30b18845bdf7 · outbound

This paper cites Neural discrete representation learning.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Neural discrete representation learning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.778537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.778537Z digest=sha256:d1f82aa13eea6bc7e09219f0d5d5dc3ba1f0653f48124c0cefeccd8b6c6544f8

Observation e18e3c3e-01c1-45dd-a062-775cf528945c · outbound

This paper cites Phenaki: Variable Length Video Generation From Open Domain Textual Description.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Phenaki: Variable Length Video Generation From Open Domain Textual Description

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.781644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.781644Z digest=sha256:cc35097027909c89f640a8c0457c04663fbfa58c3f3a2614857c92cba459a963

Observation 9b48ac4f-56d7-4001-9c42-9910aba55a29 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Emu3: Next-Token Prediction is All You Need

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.784989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.784989Z digest=sha256:b2fea2129daea3a05896c04716770a3b1b230ba0cdc853a505cbe8d6078b7253

Observation aeb2ad6d-4043-4c25-9816-35067e89fed0 · outbound

This paper cites Parallelized Autoregressive Visual Generation.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Parallelized Autoregressive Visual Generation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.788082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.788082Z digest=sha256:f889a79ec7857a2facf7c864449b778be1f2ff039afcac44416a40dab7d5ea23

Observation 10e389e7-0f61-43f6-b4b9-303926d15be5 · outbound

This paper cites Maskbit: Embedding-free image generation via bit tokens.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Maskbit: Embedding-free image generation via bit tokens

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.258779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.791122Z digest=sha256:2c7e1a2ac142e27a0cab586320eeca3aae8961648ef0bf468de1dc6f53618e74

Observation b6830a1d-9454-4d9e-9dc4-fdfa04d5e7f3 · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.793876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.793876Z digest=sha256:6bef4157c23ccf97b3354b2af73ccf3fd770296af3c07ea78d9f350273f959b2

Observation 6d49930b-816d-492d-a59a-bb82a5337018 · outbound

This paper cites VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.796917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.796917Z digest=sha256:f83070c202b231b330fb997dc8380aa06a70b624037ecb89554ed28a728ded32

Observation 3a94fa02-2984-48fc-8485-0a6fe88377f3 · outbound

This paper cites SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.799913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.799913Z digest=sha256:b5c3ea89f005450770faf3f5841e3f8ee1eb133aa6b91dcc115b2b20685a124c

Observation 17170477-082e-4982-9a7a-c920950cf2bd · outbound

This paper cites SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.803601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.803601Z digest=sha256:67395f66c7ead0e9d6e8dcea53b017e8d5161d76813c5121b183e05aa6ea7105

Observation 7d9a9c8d-afb2-4b25-9f0c-0017bc5bd809 · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.806841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.806841Z digest=sha256:c3f67101edbc7066deca5975285d91e9291c7861e81e64d551503a38eda49555

Observation 7220a54c-9e3d-467b-bfcd-98c904db1064 · outbound

This paper cites CAR: Controllable Autoregressive Modeling for Visual Generation.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer CAR: Controllable Autoregressive Modeling for Visual Generation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.810106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.810106Z digest=sha256:a47f1953c2ba51b63912f75e996bc4d6a3c6f1ea7c58051b942069deef9437d0

Observation 9d469825-40e3-496d-81ae-64fe1f51a583 · outbound

This paper cites Scaling autoregres- sive models for content-rich text-to-image generation.Trans.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Scaling autoregres- sive models for content-rich text-to-image generation.Trans

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.247703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.813094Z digest=sha256:ddb8a1eb697a264f1d2e287cf41b4b90f0e0f0971bf1c610121bd55c5f249cb9

Observation a4399186-bebf-4829-8b73-eacedc0fa28e · outbound

This paper cites Magvit: Masked generative video transformer.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Magvit: Masked generative video transformer

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.237240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.816090Z digest=sha256:6560ed35c45c73d5490aebd1ac23e0116bc5b4946e5083bdad3e661fa0d15774

Observation 852f6685-331f-4262-8912-dd1508fe325f · outbound

This paper cites An image is worth 32 tokens for reconstruction and generation.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer An image is worth 32 tokens for reconstruction and generation

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.226993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.818931Z digest=sha256:b8ad839317b6b2b80a287185bc08871d5beddf42fc29f630676c70c30ddfb3be

Observation 0cbbcf26-880c-45f5-b390-3b08383a86aa · outbound

This paper cites ShieldGemma: Generative AI Content Moderation Based on Gemma.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer ShieldGemma: Generative AI Content Moderation Based on Gemma

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.821729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.821729Z digest=sha256:5c5d94ec6b4c5e0e8dd715218ec0c3ca01aa1b11c47f4a30ac8647c8ac51ecbc

Observation 04a282f2-e224-4651-b17e-bd2a670c7af5 · outbound

This paper cites Language-Guided Image Tokenization for Generation.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Language-Guided Image Tokenization for Generation

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.825012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.825012Z digest=sha256:3e8d3b3873cbacd5ab26e0c606b4a5d6978d1c97b466c2d122f979cb95075ef5

Observation 43ad1ff8-72f0-4a82-99e6-2f3e1f952c7e · outbound

This paper cites VAR-CLIP: Text-to-Image Generator with Visual Auto-Regressive Modeling.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer VAR-CLIP: Text-to-Image Generator with Visual Auto-Regressive Modeling

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.828655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.828655Z digest=sha256:d619442ec98431afc35bf2b20b4db9c6e1869a307594cec270125f82e092e9ca

Observation 6b79ced7-fa55-4df9-a90c-e187a0b47102 · outbound

This paper cites A red heart.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer A red heart

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:41:24.215300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:41:23.831463Z digest=sha256:e35adf5572183ff611f6e762f9c651b7e68690d8334061c74349706dfd4eb0cc

Pith citing papers

Observation 70406000-a5b0-4912-b562-5cf97973bb31 · inbound

HPSv3: Towards Wide-Spectrum Human Preference Score cites this paper.

HPSv3: Towards Wide-Spectrum Human Preference Score DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-08-06T04:23:09.993343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T04:23:08.137866Z digest=sha256:999addb840cffb94447c0abebd29ed6b27722aefafdf659ff6e57f17df309bb7