Pith. sign in

Paper Citation Record · LEDGER

Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 58 inbound Pith citation observations for arXiv:2408.02657.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.02657 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 58 of 58 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:17:25.499383Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:49:30.628376Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 602fed6f-7d05-4df8-87ff-2c1d7997c77b · inbound

Bag of Design Choices for Inference of High-Resolution Masked Generative Transformer cites this paper.

Bag of Design Choices for Inference of High-Resolution Masked Generative Transformer Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T19:23:16.608097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:23:16.608097Z digest=sha256:bd073f5a85dab9aef108eef2ec9375e2793ece6e600dd8c3322c50018bcfea8f

Observation 75412ecb-9e26-49f5-adc9-a0033df6ccd8 · inbound

Continuous Speculative Decoding for Autoregressive Image Generation cites this paper.

Continuous Speculative Decoding for Autoregressive Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T18:38:40.166488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:38:40.166488Z digest=sha256:aee9366049a6e40d6b7696ba2c374c9a0a79c89c58cf5b1b6e62ee8eb300f2e6

Observation 3ab9b0db-78f6-4386-8037-753973592291 · inbound

High-Resolution Image Synthesis via Next-Token Prediction cites this paper.

High-Resolution Image Synthesis via Next-Token Prediction Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-12T14:55:42.128419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:55:42.128419Z digest=sha256:20e2ef773ec0bd7718b0902ed4e9be4db49bc1a5b9d8b4cbc03c9bad3378528e

Observation 4e718b6e-cb74-4811-b9ef-9792feeede6a · inbound

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE cites this paper.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.249652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.249652Z digest=sha256:be608b69774720aad55b03882c9e128bfdbc9cedcb8ffc67e91f69b1224b72b6

Observation 324b261a-1d65-47fa-9985-edba713facc7 · inbound

Collaborative Decoding Makes Visual Auto-Regressive Modeling Efficient cites this paper.

Collaborative Decoding Makes Visual Auto-Regressive Modeling Efficient Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T12:06:01.000349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:06:01.000349Z digest=sha256:cdd9f7d55c1578786fd17c1678bd2661b8f79ba78d8b373d401110b59c582310

Observation 10aaa647-03bb-4b2b-bdcc-5bfc0ec078e2 · inbound

OpenING: A Comprehensive Benchmark for Judging Open-ended Interleaved Image-Text Generation cites this paper.

OpenING: A Comprehensive Benchmark for Judging Open-ended Interleaved Image-Text Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T11:10:54.689429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:10:54.689429Z digest=sha256:b5c551be78a1d43b0608266761531f2c893d1de838c1c68f65486a60ac4ba7a3

Observation 6577574b-44c9-47d8-a068-36bed90e6a6a · inbound

Orthus: Autoregressive Interleaved Image-Text Generation with Modality-Specific Heads cites this paper.

Orthus: Autoregressive Interleaved Image-Text Generation with Modality-Specific Heads Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T10:35:22.651398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:35:22.651398Z digest=sha256:4d5d6172df4db07972cc55ed14245d5bc6b7b94430717a500cc57c235b4c1b71

Observation 3ca134d0-1df1-47ed-b85b-00d74882ee49 · inbound

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation cites this paper.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.145289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.145289Z digest=sha256:4c3041c97db6a5a1e8bc1702a4ecd5133eb2299d31994ef5db365c394e646ba2

Observation 31cde6aa-6ae1-4f83-842c-bddaf0fc6804 · inbound

X-Prompt: Towards Universal In-Context Image Generation in Auto-Regressive Vision Language Foundation Models cites this paper.

X-Prompt: Towards Universal In-Context Image Generation in Auto-Regressive Vision Language Foundation Models Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T00:56:59.639527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:56:59.639527Z digest=sha256:58d9613622d77b0b98c00d0faab1726da466f4455a4bd6ca93ea9a2d75b8e56f

Observation cd2555aa-4a88-491b-aa82-b381cddce663 · inbound

TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation cites this paper.

TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T22:54:20.534492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:54:20.534492Z digest=sha256:37b0bf5ce2b3d4806105313c9fd86a0c31da9520076f07170be7167136f4deef

Observation 62571cf7-2789-4a08-865d-877163f7177c · inbound

ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance cites this paper.

ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T19:29:52.643443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:29:52.643443Z digest=sha256:5f710c5b9eb3dc2ac52eea13aeca287e3bf875074908070be255c3fb3cdb89fd

Observation 5dce2f0c-126c-4b86-92b1-74f7830d77bc · inbound

Doe-1: Closed-Loop Autonomous Driving with Large World Model cites this paper.

Doe-1: Closed-Loop Autonomous Driving with Large World Model Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T16:55:42.337407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:55:42.337407Z digest=sha256:efd56701858c2ed2b3618f8608f344e9dbd524f76b97407d7f6735b2edc829a5

Observation d5b47ab1-12f1-45d0-865a-28405a8645ea · inbound

E-CAR: Efficient Continuous Autoregressive Image Generation via Multistage Modeling cites this paper.

E-CAR: Efficient Continuous Autoregressive Image Generation via Multistage Modeling Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T12:28:54.819899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:28:54.819899Z digest=sha256:54538203465400780f57d424136ab8f078a50920a7fd364245a5c79bef2fe5a2

Observation d58b3ad1-85fe-452a-9860-ccbba45bc9f2 · inbound

Parallelized Autoregressive Visual Generation cites this paper.

Parallelized Autoregressive Visual Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T11:42:31.652085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:42:31.652085Z digest=sha256:a5edcd23d99af38c9cdf79d4359ed0f9801612e3d0482c9bbf5420c07e868b3c

Observation 95406ae6-48ce-46d5-85ba-f8430b58703a · inbound

LANTERN++: Enhancing Relaxed Speculative Decoding with Static Tree Drafting for Visual Auto-regressive Models cites this paper.

LANTERN++: Enhancing Relaxed Speculative Decoding with Static Tree Drafting for Visual Auto-regressive Models Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T15:46:22.091202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:46:22.091202Z digest=sha256:fb760309fb82b03f84f9c8b882ed92f7387d7a3dd56769df75869bf6ba8ae455

Observation 6fbc2086-64f5-42a4-a939-35f47d22c474 · inbound

Lumina-Video: Efficient and Flexible Video Generation with Multi-scale Next-DiT cites this paper.

Lumina-Video: Efficient and Flexible Video Generation with Multi-scale Next-DiT Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T14:26:46.753819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:26:46.753819Z digest=sha256:0a426576f01cac26a17bdec26c48bea8fcfb156ee3d4ab6b08cf3896752baec4

Observation e773658c-3710-4996-9cd0-ffd41c0fdb40 · inbound

HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation cites this paper.

HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T20:21:30.454426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T20:21:30.454426Z digest=sha256:633e2270d4c5ba50b6bfb63ec77cec21183c6f50f4c23d60bd3ebdd98e1f5176

Observation 8d25135a-805d-4efd-97eb-8ddf379ddc46 · inbound

Personalized Text-to-Image Generation with Auto-Regressive Models cites this paper.

Personalized Text-to-Image Generation with Auto-Regressive Models Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T12:17:25.499383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:17:25.499383Z digest=sha256:702d444691d8d43fe6021cc3d1eaf92373bbe1ac32e62b85d7cce33d4459da2d

Observation 064ad557-cccb-4a1e-b1a5-6c67ea977343 · inbound

Generative Multimodal Pretraining with Discrete Diffusion Timestep Tokens cites this paper.

Generative Multimodal Pretraining with Discrete Diffusion Timestep Tokens Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T11:48:03.314257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:48:03.314257Z digest=sha256:8586d1c0bb20836e0cc5c1e1850c2305911981a73cf46a3bcf2fefa8be01bb6b

Observation 5e40635f-87ab-4833-a682-775729566e80 · inbound

Token-Shuffle: Towards High-Resolution Image Generation with Autoregressive Models cites this paper.

Token-Shuffle: Towards High-Resolution Image Generation with Autoregressive Models Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-16T10:35:43.582377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:35:43.582377Z digest=sha256:636386a61beb79511775f42258038388411897eb704db67e002e45cc2812df39

Observation 26bc7693-c9e1-4dce-ab19-04882e189e26 · inbound

IA-T2I: Internet-Augmented Text-to-Image Generation cites this paper.

IA-T2I: Internet-Augmented Text-to-Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:17:26.115967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:17:26.115967Z digest=sha256:569727215ddc390e676113937593ad6271d406eb8ddf2f6d97ac860aaa4cdc4e

Observation 70c81c4b-9bab-4fa5-a1e3-9800cc299360 · inbound

ChemMLLM: Chemical Multimodal Large Language Model cites this paper.

ChemMLLM: Chemical Multimodal Large Language Model Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:06:06.405683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:06:06.405683Z digest=sha256:98dfed1fb8d86d568b07753823a5724693a0409a11299b46d95fc6545acacb69

Observation f04e4ae7-fcef-48e0-92e0-93a264e8f051 · inbound

FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving cites this paper.

FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:19:42.913849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-15T19:19:42.748573Z digest=sha256:3a6b4afd51a8b1760c698eecd4b5a16811b3685c87f8eaedf42d7b0d454abda9

Observation 1adb9af2-121c-41eb-81d2-c622f8fed257 · inbound

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression cites this paper.

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:15.891755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:15.891755Z digest=sha256:9956735a03b334043ba955257e0b49e7e98049b2af43e8b89b4b6e8d8fdbd301

Observation 62cec599-c947-4a54-b4b7-ad3b636aa6b3 · inbound

StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation cites this paper.

StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:47.643689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:47.643689Z digest=sha256:086cd8e714a12883149eedf44621c3d0db48434df5707e5866683a693bff5f46

Observation c72fe73b-cab8-499d-8243-8f62cb7fce89 · inbound

UniRL: Self-Improving Unified Multimodal Models via Supervised and Reinforcement Learning cites this paper.

UniRL: Self-Improving Unified Multimodal Models via Supervised and Reinforcement Learning Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:53:23.757213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:53:23.757213Z digest=sha256:27db37ed430136b96fed6cf7a16efd69e2b78aba6d9d5192a74c6595eb740586

Observation d6250347-bda3-4597-adbc-03089dc1fd81 · inbound

Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model cites this paper.

Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:02:18.336941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-19T12:59:31.454155Z digest=sha256:36ac6043d096aa1f5be4e786cb305f755eb2286e099681aa09d6a0e6455899ed

Observation 49d3edac-ad4b-4652-88dc-a6fa4324f036 · inbound

Draw ALL Your Imagine: A Holistic Benchmark and Agent Framework for Complex Instruction-based Image Generation cites this paper.

Draw ALL Your Imagine: A Holistic Benchmark and Agent Framework for Complex Instruction-based Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:18:33.551160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:18:33.551160Z digest=sha256:57e8911fb581607dfa137312920bbc16def3f565764d65052c98a7764450b644

Observation 706012d4-a61f-401b-95a4-d33aad107f67 · inbound

SkipVAR: Accelerating Visual Autoregressive Modeling via Adaptive Frequency-Aware Skipping cites this paper.

SkipVAR: Accelerating Visual Autoregressive Modeling via Adaptive Frequency-Aware Skipping Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:05:00.140921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:05:00.140921Z digest=sha256:74f158b2c57c0c47aea2bdfbaee2638a683a8019661fa4b83706cde404496a55

Observation 8cce00cd-11e6-4d63-954a-b2f9934dd839 · inbound

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought cites this paper.

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:33:02.136479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-19T08:32:20.566798Z digest=sha256:b4ac9433f49fcc7d48f194313be5e7e982fb9b11c93a678c90d102c74e3fc78e

Observation 5ea88251-6d04-4cff-842b-fd6359a7cac2 · inbound

CycleVAR: Repurposing Autoregressive Model for Unsupervised One-Step Image Translation cites this paper.

CycleVAR: Repurposing Autoregressive Model for Unsupervised One-Step Image Translation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T21:49:43.534834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:49:43.534834Z digest=sha256:b47f1fa08ffb434e445679fa68f6f43155912238c81cbd6983a991462156a06c

Observation d2bd89df-c72d-4e5d-8a5c-2f7b2e9f196a · inbound

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer cites this paper.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.716761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.716761Z digest=sha256:729cb23313ae7407852f8e5efc033de4873ba8f2b2f27a5be359e6442a43c053

Observation 5bc01014-fdca-4031-bfd8-854eff4a6a24 · inbound

Resurrect Mask AutoRegressive Modeling for Efficient and Scalable Image Generation cites this paper.

Resurrect Mask AutoRegressive Modeling for Efficient and Scalable Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T16:37:35.505153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:37:35.505153Z digest=sha256:36a32922b8da2f029c810929ada2600b9d8be71844f03c0e320531f2153108f9

Observation 972c0fe8-36b9-454e-9097-5702aad21a43 · inbound

Lumina-mGPT 2.0: Stand-Alone AutoRegressive Image Modeling cites this paper.

Lumina-mGPT 2.0: Stand-Alone AutoRegressive Image Modeling Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-15T18:22:42.168483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:22:42.168483Z digest=sha256:21b648820acdecbe50052fabe2bd1c463486fbaf00f270eb1d003c3a06000440

Observation 59f2878a-3c54-45c2-be29-4e07cec1fb49 · inbound

UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation cites this paper.

UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:58.677295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:58.677295Z digest=sha256:bc6a309e62970c5066af3d24edcd2d1068eb25c7ae47f52342912877bcd5b88a

Observation 0905e324-4926-43f8-972c-984ff983303d · inbound

Exploring Autoregressive Vision Foundation Models for Image Compression cites this paper.

Exploring Autoregressive Vision Foundation Models for Image Compression Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T05:36:53.233971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:36:53.233971Z digest=sha256:0d0c69ffe9356efca35fa47bce1d90834ea9100b49d3277ced56959f6c9f522a

Observation 5a58b564-e316-4d5e-a931-3b5911feffed · inbound

Speculative Coupled Decoding for Training-Free Lossless Acceleration of Autoregressive Visual Generation cites this paper.

Speculative Coupled Decoding for Training-Free Lossless Acceleration of Autoregressive Visual Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-18T03:20:48.740044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-18T03:20:35.021120Z digest=sha256:e6594d11ecec70cb51393dab0b7f6a060be3c4fbaba87c7197db511a3ac2df2e

Observation 253ed8bd-367d-458c-8b6d-5b0899881c0e · inbound

VVS: Accelerating Speculative Decoding for Visual Autoregressive Generation via Partial Verification Skipping cites this paper.

VVS: Accelerating Speculative Decoding for Visual Autoregressive Generation via Partial Verification Skipping Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-17T21:35:17.580852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-17T21:34:14.371914Z digest=sha256:6685de5bf3c96f235d6c7d963c810a2f3e4a7e2675a1c8b9d885d683271ced5a

Observation 5d334b33-de6f-4a3d-a3ab-566b9c84a901 · inbound

dMLLM-TTS: Self-Verified and Efficient Test-Time Scaling for Diffusion Multi-Modal Large Language Models cites this paper.

dMLLM-TTS: Self-Verified and Efficient Test-Time Scaling for Diffusion Multi-Modal Large Language Models Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:24.554015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-16T20:38:06.225705Z digest=sha256:fea185c4bcf1436d6b7bca0cf2f603755fd37dc57b414c0d0e8c1e3196b67c2d

Observation 68de935d-430c-48c8-b6cd-1b6821744d63 · inbound

SJD-PAC: Accelerating Speculative Jacobi Decoding via Proactive Drafting and Adaptive Continuation cites this paper.

SJD-PAC: Accelerating Speculative Jacobi Decoding via Proactive Drafting and Adaptive Continuation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T22:30:39.049531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:30:39.049531Z digest=sha256:e18a41a51e1e4dfb6c2d9c543b1df2f8307f350b1fdfb6f4c35ad2f3890e38e1

Observation 1849eba3-8fac-4819-b16b-f7b260442cf6 · inbound

Drift-AR: Single-Step Visual Autoregressive Generation via Anti-Symmetric Drifting cites this paper.

Drift-AR: Single-Step Visual Autoregressive Generation via Anti-Symmetric Drifting Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T22:08:04.950577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-14T22:03:27.981438Z digest=sha256:b8bf0bb814f4dbb3cb377e5a8857a16753a1aca080869c0460f153a460d08bca

Observation d2a25238-8249-4e0f-a039-27e2163dfca9 · inbound

MAR-GRPO: Stabilized GRPO for AR-diffusion Hybrid Image Generation cites this paper.

MAR-GRPO: Stabilized GRPO for AR-diffusion Hybrid Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:40:58.100853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T18:00:50.105629Z digest=sha256:dda5734f7002fc2815c899fe465c77b7e9d25ffe2f8830d9ed4b5491aa105f39

Observation 47a61f61-ab51-416c-bde0-838e6bae69a0 · inbound

Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale cites this paper.

Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:14:10.770184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T05:00:31.717115Z digest=sha256:58e2f6ebc43aa8b9289e4ea4afb2f2cee96d0be9b5b4182c0bb210819dcfef55

Observation a4c460ff-9f89-481a-b00c-da3616c772c1 · inbound

TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering cites this paper.

TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:46:37.722392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-08T04:26:58.529645Z digest=sha256:0211c07ffdf3b515fcb6ce8c1420d188175107f90216db420dede5bd5319e9d6

Observation 51cdcfe1-67c0-418a-9f68-24f7917fab1b · inbound

OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation cites this paper.

OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:31:11.069089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-08T08:52:07.545191Z digest=sha256:9c85c3b52cd192e2b97899d85aaed8aa9220a66b7fb0f652d167568d52f83bee

Observation 12335dc9-37c8-408d-8b8d-b255ea492955 · inbound

CASCADE: Context-Aware Relaxation for Speculative Image Decoding cites this paper.

CASCADE: Context-Aware Relaxation for Speculative Image Decoding Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:55:55.170887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-11T02:08:27.374066Z digest=sha256:662b6a6bcd7bd221b9c1aaf77dfc5feaa06fcd3c6fa0c08ae52df5134e976151

Observation 6784e666-fbbb-45ff-a168-3c220c213ace · inbound

FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation cites this paper.

FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:21:16.463596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-12T02:20:25.071428Z digest=sha256:7934c0ef329ba8623532db57dabf4a799f5b3f15b63a579b79242bfd7d4ad84a

Observation 23a5030c-42ef-4c75-8a0a-fd9a0293c645 · inbound

FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation cites this paper.

FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:02:23.659868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-13T05:54:24.248910Z digest=sha256:efdf129f2caeb907b091157539815c6c27ac045294ac6165c2279f84a69d9bd6

Observation f8f0abe7-a300-4f66-ba5a-538d9d175aa8 · inbound

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving cites this paper.

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 125

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:31:26.229126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-12T04:13:37.421188Z digest=sha256:6d6f4c6a782a31a20c8814e4325704c77b5f71c9f36ec18ac19f180bcc4e8f39

Observation 8389d4b9-e0ff-4288-9d2a-07a3b79fa962 · inbound

Hierarchical Image Tokenization for Multi-Scale Image Super Resolution cites this paper.

Hierarchical Image Tokenization for Multi-Scale Image Super Resolution Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:15:04.280295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-30T21:11:36.019938Z digest=sha256:eab888465addcb7fe6e27755f482a5f4cc44bd083de223c6c26427880a60d667

Observation ff6b55aa-22f3-4d85-86b6-41417a31e54b · inbound

Where to Refine, When to Stop: Rethinking Redundancy via Latent Discrepancy for Efficient Visual Autoregressive Generation cites this paper.

Where to Refine, When to Stop: Rethinking Redundancy via Latent Discrepancy for Efficient Visual Autoregressive Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:42:46.717969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-28T22:38:26.147074Z digest=sha256:745345bd092a0e25b304d5eec42cdeb4ed7abd890ce56f5887bb6eb8403eb03b

Observation e145c616-d838-4db0-907f-0ff8c6e91adf · inbound

Polaris: Scaling Up Instruction-Guided Image Generation Towards Millions of Personalized Style Needs cites this paper.

Polaris: Scaling Up Instruction-Guided Image Generation Towards Millions of Personalized Style Needs Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:46:18.719326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-28T15:07:03.439928Z digest=sha256:657f110f19c9c094dac018d076b347da41bf79c832de0cf47c2d3ab84f4a003e

Observation 01af4b66-d9ae-4ba2-83dc-32fef732688e · inbound

Parallel Jacobi Decoding for Fast Autoregressive Image Generation cites this paper.

Parallel Jacobi Decoding for Fast Autoregressive Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:26:57.136218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-28T02:05:58.714029Z digest=sha256:0aea1f88901fbcb0ed840b519a418d37eb080f30b1830e5a3016e8ebedad5aad

Observation b02c2aed-bc8b-4cb6-afa6-d86d8910cb60 · inbound

Knowledge Distillation for Visual Autoregressive Models cites this paper.

Knowledge Distillation for Visual Autoregressive Models Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:36:56.883793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-28T01:59:51.409952Z digest=sha256:64a890020dd58e5635e297d416578d349569cb67141e6357f6d7cd48ed0a2598

Observation 2eb5c7f6-2ff7-426a-a41f-52ab3fd1cc5b · inbound

SSD: Spatially Speculative Decoding Accelerates Autoregressive Image Generation cites this paper.

SSD: Spatially Speculative Decoding Accelerates Autoregressive Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:49:30.630557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-26T17:36:26.762881Z digest=sha256:4f145ecd6c756d6372c3c065925beac553f403130233b3c83d8b1f1cc7bdf226

Observation 30cbcc5d-0224-4654-a5f0-3a94e05ddaf5 · inbound

From Open Loop to Closed Loop: A Test-Time Iterative Optimization Framework for Reference-Consistent Image Generation cites this paper.

From Open Loop to Closed Loop: A Test-Time Iterative Optimization Framework for Reference-Consistent Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-11T15:05:11.166049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T15:05:11.166049Z digest=sha256:341efeb4af365f4c22ad23f19dc59eb663e23842bd07964c92e9d41086ecf438

Observation d14b4f97-c2a7-4db3-acd5-ff77c4eea622 · inbound

Think, Plan, Paint: Layout-Aware Reasoning for Controllable Image Generation in Unified Models cites this paper.

Think, Plan, Paint: Layout-Aware Reasoning for Controllable Image Generation in Unified Models Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T21:07:59.570641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:07:59.570641Z digest=sha256:98f08021efc38a1dcb168fb868cb5fc1bb3884e9948427dd2b9da37c041cdff5

Observation d4586da7-a100-48dc-81c0-b86d38c43ba2 · inbound

SynVAR: Synergizing Spatial and Semantic Alignment in Visual Autoregressive Model cites this paper.

SynVAR: Synergizing Spatial and Semantic Alignment in Visual Autoregressive Model Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T00:44:45.940588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:44:45.940588Z digest=sha256:5b1ff58c5e256297486796c3ac0e602c522c23a43cf7b29fab0eb5777f979a10