Pith. sign in

Paper Citation Record · LEDGER

Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 43 inbound Pith citation observations for arXiv:2408.02657.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.02657 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 43 of 43 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T19:29:52.643443Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:49:30.628376Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 62571cf7-2789-4a08-865d-877163f7177c · inbound

ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance cites this paper.

ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T19:29:52.643443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:29:52.643443Z digest=sha256:e94d8e63d966022c50d4b3f8ce85352ea0d511d1d56684aee2e2db7e151a1130

Observation 5dce2f0c-126c-4b86-92b1-74f7830d77bc · inbound

Doe-1: Closed-Loop Autonomous Driving with Large World Model cites this paper.

Doe-1: Closed-Loop Autonomous Driving with Large World Model Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T16:55:42.337407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:55:42.337407Z digest=sha256:fc8d8a094aeb7a28a98c78ccf1f8f86c7d4073c052ac02278d07d1f65beacf8a

Observation d5b47ab1-12f1-45d0-865a-28405a8645ea · inbound

E-CAR: Efficient Continuous Autoregressive Image Generation via Multistage Modeling cites this paper.

E-CAR: Efficient Continuous Autoregressive Image Generation via Multistage Modeling Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T12:28:54.819899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:28:54.819899Z digest=sha256:c0fff6acb93a8ca28e5270716e8d4df5d293b7154ab0d119fefe403684556a7c

Observation d58b3ad1-85fe-452a-9860-ccbba45bc9f2 · inbound

Parallelized Autoregressive Visual Generation cites this paper.

Parallelized Autoregressive Visual Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T11:42:31.652085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:42:31.652085Z digest=sha256:8c4eec9e011cb7423e9140d2d6a6ee09ece7c93a7890e41898c8694790760914

Observation 95406ae6-48ce-46d5-85ba-f8430b58703a · inbound

LANTERN++: Enhancing Relaxed Speculative Decoding with Static Tree Drafting for Visual Auto-regressive Models cites this paper.

LANTERN++: Enhancing Relaxed Speculative Decoding with Static Tree Drafting for Visual Auto-regressive Models Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T15:46:22.091202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:46:22.091202Z digest=sha256:866814f72caa3d5267b80d65705cae6bd2939aab18e12e1ad0fa607cbb68eee3

Observation 6fbc2086-64f5-42a4-a939-35f47d22c474 · inbound

Lumina-Video: Efficient and Flexible Video Generation with Multi-scale Next-DiT cites this paper.

Lumina-Video: Efficient and Flexible Video Generation with Multi-scale Next-DiT Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T14:26:46.753819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:26:46.753819Z digest=sha256:d38a7ab29d73c225ccbf2711715a2a4a80d684c1f576515da11df2d48a91c017

Observation e773658c-3710-4996-9cd0-ffd41c0fdb40 · inbound

HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation cites this paper.

HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T20:21:30.454426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T20:21:30.454426Z digest=sha256:7f8ba29e16e6822cc3909b8ecb7ce20682d71cbf72115da710859bf6e3d19db4

Observation 26bc7693-c9e1-4dce-ab19-04882e189e26 · inbound

IA-T2I: Internet-Augmented Text-to-Image Generation cites this paper.

IA-T2I: Internet-Augmented Text-to-Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:17:26.115967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:17:26.115967Z digest=sha256:eec5cc3d4fdbfd9a89e1932839ef54c4465d44f33a4938f6605b6b07481ed748

Observation 70c81c4b-9bab-4fa5-a1e3-9800cc299360 · inbound

ChemMLLM: Chemical Multimodal Large Language Model cites this paper.

ChemMLLM: Chemical Multimodal Large Language Model Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:06:06.405683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:06:06.405683Z digest=sha256:11a090dbb29511f41bb7f5c7774912a1dc396275e491c741f3fbe52577c4d121

Observation f04e4ae7-fcef-48e0-92e0-93a264e8f051 · inbound

FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving cites this paper.

FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:19:42.913849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T19:19:42.748573Z digest=sha256:23e327ba0c6a6b41d0bdf628b4fc5913996981389fe4b22331b4f4be364923a8

Observation 1adb9af2-121c-41eb-81d2-c622f8fed257 · inbound

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression cites this paper.

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:15.891755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:15.891755Z digest=sha256:92e4a78ee1ec2d9259dbc106d20e0762798f959e7d3115b1861c58cb990985ec

Observation 62cec599-c947-4a54-b4b7-ad3b636aa6b3 · inbound

StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation cites this paper.

StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:47.643689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:47.643689Z digest=sha256:6bcf6fb2ff52d88f07cf4b1053ad809ca26cd7122580f80349de02befdb81064

Observation c72fe73b-cab8-499d-8243-8f62cb7fce89 · inbound

UniRL: Self-Improving Unified Multimodal Models via Supervised and Reinforcement Learning cites this paper.

UniRL: Self-Improving Unified Multimodal Models via Supervised and Reinforcement Learning Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:53:23.757213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:53:23.757213Z digest=sha256:a36609dc095d665d2f8304f28ad7ccbc8b4b1193d128d450a80bd7db1d7e0b24

Observation d6250347-bda3-4597-adbc-03089dc1fd81 · inbound

Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model cites this paper.

Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:02:18.336941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-19T12:59:31.454155Z digest=sha256:a8400bd0e044f3f84032fa54f0f641c3f64aa00b946030f3aa6f2407e61b2d8f

Observation 49d3edac-ad4b-4652-88dc-a6fa4324f036 · inbound

Draw ALL Your Imagine: A Holistic Benchmark and Agent Framework for Complex Instruction-based Image Generation cites this paper.

Draw ALL Your Imagine: A Holistic Benchmark and Agent Framework for Complex Instruction-based Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:18:33.551160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:18:33.551160Z digest=sha256:e4441e0651190ff361621622c2153fd9e8c89b5187a3b5028a01b46736af8860

Observation 706012d4-a61f-401b-95a4-d33aad107f67 · inbound

SkipVAR: Accelerating Visual Autoregressive Modeling via Adaptive Frequency-Aware Skipping cites this paper.

SkipVAR: Accelerating Visual Autoregressive Modeling via Adaptive Frequency-Aware Skipping Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:05:00.140921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:05:00.140921Z digest=sha256:94932db7c84e693d3687dc5b2ab828c8f256507347b0dd678bf39655f7451654

Observation 8cce00cd-11e6-4d63-954a-b2f9934dd839 · inbound

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought cites this paper.

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:33:02.136479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-19T08:32:20.566798Z digest=sha256:44dc073e6a4f531b2ad8ba005ea6ec0526eeda7aa25a56301d6b5f129340023f

Observation 5ea88251-6d04-4cff-842b-fd6359a7cac2 · inbound

CycleVAR: Repurposing Autoregressive Model for Unsupervised One-Step Image Translation cites this paper.

CycleVAR: Repurposing Autoregressive Model for Unsupervised One-Step Image Translation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T21:49:43.534834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:49:43.534834Z digest=sha256:afa11f1b1f4bbed92f2c51b9eb1ad7a4090fd9a878acebee08cc8be89cb8a6ed

Observation d2bd89df-c72d-4e5d-8a5c-2f7b2e9f196a · inbound

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer cites this paper.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.716761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.716761Z digest=sha256:0e687c40704f497ff6da12c638c2c9d7ebe4094b701f2c2dd84c66c990dd9ce1

Observation 5bc01014-fdca-4031-bfd8-854eff4a6a24 · inbound

Resurrect Mask AutoRegressive Modeling for Efficient and Scalable Image Generation cites this paper.

Resurrect Mask AutoRegressive Modeling for Efficient and Scalable Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T16:37:35.505153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:37:35.505153Z digest=sha256:d6bb5020bc7d4da9777cd204c13fd422234a1e1994ea3990e4695fbbf35035f5

Observation 59f2878a-3c54-45c2-be29-4e07cec1fb49 · inbound

UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation cites this paper.

UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:58.677295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:58.677295Z digest=sha256:209b1c59e8ff709c0b69fbb4ddd8b159396707e70b31323de2dab2f52dd6c8ee

Observation 0905e324-4926-43f8-972c-984ff983303d · inbound

Exploring Autoregressive Vision Foundation Models for Image Compression cites this paper.

Exploring Autoregressive Vision Foundation Models for Image Compression Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T05:36:53.233971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:36:53.233971Z digest=sha256:8a3b07194fc85902c74574d73a325bfab8d3e1132ae7aa72cdb38046e76a38e1

Observation 5a58b564-e316-4d5e-a931-3b5911feffed · inbound

Speculative Coupled Decoding for Training-Free Lossless Acceleration of Autoregressive Visual Generation cites this paper.

Speculative Coupled Decoding for Training-Free Lossless Acceleration of Autoregressive Visual Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-18T03:20:48.740044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T03:20:35.021120Z digest=sha256:4b4e4c8780eff0f2abfbd8b2ea9e24c70462b3528ab98e65e4151d26ba6bffc2

Observation 253ed8bd-367d-458c-8b6d-5b0899881c0e · inbound

VVS: Accelerating Speculative Decoding for Visual Autoregressive Generation via Partial Verification Skipping cites this paper.

VVS: Accelerating Speculative Decoding for Visual Autoregressive Generation via Partial Verification Skipping Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-17T21:35:17.580852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-17T21:34:14.371914Z digest=sha256:308a88dc5101699f5fe2fd7391a69d4bcb7c16e3824069d8bc786ae2b7887916

Observation 5d334b33-de6f-4a3d-a3ab-566b9c84a901 · inbound

dMLLM-TTS: Self-Verified and Efficient Test-Time Scaling for Diffusion Multi-Modal Large Language Models cites this paper.

dMLLM-TTS: Self-Verified and Efficient Test-Time Scaling for Diffusion Multi-Modal Large Language Models Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:24.554015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-16T20:38:06.225705Z digest=sha256:9a03da9abe66aaa29f8108bbd6938de9121ca343c332b23d9bd514988038a8dd

Observation 68de935d-430c-48c8-b6cd-1b6821744d63 · inbound

SJD-PAC: Accelerating Speculative Jacobi Decoding via Proactive Drafting and Adaptive Continuation cites this paper.

SJD-PAC: Accelerating Speculative Jacobi Decoding via Proactive Drafting and Adaptive Continuation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T22:30:39.049531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:30:39.049531Z digest=sha256:6313631ca306bf82df33fbdde3cccc66682a75659e2273bce0a7d4725e0eb2cf

Observation 1849eba3-8fac-4819-b16b-f7b260442cf6 · inbound

Drift-AR: Single-Step Visual Autoregressive Generation via Anti-Symmetric Drifting cites this paper.

Drift-AR: Single-Step Visual Autoregressive Generation via Anti-Symmetric Drifting Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T22:08:04.950577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-14T22:03:27.981438Z digest=sha256:0feb1b00015bd7ba9db7e4d43ea5a7165dd7aeec90e8c515ff5ba44e99d7875a

Observation d2a25238-8249-4e0f-a039-27e2163dfca9 · inbound

MAR-GRPO: Stabilized GRPO for AR-diffusion Hybrid Image Generation cites this paper.

MAR-GRPO: Stabilized GRPO for AR-diffusion Hybrid Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:40:58.100853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T18:00:50.105629Z digest=sha256:81d4ddc681f9fda08f9c36bf2288a754611042746b15a5400a95055204ba940f

Observation 47a61f61-ab51-416c-bde0-838e6bae69a0 · inbound

Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale cites this paper.

Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:14:10.770184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T05:00:31.717115Z digest=sha256:bf97426f97da3a171660d508d78619eaa4f712f41d92a25dfc5061d7f369fa8b

Observation a4c460ff-9f89-481a-b00c-da3616c772c1 · inbound

TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering cites this paper.

TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:46:37.722392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T04:26:58.529645Z digest=sha256:e27eb2b55977214d3a580da9baee180f352d9e36215cdf2ec3c32d53dc1ffa33

Observation 51cdcfe1-67c0-418a-9f68-24f7917fab1b · inbound

OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation cites this paper.

OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:31:11.069089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T08:52:07.545191Z digest=sha256:5eb50184e783e1980be0e467dbbb089d6ccc4622aa36b9036de02a0f84c27e9b

Observation 12335dc9-37c8-408d-8b8d-b255ea492955 · inbound

CASCADE: Context-Aware Relaxation for Speculative Image Decoding cites this paper.

CASCADE: Context-Aware Relaxation for Speculative Image Decoding Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:55:55.170887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-11T02:08:27.374066Z digest=sha256:a7ec4ce7c2840c9cd44e9c5173d966b54512e9714cb5981a62a5a827571bf847

Observation 6784e666-fbbb-45ff-a168-3c220c213ace · inbound

FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation cites this paper.

FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:21:16.463596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-12T02:20:25.071428Z digest=sha256:645226217b24d0fe4ae6547b00182d36dc65467d67e033aa70d352d6b35d7d69

Observation 23a5030c-42ef-4c75-8a0a-fd9a0293c645 · inbound

FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation cites this paper.

FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:02:23.659868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-13T05:54:24.248910Z digest=sha256:3e12306d9db81f2594b11be9807e49cb040725986f2f683271b9c61dda6a42eb

Observation f8f0abe7-a300-4f66-ba5a-538d9d175aa8 · inbound

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving cites this paper.

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 125

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:31:26.229126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T04:13:37.421188Z digest=sha256:b2454a14411bc290975261d85d291cec5f03876f5683bf027a03d1b9df38f83e

Observation 8389d4b9-e0ff-4288-9d2a-07a3b79fa962 · inbound

Hierarchical Image Tokenization for Multi-Scale Image Super Resolution cites this paper.

Hierarchical Image Tokenization for Multi-Scale Image Super Resolution Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:15:04.280295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-30T21:11:36.019938Z digest=sha256:e3f748ebd6426515d4befa6877a34339207324ccc72c61932d4f2e423aa816f3

Observation ff6b55aa-22f3-4d85-86b6-41417a31e54b · inbound

Where to Refine, When to Stop: Rethinking Redundancy via Latent Discrepancy for Efficient Visual Autoregressive Generation cites this paper.

Where to Refine, When to Stop: Rethinking Redundancy via Latent Discrepancy for Efficient Visual Autoregressive Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:42:46.717969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-28T22:38:26.147074Z digest=sha256:36a2908a74e04a04f3422f04196f02549e03510af80f4271dfd494b87cbb74a2

Observation e145c616-d838-4db0-907f-0ff8c6e91adf · inbound

Polaris: Scaling Up Instruction-Guided Image Generation Towards Millions of Personalized Style Needs cites this paper.

Polaris: Scaling Up Instruction-Guided Image Generation Towards Millions of Personalized Style Needs Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:46:18.719326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T15:07:03.439928Z digest=sha256:40b3a9a5bf58a8f35749ee3d36a2a7fea81c4c1c316812fee8f8c312a909be67

Observation 01af4b66-d9ae-4ba2-83dc-32fef732688e · inbound

Parallel Jacobi Decoding for Fast Autoregressive Image Generation cites this paper.

Parallel Jacobi Decoding for Fast Autoregressive Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:26:57.136218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-28T02:05:58.714029Z digest=sha256:5a63fa7b4e7405337acba7f8a534378358cb06f052844c74fd48de77aa83e63c

Observation b02c2aed-bc8b-4cb6-afa6-d86d8910cb60 · inbound

Knowledge Distillation for Visual Autoregressive Models cites this paper.

Knowledge Distillation for Visual Autoregressive Models Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:36:56.883793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-28T01:59:51.409952Z digest=sha256:56616c97f21058e27e3e46189d37763d279f829c3210be161e9290fe69001244

Observation 2eb5c7f6-2ff7-426a-a41f-52ab3fd1cc5b · inbound

SSD: Spatially Speculative Decoding Accelerates Autoregressive Image Generation cites this paper.

SSD: Spatially Speculative Decoding Accelerates Autoregressive Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:49:30.630557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-26T17:36:26.762881Z digest=sha256:7dba4f21b747a7afbb9d0070695bb16334f548511754153e29383c6c3e639ce8

Observation 30cbcc5d-0224-4654-a5f0-3a94e05ddaf5 · inbound

From Open Loop to Closed Loop: A Test-Time Iterative Optimization Framework for Reference-Consistent Image Generation cites this paper.

From Open Loop to Closed Loop: A Test-Time Iterative Optimization Framework for Reference-Consistent Image Generation Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-11T15:05:11.166049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T15:05:11.166049Z digest=sha256:00829b59d087dd13c3627d3938c51fed97529191335b48b461fc695727fcc765

Observation d14b4f97-c2a7-4db3-acd5-ff77c4eea622 · inbound

Think, Plan, Paint: Layout-Aware Reasoning for Controllable Image Generation in Unified Models cites this paper.

Think, Plan, Paint: Layout-Aware Reasoning for Controllable Image Generation in Unified Models Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T21:07:59.570641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:07:59.570641Z digest=sha256:f6d455cc59bf6ce1394263340148940f3c34e12b178fe998a7c34432ef6db45b