Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-15T14:50:59.661153Z
Paper Citation Record · LEDGER
As of 3 August 2026, this Paper Citation Record lists 100 of 126 outbound references and 69 inbound Pith citation observations for arXiv:2505.15809.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-15T14:50:59.661153Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T16:21:38.232907Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-11T03:07:51.413936Z
100 of 126 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation fb949630-e33e-4d9c-b529-7fe530ed1b86 · outbound
MMaDA: Multimodal Large Diffusion Language Models Improving language understanding by generative pre-training
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a2fb3350-ce98-49dc-8bf6-8228dc4af0f7 · outbound
MMaDA: Multimodal Large Diffusion Language Models Language models are few-shot learners
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 921f9849-59d5-4d60-a968-b371ad70fa52 · outbound
MMaDA: Multimodal Large Diffusion Language Models OpenAI o1 System Card
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0d456aa7-05ec-41fe-a08f-a3a0500c832c · outbound
MMaDA: Multimodal Large Diffusion Language Models VL-GPT: A Generative Pre-trained Transformer for Vision and Language Understanding and Generation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 24ae7dfa-e2f3-4a95-857e-278095069e6f · outbound
MMaDA: Multimodal Large Diffusion Language Models Emu: Generative Pretraining in Multimodality
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ec11eb0b-6777-46b3-b61b-780cc315fee6 · outbound
MMaDA: Multimodal Large Diffusion Language Models Generative Multimodal Models are In-Context Learners
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d9716231-2902-45fa-84b5-fae5bfbace46 · outbound
MMaDA: Multimodal Large Diffusion Language Models Chameleon: Mixed-Modal Early-Fusion Foundation Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0e418431-77d6-4234-87bb-f6c2b035c798 · outbound
MMaDA: Multimodal Large Diffusion Language Models World model on million-length video and language with ringattention
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 571b9114-4fb5-4417-8ff5-6fef32f79f08 · outbound
MMaDA: Multimodal Large Diffusion Language Models VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ec1089e9-f5f8-4811-9685-b0ca27da6215 · outbound
MMaDA: Multimodal Large Diffusion Language Models Emu: Generative pretraining in multimodality
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 90dc9801-715f-437f-8f3a-fda1bd743430 · outbound
MMaDA: Multimodal Large Diffusion Language Models Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b996e979-8360-48f4-b03c-9fd914e15219 · outbound
MMaDA: Multimodal Large Diffusion Language Models Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7c44082a-23b0-4d85-9b1e-dffd36223999 · outbound
MMaDA: Multimodal Large Diffusion Language Models Gpt-4 technical report
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 19b8fbdd-2b1d-4fe5-9a6b-71628249ff35 · outbound
MMaDA: Multimodal Large Diffusion Language Models DreamLLM: Synergistic multimodal comprehension and creation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6053d725-9f06-4d06-8744-80528a8cd981 · outbound
MMaDA: Multimodal Large Diffusion Language Models Generating images with multimodal language models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1131beae-9237-465c-aa6b-c535e9745890 · outbound
MMaDA: Multimodal Large Diffusion Language Models Llava-plus: Learning to use tools for creating multimodal agents
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9e61143c-38eb-40b6-b20d-e5576defed9b · outbound
MMaDA: Multimodal Large Diffusion Language Models SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0b0ba415-f8e8-4a23-bfa3-0bf9125946ef · outbound
MMaDA: Multimodal Large Diffusion Language Models Show-o: One Single Transformer to Unify Multimodal Understanding and Generation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 867ee91a-bbfd-41da-8309-17a8fbfb801b · outbound
MMaDA: Multimodal Large Diffusion Language Models Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f335ff4b-d157-4cd3-8d00-4fcb2a3e26d2 · outbound
MMaDA: Multimodal Large Diffusion Language Models Large Language Diffusion Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 62406cb4-bb3f-4aa6-a51d-0720ff0926b3 · outbound
MMaDA: Multimodal Large Diffusion Language Models Denoising diffusion probabilistic models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 47ddc277-904f-43b0-a2d7-987c74eeb0ae · outbound
MMaDA: Multimodal Large Diffusion Language Models Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 182cc9b2-c698-4796-a379-09b667703afc · outbound
MMaDA: Multimodal Large Diffusion Language Models Emu3: Next-Token Prediction is All You Need
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4731bda0-a302-4137-ad12-e883f6e02947 · outbound
MMaDA: Multimodal Large Diffusion Language Models Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 702b57a8-b5f9-40ad-983f-3436b1ea454b · outbound
MMaDA: Multimodal Large Diffusion Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d1574be5-1323-4b1a-9bb0-075f3eb242ba · outbound
MMaDA: Multimodal Large Diffusion Language Models d1: Scaling reasoning in diffusion large language models via reinforcement learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ed9aa9d7-c1bd-423b-9044-166a6e6b1be7 · outbound
MMaDA: Multimodal Large Diffusion Language Models Xing, and Liang Lin
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1fe952ff-2cde-489a-a6b1-8c4983afaec8 · outbound
MMaDA: Multimodal Large Diffusion Language Models Lawrence Zitnick, and Ross Girshick
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 3e29b659-8c9e-4825-b0c3-48b94a10db48 · outbound
MMaDA: Multimodal Large Diffusion Language Models Improved baselines with visual instruction tuning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ff1def59-772d-428e-a4ef-0569566a1f0d · outbound
MMaDA: Multimodal Large Diffusion Language Models Instructblip: Towards general-purpose vision-language models with instruction tuning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 759a2b7b-9112-4ba1-bd89-794b230386cf · outbound
MMaDA: Multimodal Large Diffusion Language Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8a4a4225-c644-40ef-85da-e8588c7d0ea0 · outbound
MMaDA: Multimodal Large Diffusion Language Models mplug-owl2: Revolutionizing multi-modal large language model with modality collaboration
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 51dbc47a-cd85-4cfb-9901-3c1590ec3fea · outbound
MMaDA: Multimodal Large Diffusion Language Models Llava-phi: Efficient multi-modal assistant with small language model
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 31acadb4-d211-49c0-8d7b-30780d80408c · outbound
MMaDA: Multimodal Large Diffusion Language Models The refinedweb dataset for falcon LLM: outperforming curated corpora with web data only
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 23f19fd7-322a-4a97-a093-5d810bbad181 · outbound
MMaDA: Multimodal Large Diffusion Language Models Imagenet: A large-scale hierarchical image database
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e838c433-1215-44f3-9ce2-9bf5ff23d73e · outbound
MMaDA: Multimodal Large Diffusion Language Models Conceptual 12m: Pushing web-scale image-text pre-training to recognize long-tail visual concepts
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 722114fa-ddfd-415f-8ec7-eacb352f28a5 · outbound
MMaDA: Multimodal Large Diffusion Language Models Segment anything
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f25b407c-9c2b-4b1b-8e29-f0758e096c08 · outbound
MMaDA: Multimodal Large Diffusion Language Models laion-aesthetics-12m-umap
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation cda28788-7f4e-42f9-a41b-6a09ceb63cde · outbound
MMaDA: Multimodal Large Diffusion Language Models Journeydb: A benchmark for generative image understanding
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 98246a7a-4802-440b-951e-9e615c656d70 · outbound
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a0c2d6af-8636-467e-9d76-cdcf6fdaf682 · outbound
MMaDA: Multimodal Large Diffusion Language Models ReasonFlux: Hierarchical LLM Reasoning via Scaling Thought Templates
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 78a86927-f67a-4792-b7c6-3df24d26a94a · outbound
MMaDA: Multimodal Large Diffusion Language Models Limo: Less is more for reasoning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4ef0f7f0-3045-41d0-87c7-051708ccf443 · outbound
MMaDA: Multimodal Large Diffusion Language Models s1: Simple test-time scaling
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 212ea4a7-d9a5-4017-9faa-0ffc325155ef · outbound
MMaDA: Multimodal Large Diffusion Language Models Open Thoughts
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f601dd0e-860b-49b8-9bc0-d5bc2e3c78c5 · outbound
MMaDA: Multimodal Large Diffusion Language Models Acemath: Advancing frontier math reasoning with post-training and reward modeling
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5ef9c87f-ce9e-4e93-b53f-8e5202aede9b · outbound
MMaDA: Multimodal Large Diffusion Language Models Lmm-r1: Empowering 3b lmms with strong reasoning abilities through two-stage rule-based rl
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 33dcdefe-e59e-430d-8966-2633d29c5969 · outbound
MMaDA: Multimodal Large Diffusion Language Models Training Verifiers to Solve Math Word Problems
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5311a623-c235-45d2-bc4f-373c347057c0 · outbound
MMaDA: Multimodal Large Diffusion Language Models Learning transferable visual models from natural language supervision
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 07cd823d-cbc9-4016-a31d-a06def6de80e · outbound
MMaDA: Multimodal Large Diffusion Language Models Imagereward: Learning and evaluating human preferences for text-to-image generation.Advances in Neural Information Processing Systems, 36
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation fa6e826d-f70f-4b22-99ca-3c632c35d5f4 · outbound
MMaDA: Multimodal Large Diffusion Language Models Geneval: An object-focused framework for evaluating text-to-image alignment
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5d67eccc-d9d6-45be-92dc-f8b664f910f4 · outbound
MMaDA: Multimodal Large Diffusion Language Models WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6127edc3-50d1-463e-8830-0832603bb267 · outbound
MMaDA: Multimodal Large Diffusion Language Models High-resolution image synthesis with latent diffusion models
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b57b3c1b-f38c-43cc-84c9-7533e36819c4 · outbound
MMaDA: Multimodal Large Diffusion Language Models SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 510bc526-9323-432f-a5c1-cf8a86e89652 · outbound
MMaDA: Multimodal Large Diffusion Language Models Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e6b68b22-0eb4-4371-9e7a-5af03bd4b398 · outbound
MMaDA: Multimodal Large Diffusion Language Models VARGPT: Unified Understanding and Generation in a Visual Autoregressive Multimodal Large Language Model
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7cf3fa26-7b8a-44aa-b7fb-8a79f285668d · outbound
MMaDA: Multimodal Large Diffusion Language Models Gemini: A Family of Highly Capable Multimodal Models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9c078740-7d9f-4a6b-8861-875e33ca3925 · outbound
MMaDA: Multimodal Large Diffusion Language Models Openai o1 system card.preprint
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b4fae0d7-6a48-4de8-8ee4-c4d75be11488 · outbound
MMaDA: Multimodal Large Diffusion Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 92383866-01a2-4c19-959c-3a824be15364 · outbound
MMaDA: Multimodal Large Diffusion Language Models Multimodal foundation models: From specialists to general-purpose assistants.Foundations and Trends® in Computer Graphics and Vision, 16(1-2):1–214
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 253b4759-b89e-45ff-bce7-2c1833eac5e3 · outbound
MMaDA: Multimodal Large Diffusion Language Models A Survey on Multimodal Large Language Models
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 37c64dfe-60db-43e3-b996-6bb00ec520d7 · outbound
MMaDA: Multimodal Large Diffusion Language Models Hallucination of Multimodal Large Language Models: A Survey
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 71799f63-ff7e-4461-8c58-c67cb357dcb4 · outbound
MMaDA: Multimodal Large Diffusion Language Models Visual instruction tuning.NeurIPS, 36
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation eb458215-0d14-4806-904f-b4a4fbf4b0c8 · outbound
MMaDA: Multimodal Large Diffusion Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2c1fac96-d5eb-49f7-879a-bc2c985ea8ee · outbound
MMaDA: Multimodal Large Diffusion Language Models Qwen Technical Report
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0c45f6b8-6c9c-4d34-beed-57b63926474f · outbound
MMaDA: Multimodal Large Diffusion Language Models MM1: Methods, Analysis & Insights from Multimodal LLM Pre-training
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation bbd1b1fa-0b9c-4025-be4d-057ddaff9198 · outbound
MMaDA: Multimodal Large Diffusion Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d9dfa50b-319f-4700-b3da-9d70218584c3 · outbound
MMaDA: Multimodal Large Diffusion Language Models Hierarchical Text-Conditional Image Generation with CLIP Latents
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1794a132-4200-4fb3-bc8e-bf60d18828b9 · outbound
MMaDA: Multimodal Large Diffusion Language Models Kwok, Ping Luo, Huchuan Lu, and Zhenguo Li
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 248141da-dd18-4816-a150-b61bbafd69f5 · outbound
MMaDA: Multimodal Large Diffusion Language Models GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8e93e53e-8629-4271-8caf-c93d509f8562 · outbound
MMaDA: Multimodal Large Diffusion Language Models Raphael: Text-to- image generation via large mixture of diffusion paths.NeurIPS, 36
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 949f6199-9132-4b85-a745-fcc2a450ad08 · outbound
MMaDA: Multimodal Large Diffusion Language Models Mastering text-to-image diffusion: Recaptioning, planning, and generating with multimodal llms
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0d88f640-9c5a-4b81-9635-46bea4750d2e · outbound
MMaDA: Multimodal Large Diffusion Language Models Itercomp: Iterative composition-aware feedback learning from model gallery for text-to-image generation
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f7995592-57e2-4dfd-b7cc-6b9e0d61a067 · outbound
MMaDA: Multimodal Large Diffusion Language Models An Overview of Diffusion Models: Applications, Guided Generation, Statistical Rates and Optimization
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 3df2197d-09f1-4cb0-9dbd-59dcb47bd885 · outbound
MMaDA: Multimodal Large Diffusion Language Models Diffusion-Sharpening: Fine-tuning Diffusion Models with Denoising Trajectory Sharpening
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d89783d8-ee86-447f-82de-4b03473bc5e9 · outbound
MMaDA: Multimodal Large Diffusion Language Models Reward-directed conditional diffusion: Provable distribution estimation and reward improvement.Advancesin Neural Information Processing Systems, 36:60599–60635
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 21ad3f4a-8f13-4fb1-9598-168fdd07bd19 · outbound
MMaDA: Multimodal Large Diffusion Language Models Gradient Guidance for Diffusion Models: An Optimization Perspective
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b4237290-8605-422c-91d7-6b17f0d12439 · outbound
MMaDA: Multimodal Large Diffusion Language Models Score approximation, estimation and distribution recovery of diffusion models on low-dimensional data
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c5430022-d2a2-46f6-8bb5-2dcde028a193 · outbound
MMaDA: Multimodal Large Diffusion Language Models Rectified Diffusion: Straightness Is Not Your Need in Rectified Flow
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a6f5c7d5-a4d4-4d66-857c-e97fe9dae415 · outbound
MMaDA: Multimodal Large Diffusion Language Models Improving diffusion-based image synthesis with context prediction.Advances in Neural Information Processing Systems, 36:37636–37656
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 543b501d-f5d6-4375-bfba-703bac231bfa · outbound
MMaDA: Multimodal Large Diffusion Language Models Structure-guided adversarial training of diffusion models
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 76aaedea-6f6d-4064-a725-783986063885 · outbound
MMaDA: Multimodal Large Diffusion Language Models Videotetris: Towards compositional text-to-video generation.Advancesin Neural Information Processing Systems, 37:29489–29513
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 049a8ebd-8223-4cb0-afd1-22fd19b05dfc · outbound
MMaDA: Multimodal Large Diffusion Language Models Structured denoising diffusion models in discrete state-spaces.NeurIPS, pages 17981–17993
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 790d0e44-e598-49d6-b8ea-bdfc16c93625 · outbound
MMaDA: Multimodal Large Diffusion Language Models Vector quantized diffusion model for text-to-image synthesis
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2362884e-f3b6-4043-a4d3-5e4239975d83 · outbound
MMaDA: Multimodal Large Diffusion Language Models Gritsenko, Jasmijn Bastings, Ben Poole, Rianne van den Berg, and Tim Salimans
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 59ab54db-1ca8-4f3a-8657-8526c0e9bb87 · outbound
MMaDA: Multimodal Large Diffusion Language Models Murphy.Probabilistic Machine Learning: AdvancedTopics
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation fdd5bb84-3abf-41d9-9342-1657e0830363 · outbound
MMaDA: Multimodal Large Diffusion Language Models Rethinking the objectives of vector- quantized tokenizers for image synthesis
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 19c4bfa5-7d2b-4891-84d7-e719c1307237 · outbound
MMaDA: Multimodal Large Diffusion Language Models Attention is all you need.NeurIPS, 30
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 09b78d8a-4575-4647-a26d-de6db18337e0 · outbound
MMaDA: Multimodal Large Diffusion Language Models Exploring the limits of transfer learning with a unified text-to-text transformer.Journal of machine learning research, 21(140):1–67
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation fb099ec4-5ddd-4b41-b20b-b2fbd04e324c · outbound
MMaDA: Multimodal Large Diffusion Language Models LLaMA: Open and Efficient Foundation Language Models
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0f1742aa-d590-4dc5-87b0-f2e39c33fe95 · outbound
MMaDA: Multimodal Large Diffusion Language Models Image transformer
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 25311f4d-123b-48d5-9cc9-ba4344ced037 · outbound
MMaDA: Multimodal Large Diffusion Language Models Taming transformers for high-resolution image synthesis
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ad00e2d3-c149-4788-9428-78d0128f8b05 · outbound
MMaDA: Multimodal Large Diffusion Language Models Classification accuracy score for conditional generative models.NeurIPS, 32
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 03201462-0c46-4cba-bc4f-0974193d3968 · outbound
MMaDA: Multimodal Large Diffusion Language Models Generative pretraining from pixels
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e9eef0f7-ce2b-4a08-8c80-2884929218b5 · outbound
MMaDA: Multimodal Large Diffusion Language Models VideoPoet: A Large Language Model for Zero-Shot Video Generation
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 85e03685-b3c2-428e-9944-0f96f0b08232 · outbound
MMaDA: Multimodal Large Diffusion Language Models Visual autoregressive modeling: Scalable image generation via next-scale prediction.Advancesin neural information processing systems, 37:84839–84865
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation bfb03666-dc73-483e-a948-ce00f2e52775 · outbound
MMaDA: Multimodal Large Diffusion Language Models NExT-GPT: Any-to-Any Multimodal LLM
Reference 101
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d0aea713-0a11-434f-91c3-b8bf8e3b0ae0 · outbound
MMaDA: Multimodal Large Diffusion Language Models Any-to-any generation via composable diffusion
Reference 102
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation abff7c5e-1703-4c7e-8ceb-c800c58692f3 · outbound
MMaDA: Multimodal Large Diffusion Language Models X-VILA: Cross-Modality Alignment for Large Language Model
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 427529eb-446c-43df-a223-7c2d03618e88 · outbound
MMaDA: Multimodal Large Diffusion Language Models Jointly training large autoregressive multimodal models
Reference 104
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f698e10f-5ba2-467b-8996-e285e6f2fe01 · outbound
MMaDA: Multimodal Large Diffusion Language Models Vector quantized diffusion model for text-to-image synthesis
Reference 105
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1b7c86df-7425-4b28-8daa-40d4aa28a340 · inbound
Show-o2: Improved Native Unified Multimodal Models MMaDA: Multimodal Large Diffusion Language Models
Reference 135
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ab559257-454d-43bf-8b9b-c0bcee9e46ca · inbound
GIFT: Guided Importance-Aware Fine-Tuning for Diffusion Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation fef952f2-3797-49a6-a50f-cf81e24ba9ed · inbound
Discrete Guidance Matching: Exact Guidance for Discrete Flow Matching MMaDA: Multimodal Large Diffusion Language Models
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation cad80625-c2a5-4ed7-ad51-3222285652f4 · inbound
Motus: A Unified Latent Action World Model MMaDA: Multimodal Large Diffusion Language Models
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1dd19a22-ab81-4af8-b0e9-42b4d1a2685c · inbound
Sparse-LaViDa: Sparse Multimodal Discrete Diffusion Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32197b4c-006c-4b17-ad27-5c150d8fc6dd · inbound
Efficient-DLM: From Autoregressive to Diffusion Language Models, and Beyond in Speed MMaDA: Multimodal Large Diffusion Language Models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 56a7a3e3-ce11-432b-839d-baaca9810c21 · inbound
ViewMask-1-to-3: Multi-View Consistent Image Generation via Multimodal Discrete Diffusion Models MMaDA: Multimodal Large Diffusion Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df090975-d48e-4dbf-acf9-6cc62755331b · inbound
dMLLM-TTS: Self-Verified and Efficient Test-Time Scaling for Diffusion Multi-Modal Large Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 13f313ab-603b-4cef-b83e-0c3c78dcc9b2 · inbound
The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c3f005f-105c-4236-a0e2-1d02d5f23d12 · inbound
ChatUMM: Robust Context Tracking for Conversational Interleaved Generation MMaDA: Multimodal Large Diffusion Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52ce73a6-7591-4de5-a891-63e8b6af6cd8 · inbound
Improving Sampling for Masked Diffusion Models via Information Gain MMaDA: Multimodal Large Diffusion Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 11c44097-da16-4beb-934e-295162996ef0 · inbound
Improving Full Waveform Inversion in Large Model Era MMaDA: Multimodal Large Diffusion Language Models
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5160d785-f484-4f58-a685-23dfa6d5bbcd · inbound
Omni-Diffusion: Unified Multimodal Understanding and Generation with Masked Discrete Diffusion MMaDA: Multimodal Large Diffusion Language Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43ecc968-9484-4cf2-85df-ffd568c4061b · inbound
Reinforcement Learning for Diffusion LLMs with Entropy-Guided Step Selection and Stepwise Advantages MMaDA: Multimodal Large Diffusion Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4595ef8c-3fe9-434d-84a8-0f7ffbe37779 · inbound
Thinking Diffusion: Penalize and Guide Visual-Grounded Reasoning in Diffusion Multimodal Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 12d2da17-a5d3-448c-9437-594b4866016d · inbound
DMax: Aggressive Parallel Decoding for dLLMs MMaDA: Multimodal Large Diffusion Language Models
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ecbb25a0-6515-4899-ae16-743bcb6f6f51 · inbound
DMax: Aggressive Parallel Decoding for dLLMs MMaDA: Multimodal Large Diffusion Language Models
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 70f94ca4-0977-44d6-ae4c-89a99b862690 · inbound
Learning Preference-Based Objectives from Clinical Narratives for Dynamic Sepsis Treatment MMaDA: Multimodal Large Diffusion Language Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0087f75-aa6b-41d7-add5-d7a9910619cd · inbound
TorchUMM: A Unified Multimodal Model Codebase for Evaluation, Analysis, and Post-training MMaDA: Multimodal Large Diffusion Language Models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8128049b-d782-4112-a1c9-3daee9e4283c · inbound
TorchUMM: A Unified Multimodal Model Codebase for Evaluation, Analysis, and Post-training MMaDA: Multimodal Large Diffusion Language Models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation abe51ad2-db00-45d6-8aee-0aaa5fbd0530 · inbound
BARD: Bridging AutoRegressive and Diffusion Vision-Language Models Via Highly Efficient Progressive Block Merging and Stage-Wise Distillation MMaDA: Multimodal Large Diffusion Language Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 241ce839-64e1-4fc9-bf7c-a547a7d67a7d · inbound
Stability-Weighted Decoding for Diffusion Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9f9a43b9-bc4d-49ff-9e33-9b4302fa9024 · inbound
UniGenDet: A Unified Generative-Discriminative Framework for Co-Evolutionary Image Generation and Generated Image Detection MMaDA: Multimodal Large Diffusion Language Models
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation bf864c24-1876-4b1e-a4ce-c555a257bbdd · inbound
dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model MMaDA: Multimodal Large Diffusion Language Models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation de4bc6a6-a86c-4a91-a960-b3f01ce383a9 · inbound
Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation MMaDA: Multimodal Large Diffusion Language Models
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9e3482fc-0701-46c0-8218-8693526b6cbe · inbound
Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation MMaDA: Multimodal Large Diffusion Language Models
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a9b9f25b-f62e-4c85-91d7-af8e98de2a9a · inbound
Why Does Grounding Hurt Medical VQA? Benchmarking, Diagnosis, and Fine-Tuning of Vision-Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation afa1081a-abf3-4e4c-8977-380af283f1dc · inbound
Why Does Grounding Hurt Medical VQA? Benchmarking, Diagnosis, and Fine-Tuning of Vision-Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22896690-8163-4db9-82d3-8bc8dd727ee6 · inbound
Break the Block: Dynamic-size Reasoning Blocks for Diffusion Large Language Models via Monotonic Entropy Descent with Reinforcement Learning MMaDA: Multimodal Large Diffusion Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5ee87dd0-caaf-4517-9c38-3bbcb043155c · inbound
Break the Block: Dynamic-size Reasoning Blocks for Diffusion Large Language Models via Monotonic Entropy Descent with Reinforcement Learning MMaDA: Multimodal Large Diffusion Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b100ee1d-9e9c-4680-8b07-6c8ec67f1723 · inbound
NoiseRater: Meta-Learned Noise Valuation for Diffusion Model Training MMaDA: Multimodal Large Diffusion Language Models
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 538becbd-efb7-4b90-8fc1-932698fa52a8 · inbound
Discrete Langevin-Inspired Posterior Sampling MMaDA: Multimodal Large Diffusion Language Models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation dc424ff0-e64d-469c-81c9-73bb81583ce7 · inbound
Relative Score Policy Optimization for Diffusion Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b9d10b12-947a-46c1-8d4c-826d241f0725 · inbound
UniPath: Adaptive Coordination of Understanding and Generation for Unified Multimodal Reasoning MMaDA: Multimodal Large Diffusion Language Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 869fb5ae-7717-4b82-9fe7-66b12b308dec · inbound
Block-R1: Rethinking the Role of Block Size in Multi-domain Reinforcement Learning for Diffusion Large Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4793b85f-c456-4ab3-bd5b-0d9c5a154f2f · inbound
Block-R1: Rethinking the Role of Block Size in Multi-domain Reinforcement Learning for Diffusion Large Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation cbf8dbdc-c4e7-4fc5-9623-f054887d5dbf · inbound
Language Generation as Optimal Control: Closed-Loop Diffusion in Latent Control Space MMaDA: Multimodal Large Diffusion Language Models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e28f85a9-b768-42a9-bc43-e4bb70a80e15 · inbound
Language Generation as Optimal Control: Closed-Loop Diffusion in Latent Control Space MMaDA: Multimodal Large Diffusion Language Models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1a888b34-b3c8-48f3-a060-0a77a183a5a0 · inbound
Language Generation as Optimal Control: Closed-Loop Diffusion in Latent Control Space MMaDA: Multimodal Large Diffusion Language Models
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a88484b6-616b-40f8-bb85-b1c46bcb2aac · inbound
Roll Out and Roll Back: Diffusion LLMs are Their Own Efficiency Teachers MMaDA: Multimodal Large Diffusion Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 358a728a-f7ca-42f1-abd3-bc76902c247c · inbound
Fast-dDrive: Efficient Block-Diffusion VLM for Autonomous Driving MMaDA: Multimodal Large Diffusion Language Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e667c714-65ec-4adf-a8d0-f821f5f87e9b · inbound
Fast-dDrive: Efficient Block-Diffusion VLM for Autonomous Driving MMaDA: Multimodal Large Diffusion Language Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 086e7709-73b4-463a-b55b-8fdc273f2be7 · inbound
Visual-Redundancy-Controlled Parallel Decoding for Diffusion-Based Multimodal Large Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 600fe1e3-ad13-4431-b520-e22a69821dea · inbound
Guidance Contrastive Token Credit Assignment for Discrete Policy Optimization MMaDA: Multimodal Large Diffusion Language Models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a827c1ca-a1b2-4d91-ada0-e3a5a414332c · inbound
GDSD: Reinforcement Learning as Guided Denoiser Self-Distillation for Diffusion Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6939d29e-488c-4439-aea2-ed4237fc773b · inbound
dMoE: dLLMs with Learnable Block Experts MMaDA: Multimodal Large Diffusion Language Models
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7f5534d9-0081-4af6-b3be-3e7c875ee5bd · inbound
Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models MMaDA: Multimodal Large Diffusion Language Models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b0f2c5cc-9191-4ff2-8abc-c9f2a8921161 · inbound
SimSD: Simple Speculative Decoding in Diffusion Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 88613756-b150-49be-8009-094b86058f00 · inbound
MaskForge: Structure-Aware Adaptive Attacks for Jailbreaking Diffusion Large Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation aa8649db-da31-4f93-b55e-85eee930f019 · inbound
UniCanvas: A Diffusion-base Unified Model for Text-in-Image Joint Generation MMaDA: Multimodal Large Diffusion Language Models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 778bbb38-14fc-4d4c-b762-eec8e2a7b276 · inbound
Dynamic Infilling Anchors for Format-Constrained Generation in Diffusion Large Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 010827f9-95ea-4d51-816e-476b412c26f5 · inbound
Back on Track: Aligning Rewards and States for Reasoning in Diffusion Large Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 60b3ec82-4b0b-4ddc-baa1-f7972b607dd4 · inbound
TimeROME-DLM: Temporal Causal Tracing and Low-Rank Inference-Time Knowledge Editing for Masked Diffusion Language Models MMaDA: Multimodal Large Diffusion Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a0b40dd6-1ac5-4f2c-a8c8-aa80f431ff51 · inbound
InterleaveThinker: Reinforcing Agentic Interleaved Generation MMaDA: Multimodal Large Diffusion Language Models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 95866b58-e1d8-4d40-9303-159209c7b898 · inbound
DiPOD: Diffusion Policy Optimization without Drifting Apart MMaDA: Multimodal Large Diffusion Language Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0a64f099-d9aa-4a2c-810e-8af79b3dfc97 · inbound
UniTranslator: A Unified Multi-modal Framework for End-to-end In-Image Machine Translation MMaDA: Multimodal Large Diffusion Language Models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation cd5ff93b-0488-4833-8a04-b6b31e6082d8 · inbound
Improved Large Language Diffusion Models MMaDA: Multimodal Large Diffusion Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 595926b4-c45a-4e9c-8f11-1313fd0b44cd · inbound
LithoDreamer: A Physics-Informed World Model for Multi-Stage Computational Lithography MMaDA: Multimodal Large Diffusion Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c2d8373b-64d8-4ad0-aadf-7063bd4454b7 · inbound
Nemotron-Labs-Diffusion-Image: Advancing Masked Discrete Diffusion for High-Resolution Image Synthesis MMaDA: Multimodal Large Diffusion Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 78dcc2da-fea6-4cd4-89cc-8bbba6b9a481 · inbound
Nemotron-Labs-Diffusion-Image: Advancing Masked Discrete Diffusion for High-Resolution Image Synthesis MMaDA: Multimodal Large Diffusion Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbc51968-a12b-4d36-b0b9-dd53948f64b0 · inbound
Illuminating Unified Multimodal Model for Free-form Interleaved Text-Image Generation MMaDA: Multimodal Large Diffusion Language Models
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a4e419d1-7ebb-46d6-bc5c-eb5a0328f9a7 · inbound
Bridging Video Understanding and Generation in a Unified Framework MMaDA: Multimodal Large Diffusion Language Models
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation dfd29b72-a0f2-406a-9f9b-dc0bbda9c39c · inbound
TACG: Trajectory-Aware Commit Gating for Diffusion Language Model Decoding MMaDA: Multimodal Large Diffusion Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f547940-83d8-426b-85f1-49597944d867 · inbound
Transferability Between Understanding and Generation in Unified Multimodal Models MMaDA: Multimodal Large Diffusion Language Models
Reference 102
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a28d3a0d-2935-4b6f-9904-97d996fbf086 · inbound
Nemotron-Labs-Diffusion: A Tri-Mode Language Model Unifying Autoregressive, Diffusion, and Self-Speculation Decoding MMaDA: Multimodal Large Diffusion Language Models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 04cb6e76-ea97-4900-bd0b-80bd18c98368 · inbound
Discrete Diffusion Models: A Unified Framework from Tokenization to Generation MMaDA: Multimodal Large Diffusion Language Models
Reference 190
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43a35c33-fa2b-4b4e-9ef5-bb071eb4c73a · inbound
Seeing the End at Step Zero: Accelerating Diffusion MLLMs via MLP Sparsity-Aware Truncation MMaDA: Multimodal Large Diffusion Language Models
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47df349c-489c-436f-aec7-59f53a5cbe36 · inbound
ST-Veto: Spatio-Temporal Token Veto for Diffusion MLLMs via Taylor Prediction and Visual Grounding MMaDA: Multimodal Large Diffusion Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 231c1717-8a8d-452d-8840-1cedb5534a6d · inbound
Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation MMaDA: Multimodal Large Diffusion Language Models
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.