Pith. sign in

Paper Citation Record · LEDGER

Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2406.18583.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.18583 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T14:26:46.901368Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T00:04:22.362322Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5afe3a67-b09f-4620-80e0-6f5abf963128 · inbound

Emu3: Next-Token Prediction is All You Need cites this paper.

Emu3: Next-Token Prediction is All You Need Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 110

Resolution
malformed identifier
arxiv_id, observed 2026-05-11T10:56:09.596195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T10:56:06.418360Z digest=sha256:d7d01832795a11c2e626134ed2231815a88745e4f1f5c75a7051c0aac0f8cacc

Observation d65585ac-45e1-4bef-83cb-8bc6d1032b8e · inbound

SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers cites this paper.

SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:56:50.054647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T00:56:50.009149Z digest=sha256:86d93205334ac8fc75e5c9fcc6d1e932b4e1d2880735346d72677d67bf954a42

Observation 6dcbc11a-5b14-4933-8af5-0903233ddb4b · inbound

Autoregressive Video Generation without Vector Quantization cites this paper.

Autoregressive Video Generation without Vector Quantization Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T15:07:39.842642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T15:07:39.718555Z digest=sha256:9fbec2da01515512ca52d462548908642140a616257271f7a931344aa8b619db

Observation 56ef134c-42d5-4f56-a0f2-add6fdb7e4a1 · inbound

Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling cites this paper.

Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:14:52.933348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T08:14:52.890145Z digest=sha256:0a11317dca38c7710b2cff8889dfc353894e34065f24a358b01654bfc7f70618

Observation 54140678-5bd4-4a29-9ec2-87b6799cfe9b · inbound

Lumina-Video: Efficient and Flexible Video Generation with Multi-scale Next-DiT cites this paper.

Lumina-Video: Efficient and Flexible Video Generation with Multi-scale Next-DiT Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-08T14:26:46.901368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:26:46.901368Z digest=sha256:ef450248250e73c2372599ec2f395c42d383a8fad4498187c5cd4618309489f0

Observation d64f56ad-b319-4dbd-9e1b-6bf3a04438fd · inbound

RectifiedHR: Enable Efficient High-Resolution Synthesis via Energy Rectification cites this paper.

RectifiedHR: Enable Efficient High-Resolution Synthesis via Energy Rectification Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-23T01:35:19.417810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T01:33:49.962427Z digest=sha256:2d366f7d75acbe7ed0dadcd42daf626f203a7817a0a29d06c343bc5553133e08

Observation 7faa02bc-a650-4fd8-98be-3d7b2c618b25 · inbound

Transfer between Modalities with MetaQueries cites this paper.

Transfer between Modalities with MetaQueries Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:49:23.134684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T22:49:23.074271Z digest=sha256:c49a46666525e24da4bdc97725f9601ed3d1b1df29ceee5fe67d446e10138d63

Observation 4ead44de-aedb-43fe-a330-35bcce928279 · inbound

Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation cites this paper.

Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-17T07:24:04.532924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T07:24:04.460276Z digest=sha256:6d344ea89b1b0a3f89beeff7d60941bbeb823eeb666c32d1d3d9813e0a8f7fbe

Observation f400f3fe-1e13-4927-8cfd-aeba81746a5c · inbound

BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset cites this paper.

BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:34:26.968355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T23:34:26.878354Z digest=sha256:f17587415d0064a21a9de0caca08d7c484daaf9fd2e1683b9dd893bdc8624914

Observation bb5c5c5f-b224-4f4c-8242-18951b4f5ea0 · inbound

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities cites this paper.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 128

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:06.585905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:06.585905Z digest=sha256:0be26ab4927d78ac4d9ae2cfe7c7ef3dd88c6479716c7c743beea108e0f56de5

Observation f46da92d-f9b5-4608-99af-db1e4c6520d2 · inbound

OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation cites this paper.

OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:16.743577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:44:16.743577Z digest=sha256:0069931809778204143c66775881c0682789ab35337a8229f6be931870c948f4

Observation a4fe494c-7197-49e9-8be6-951aa0ace63f · inbound

Synthesis of discrete-continuous quantum circuits with multimodal diffusion models cites this paper.

Synthesis of discrete-continuous quantum circuits with multimodal diffusion models Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:42:15.931213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T11:39:14.375702Z digest=sha256:35564084efd160d6e926ff1b6e7ee914224c59f3f11218776b945efda7439408

Observation 89787206-df16-40ae-a48d-3588a67b81ed · inbound

RelationAdapter: Learning and Transferring Visual Relation with Diffusion Transformers cites this paper.

RelationAdapter: Learning and Transferring Visual Relation with Diffusion Transformers Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:30.614880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:30.614880Z digest=sha256:23430dd7a6b42ea0419c078073e934a854a5be6a2262037ff6831ba59df4f2d4

Observation ae66d45d-f3f5-45a0-91f6-994ca9ad6a26 · inbound

A Comprehensive Study of Decoder-Only LLMs for Text-to-Image Generation cites this paper.

A Comprehensive Study of Decoder-Only LLMs for Text-to-Image Generation Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T05:20:51.240632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:20:51.240632Z digest=sha256:14c5d6281ccb3d194910e130a805c346a92faa98a8c4b518e1f8fbb32390674f

Observation 6e8b3b05-e52d-4294-b1d1-4cad0a13b504 · inbound

ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation cites this paper.

ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T23:28:09.448639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:28:09.448639Z digest=sha256:04de6479cf89f0b99f83a0a853d030f8a0270c7140070aa7ee7ccd3a197cc573

Observation 96bb966d-d105-44db-a794-0efa926f9679 · inbound

PixelDiT: Pixel Diffusion Transformers for Image Generation cites this paper.

PixelDiT: Pixel Diffusion Transformers for Image Generation Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:31:31.212488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T04:30:07.417197Z digest=sha256:b231be7f12197eda1465c027f51e00b243ac0f9cb42d73d758c56f979ff12c8c

Observation 10ebfc64-4b09-49e2-bad3-568454c0fe8c · inbound

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices cites this paper.

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T20:03:43.892765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T20:00:27.987481Z digest=sha256:30f5e7522425c46ab6d3380b47d49222a38a797bef39a5bc7037b7f90032930d

Observation 82e8e2fd-508b-4351-9a9d-f005f652d3ce · inbound

MONET: A Massive, Open, Non-redundant and Enriched Text-to-image dataset cites this paper.

MONET: A Massive, Open, Non-redundant and Enriched Text-to-image dataset Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 114

Resolution
malformed identifier
arxiv_id, observed 2026-05-21T05:23:58.444159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T05:21:18.369534Z digest=sha256:7140d3432ea61cb298012cab61aa79cb044a4a5d1dcaa86abbfcfaa39f3ebfed

Observation fee82bc0-6dc0-45e1-8278-a456b7da04f0 · inbound

Compositional Text-to-Image Generation Via Region-aware Bimodal Direct Preference Optimization cites this paper.

Compositional Text-to-Image Generation Via Region-aware Bimodal Direct Preference Optimization Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:23:28.299117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T13:15:24.299457Z digest=sha256:a9c16fdc9c1c5b53a4274404260692a769add7ac748a10ee89bfcd8d8163a41f

Observation f89bff8b-6c0d-4f58-9ff3-e05e9ef05ff7 · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 181

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T00:04:22.363642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-07T23:59:38.702609Z digest=sha256:469a05b5563802c2ec15afb49c8f2ab64889fc1d44e971ca54e5abf5a9246061

Observation cfc7d972-a5e8-41bb-ac42-625fdfdae93c · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT

Reference 181

Resolution
unresolved
no resolver link, observed 2026-07-11T07:46:49.059192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T07:46:49.059192Z digest=sha256:22b0237ea51d2f10c725827a61b011c697b0c51a5fa199e935e61116b541b784