Pith. sign in

Paper Citation Record · LEDGER

MMGen: Unified Multi-modal Image Generation and Understanding in One Go

As of 24 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2503.20644.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.20644 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:25:22.539339Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-07T14:33:54.337921Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a9fa7bfd-06dc-4760-afd8-4232533b08e8 · inbound

Jodi: Unification of Visual Generation and Understanding via Joint Modeling cites this paper.

Jodi: Unification of Visual Generation and Understanding via Joint Modeling MMGen: Unified Multi-modal Image Generation and Understanding in One Go

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:22.539339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:22.539339Z digest=sha256:d65ea7c359e2ba6e1c6400f096697648c5d9c5c14081fdbcd309e3096bd89a0c

Observation e34ac4cf-7c1b-4b97-8151-a501a01582d2 · inbound

CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos cites this paper.

CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos MMGen: Unified Multi-modal Image Generation and Understanding in One Go

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-16T13:47:57.548748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-16T13:43:26.460480Z digest=sha256:8d4b5947b33550616d710f626264d2fb44e67758fb4c8df9c34885dab35361d7

Observation 80b00ac2-d3b6-4271-9e71-96f9dd102d15 · inbound

PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space cites this paper.

PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space MMGen: Unified Multi-modal Image Generation and Understanding in One Go

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-07-07T14:33:54.339211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-07-07T14:29:23.795094Z digest=sha256:3f1473107fcbe649e5657d7f8c6319a346e42adb8cf9a42c3ef5c92278e03d58

Observation b9a71ed7-efba-496f-bfd5-792cbdb5e584 · inbound

MMOE: Modernizing Diffusion Transformers with Efficient Expert Design cites this paper.

MMOE: Modernizing Diffusion Transformers with Efficient Expert Design MMGen: Unified Multi-modal Image Generation and Understanding in One Go

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-31T08:31:50.142918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T08:31:50.142918Z digest=sha256:4d882a7cc455b05e683935f1e7e74dbb985f740d5ff0d654ab0a3cc13102d61c