Pith. sign in

Paper Citation Record · LEDGER

MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2407.07614.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.07614 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T11:37:10.964681Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T21:27:24.441164Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c71919a8-492b-4e79-9bb8-f368a347f265 · inbound

$\pi_0$: A Vision-Language-Action Flow Model for General Robot Control cites this paper.

$\pi_0$: A Vision-Language-Action Flow Model for General Robot Control MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:38:24.472611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T12:38:24.425784Z digest=sha256:ff5eda0e6434a0dca477383e5d96bc210e368f4071177c48b75fed5ab9c4bd4d

Observation 4a6574d1-2d8d-4be1-8f5a-97665eaab9a7 · inbound

Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models cites this paper.

Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T02:48:44.995289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T02:48:44.900467Z digest=sha256:cee3fb56dbb42d5defcad4af7196b84ce6af3c5838d6b14043b728a5907b0ec8

Observation 61fe701b-d584-442e-a2fd-5a71f2a9ff36 · inbound

T2I-FactualBench: Benchmarking the Factuality of Text-to-Image Models with Knowledge-Intensive Concepts cites this paper.

T2I-FactualBench: Benchmarking the Factuality of Text-to-Image Models with Knowledge-Intensive Concepts MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:02:43.319764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-23T08:00:12.781392Z digest=sha256:87eeb6acd0edb041fb2c59ec100bc36fdc413bf14b10bf2a4be0dc53fb36a9ea

Observation 80703a0e-fd71-4e46-9967-6a9dceae9847 · inbound

LMFusion: Adapting Pretrained Language Models for Multimodal Generation cites this paper.

LMFusion: Adapting Pretrained Language Models for Multimodal Generation MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-11T11:37:10.964681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:37:10.964681Z digest=sha256:6fb3ba7be91dccbf44ea7f3829afbc93e48c27014fd93e7189e63396ee987987

Observation 779a674e-f8fa-4631-9a2d-5bf3291d5568 · inbound

Text2Earth: Unlocking Text-driven Remote Sensing Image Generation with a Global-Scale Dataset and a Foundation Model cites this paper.

Text2Earth: Unlocking Text-driven Remote Sensing Image Generation with a Global-Scale Dataset and a Foundation Model MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T22:44:30.545433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:44:30.545433Z digest=sha256:2c632523e1dae9452554d599791c361d28e8dd2e8411e9e8284357625cfc53b3

Observation 16726869-fae9-4895-8147-a60da36e5519 · inbound

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity cites this paper.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.681892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.681892Z digest=sha256:2710c63fbf50f95352c269ba4a64d509647dafded0c1ef9745917b7f446df68a

Observation fbc237b3-b337-479e-8e9a-c9f999690795 · inbound

A Comprehensive Study of Decoder-Only LLMs for Text-to-Image Generation cites this paper.

A Comprehensive Study of Decoder-Only LLMs for Text-to-Image Generation MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:20:51.042386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:20:51.042386Z digest=sha256:c5e2f8262ba5d40178a7802569d5e6879118a2b6238148f4182ecf17e30b7721

Observation 2bd4c409-f80f-49a6-a54a-babef291fd0a · inbound

Discrete Noise Inversion for Next-scale Autoregressive Text-based Image Editing cites this paper.

Discrete Noise Inversion for Next-scale Autoregressive Text-based Image Editing MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T12:07:32.725814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:07:32.725814Z digest=sha256:e77d2db8cfb996c9ca756c6e5c500857e689545cba1cb55465fd34ef49377fe6

Observation df9c6714-1456-44b4-a6fd-4d87ef5e18a3 · inbound

Compositional Text-to-Image Generation Via Region-aware Bimodal Direct Preference Optimization cites this paper.

Compositional Text-to-Image Generation Via Region-aware Bimodal Direct Preference Optimization MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:23:28.310544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T13:15:24.299457Z digest=sha256:697c47906698a317358ad87ac68c91976f01ae9d7d7db282a869533304d7c205

Observation ea719a7f-f3e4-479f-8440-f7cf414ba219 · inbound

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling cites this paper.

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:27:24.442554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-27T19:46:43.514413Z digest=sha256:c64bafa3a935df199069853223163bc0d1cf37ab603a3923b6f0d5c29fd5b918