Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2309.02591.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:50:00.168507Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
27
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 739d5dad-b6b6-47d5-99ac-9a6729b8ee4d · inbound
Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 214
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2e63b695-acb3-402b-8270-6b1634298980 · inbound
SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a157ff30-ed01-4622-8d76-af1ffb73c384 · inbound
Chameleon: Mixed-Modal Early-Fusion Foundation Models Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d85a6cb4-c31f-454c-a2d5-09905940a6cd · inbound
PaliGemma: A versatile 3B VLM for transfer Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 157
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3cdb2fec-7473-4b18-a944-37d7bf23c709 · inbound
Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 50606391-af01-487d-a526-855b56149c58 · inbound
VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7554d3b9-a55c-47e6-babf-3ee709a2ed4a · inbound
Emu3: Next-Token Prediction is All You Need Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d99ca5f-b621-4cb0-8da2-388754be6df9 · inbound
Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a911f41a-9533-4efb-ade8-38a6107d7050 · inbound
DualToken: Towards Unifying Visual Understanding and Generation with Dual Visual Vocabularies Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1f6f407e-d188-444d-b36f-0b6e4a9b9b6d · inbound
Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0349970a-a02d-4bfb-b826-ce4611bc89f9 · inbound
AR-RAG: Autoregressive Retrieval Augmentation for Image Generation Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cddd270a-1f4d-403d-a095-4b938064210e · inbound
CuRe: Cultural Gaps in the Long Tail of Text-to-Image Systems Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9b4edd9-af5c-4184-8228-db135bf16a19 · inbound
Pisces: An Auto-regressive Foundation Model for Image Understanding and Generation Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e051434-6923-46ac-b071-f1409063ba81 · inbound
Transition Matching: Scalable and Flexible Generative Modeling Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3d45e71-80cd-47f3-bbb0-101445996b39 · inbound
Beyond Patches: Global-aware Autoregressive Model for Multimodal Few-Shot Font Generation Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c1c71beb-6d0d-4eec-b642-46710fa28895 · inbound
Mirai: Autoregressive Visual Generation Needs Foresight Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a8c6eaa1-d52b-47f3-9230-edd866b78180 · inbound
Shape of Thought: Progressive Object Assembly via Visual Chain-of-Thought Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c18d0c1-30cf-445e-9666-814e7aa69a2a · inbound
Token by Token, Compromised: Backdoor Vulnerabilities in Unified Autoregressive Models Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eb19b34b-f87e-48c9-9918-9b17883db5ca · inbound
FullFlow: Upgrading Text-to-Image Flow Matching Models for Bidirectional Vision--Language Generation Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7533d770-7924-41f1-a666-d1af27ffe8ff · inbound
InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 198
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c5eac19c-a2da-47f4-9386-824529c178d0 · inbound
SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b62c772c-45f8-4b51-b1fc-f3b5fa752002 · inbound
SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 80d14c5e-5aca-41bd-bf26-dc472960ebc6 · inbound
Obliviate: Erasing Concepts from Autoregressive Image Generation Models Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a6b91e9c-04c0-486f-8eba-a1f81998cbda · inbound
HELP: Human-Efficient Large-Scale Robot Post-Training with Rollout Segmentation Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a7d5a16-0de4-403a-bb72-3732166dc2cb · inbound
HELP: Human-Efficient Large-Scale Robot Post-Training with Rollout Segmentation Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18ce830d-80e2-4c80-87b4-e5957605ca13 · inbound
Instruction-based Image Editing: A Survey on Data, Models, Evaluation, and Applications Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.