Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 59 inbound Pith citation observations for arXiv:2301.00704.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:39:24.622157Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
119
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 9232872e-81f2-4cf9-80d4-dadf982e014b · inbound
Scaling Robot Learning with Semantically Imagined Experience Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0f796b13-7c6f-4551-8c62-4e22c30ac520 · inbound
DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 129
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a694b0ae-382a-41c6-b127-2f87bba00510 · inbound
Finite Scalar Quantization: VQ-VAE Made Simple Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c77dcd03-48dd-492f-9de5-fdf451d90738 · inbound
Learning Interactive Real-World Simulators Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 173
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a07e5ce6-1751-4c76-9b09-d8f18e2bd385 · inbound
VideoCrafter1: Open Diffusion Models for High-Quality Video Generation Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c40f8024-858c-4265-a131-657112c79235 · inbound
VideoPoet: A Large Language Model for Zero-Shot Video Generation Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5b3eb7f1-a16f-4980-8df3-66b9a26b42c7 · inbound
Controllable Image Generation with Composed Parallel Token Prediction Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 53a8d595-7b89-4e86-9456-b2e176e8dd26 · inbound
Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 131f433b-66db-4afd-a722-86e083c996f1 · inbound
Show-o: One Single Transformer to Unify Multimodal Understanding and Generation Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 80dc1c22-9878-45d2-9c2a-0b0e873d3836 · inbound
Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 66e0d9ea-dd01-44ae-baca-dd2d23af5fab · inbound
Autoregressive Video Generation without Vector Quantization Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e4c62dd1-97f8-43f1-b85b-6c6e62cec95a · inbound
Large Language Diffusion Models Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2a1143d1-7a46-454e-9f27-222bada2c9f2 · inbound
MSDformer: Multi-scale Discrete Transformer For Time Series Generation Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 25db7335-c11d-44ce-a2bb-695de5c292dd · inbound
MMaDA: Multimodal Large Diffusion Language Models Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 107
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e77c68c5-45da-461b-aeba-416f1234010b · inbound
SpecMaskFoley: Steering Pretrained Spectral Masked Generative Transformer Toward Synchronized Video-to-audio Synthesis via ControlNet Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e99710c-9d77-427f-8763-c4e96e5bdcd7 · inbound
LaViDa: A Large Diffusion Language Model for Multimodal Understanding Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f51b3af5-c512-4de2-977c-a437fe25ccf4 · inbound
Semantics-Aware Human Motion Generation from Audio Instructions Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20a3d936-6f55-46db-9e63-6e5bf523bf11 · inbound
IMPACT: Iterative Mask-based Parallel Decoding for Text-to-Audio Generation with Diffusion Modeling Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36887104-8b08-42f6-9869-9f24112971e3 · inbound
Native-Resolution Image Synthesis Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b2b60c7-2598-43c8-ad79-1de3049e13ba · inbound
HMAR: Efficient Hierarchical Masked Auto-Regressive Image Generation Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3712f10-5504-470b-a3ff-dff63c008640 · inbound
Noise Consistency Regularization for Improved Subject-Driven Image Synthesis Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9c346ec-e6fd-49ab-adf8-2a271f52af36 · inbound
MapBERT: Bitwise Masked Modeling for Real-Time Semantic Mapping Generation Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff38e35f-58a8-4028-8e43-25e4f3c590c4 · inbound
Output Scaling: YingLong-Delayed Chain of Thought in a Large Pretrained Time Series Forecasting Model Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93898f86-93a1-4007-bbe9-fd4ddcd0c70e · inbound
MARch\'e: Fast Masked Autoregressive Image Generation with Cache-Aware Attention Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f962a37-9a2e-400b-9a7d-5a0dbce28b9d · inbound
Learning golf swing signatures from a single wrist-worn inertial sensor Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8f6cd43-d7dc-455d-a05e-11b2a34633df · inbound
MVGBench: Comprehensive Benchmark for Multi-view Generation Models Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53d434c6-451e-4728-bb7a-2559e37af9e4 · inbound
Is Visual in-Context Learning for Compositional Medical Tasks within Reach? Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc6d1286-0035-4aa1-9ff1-5e25eab8a452 · inbound
CI-VID: A Coherent Interleaved Text-Video Dataset Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0757d32b-5882-4aac-a7ab-e83195aef7f6 · inbound
DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed205e6c-4d82-4db4-bc82-fb3e72b207a6 · inbound
Room Impulse Response Generation Conditioned on Acoustic Parameters Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07764351-64d3-4d07-9710-39ec82cf5df2 · inbound
MADI: Masking-Augmented Diffusion with Inference-Time Scaling for Visual Editing Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3469a9a0-887f-4059-ad72-6d7dc63615e5 · inbound
Learning neuro-symbolic convergent term rewriting systems Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7235f01d-ba09-49de-a38a-6496680f10d2 · inbound
Zero-Residual Concept Erasure via Progressive Alignment in Text-to-Image Model Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0233a12-2300-4dcf-8ad9-faf649f90181 · inbound
FBI: Learning Dexterous In-hand Manipulation with Dynamic Visuotactile Shortcut Policy Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03f319a0-c123-4471-9679-1616fb66726a · inbound
Layout-Conditioned Autoregressive Text-to-Image Generation via Structured Masking Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3eeabf2-40d5-4655-b1f5-f0bdfb7aee7c · inbound
Lavida-O: Elastic Large Masked Diffusion Models for Unified Multimodal Understanding and Generation Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b9524c7-760c-40a4-b39b-46c5de90dd0f · inbound
IAR2: Improving Autoregressive Visual Generation with Semantic-Detail Associated Token Prediction Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71063a29-1ae8-45a6-87fa-202e586ea3d0 · inbound
RubricRL: Simple Generalizable Rewards for Text-to-Image Generation Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab85a340-a5c4-4404-a48d-c2a01f7d38e5 · inbound
Sparse-LaViDa: Sparse Multimodal Discrete Diffusion Language Models Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e823a537-1fdc-4c00-bb45-09ecb3b07f47 · inbound
The Script is All You Need: An Agentic Framework for Long-Horizon Dialogue-to-Cinematic Video Generation Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8362087-437c-4cbf-a03b-83b1d4e73e9b · inbound
Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a83873ff-e267-4133-8856-ff170a96f15b · inbound
Unifying Contrastive and Generative Objectives for Visual Understanding and Text-to-Image Generation Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ef4784c5-9a87-4b1d-9891-5fb6a6b4e38f · inbound
Controllable Image Generation with Composed Parallel Token Prediction Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 14212172-9d6b-4810-9261-eae613d889da · inbound
UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7d97494e-1a26-412c-a643-e19525e82873 · inbound
UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 83ed988c-ef11-45c2-a602-6aad93d637dc · inbound
Coupling Models for One-Step Discrete Generation Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9d6ce223-0082-4686-ae56-d033d04f4f12 · inbound
VAGS: Velocity Adaptive Guidance Scale for Image Editing and Generation Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 609ca211-2bb8-4d53-b578-5bbb2630e701 · inbound
Dimension-Free Convergence of Discrete Diffusion Models: Adjoint Equations Induce the Right Space Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 04ea7e06-9606-4226-9f00-13106dcbd614 · inbound
Dimension-Free Convergence of Discrete Diffusion Models: Adjoint Equations Induce the Right Space Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 281bbec6-d233-4b26-9316-e46384d8f824 · inbound
Dimension-Free Convergence of Discrete Diffusion Models: Adjoint Equations Induce the Right Space Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc82891e-2ef2-4cef-886d-9807b393be94 · inbound
Going PLACES: Participatory Localized Red Teaming for Text-to-Image Safety in the Global South Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 41295a1b-dadb-4525-8144-73c64613c3ba · inbound
SplitAvatar: One-shot Head Avatar with Autoregressive Gaussian Splitting Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 98eb01a1-65d2-4d36-b016-935ee963d365 · inbound
Diffusing in the Right Space: A Systematic Study of Latent Diffusability Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b7772da1-d62b-4175-b2b5-9909ce03c38b · inbound
What Type of Inference is Active Inference? Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 05d80a24-c1eb-495f-8d8c-cf224c716547 · inbound
ARM: An AutoRegressive Large Multimodal Model with Unified Discrete Representations Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a832f26-c88c-40e4-81b6-d31f4a969b93 · inbound
Expected Free Energy-based Planning as Variational Inference Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4e6b0e0c-9da9-4d51-bef2-c9129619da42 · inbound
Co-occurring associated retained concepts in Diffusion Unlearning Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7d4417cc-c09e-40ff-9e0e-29344603c31c · inbound
FlexiSLM: A Dynamic and Controllable Frame Rate Spoken Language Model Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 153
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bda92eb2-a1ff-4c83-8505-db6d71b2ede1 · inbound
MentalThink: Shaping Thoughts in Mental SVG World Muse: Text-To-Image Generation via Masked Generative Transformers
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.