Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T13:10:14.308216Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 0 inbound Pith citation observations for arXiv:2606.11096.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T13:10:14.308216Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
75 of 75 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1ab62cca-5d5f-41ad-b3fa-81091aad638d · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 07ad680f-c33f-452d-b693-145cfcb517bc · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Perception encoder: The best visual embeddings are not at the output of the network
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5749fc2b-6f54-41a4-8636-6289b596f651 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Perception encoder: The best visual embeddings are not at the output of the network
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 454d1247-f1ab-4ff3-94dd-f056b3f411d1 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Emerging properties in self-supervised vision transformers
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8beec1cc-4d1c-4b4c-aee1-66d02d65ec5b · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Maskgit: Masked generative image transformer
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d778a25-db20-48c7-afae-87c9e024a33f · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Enhancing Vision Foundation Models via Multimodal Continual Pre-Training
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 22f0a495-3c75-4f05-baaf-050fe6b32d74 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Vision transformers need registers
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b015d138-e0cc-42db-9987-83281921b62b · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder author Dong, W
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 23545e46-2959-463f-8052-510b0feae8c6 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder arXiv preprint arXiv:2511.23386 (2025) 4, 7, 9
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 705bcac2-f4b6-4c43-8c5e-81cbf877b30d · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Taming transformers for high-resolution image synthesis
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96559e00-7569-4b47-a837-4d40fda3ee64 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2720aa33-4bf1-4df2-9ede-0fe84d596420 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder One layer is enough: Adapting pretrained visual encoders for image generation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f0612a31-cdc8-4585-8221-891bffb55bc5 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Long Video Generation with Time-Agnostic VQGAN and Time-Sensitive Transformer
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e38956fe-f930-4997-938f-1ba8d2f9deaf · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Gans trained by a two time-scale update rule converge to a local nash equilibrium
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0b29bcb-567e-4895-8630-319c76e13cbb · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder beta-vae: Learning basic visual concepts with a constrained variational framework
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e56af5a1-2663-49f7-bc14-9c012f43a31f · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder GQA: A New Dataset for Real-World Visual Reasoning and Compositional Question Answering
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c6ba7bee-4500-4af8-926c-d3803b477c5e · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Image-to-imagetranslationwithconditionaladversarial networks
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffcb0b9a-e1d7-4277-bca8-0e8a45bf70f2 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Dino-tok: Adapting dino for visual tokenizers
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 15cac6ff-370f-4925-956d-f63befdbadc3 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder IEEE Trans
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6c2d7fc6-2cd3-4557-9d91-6999b73609ac · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Auto-encoding variational bayes
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61f31447-932c-481b-ab85-22acfc9a264f · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder International Journal of Computer Vision 128(7), 1956–1981 (Mar 2020)
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1eab0458-211a-4198-96b4-33a72c865654 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Improved Precision and Recall Metric for Assessing Generative Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5427997b-3428-4439-85d4-924984e7d9bc · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Autoregressive Image Generation using Residual Quantization
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 35d9ed1b-5a95-462a-9335-da525c8cd3a8 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 280769d0-f1ae-473a-925a-d3bc569611a0 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder ImageFolder: Autoregressive Image Generation with Folded Tokens
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0770a158-b49d-4e95-b517-b1abafc45488 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Evaluating Object Hallucination in Large Vision-Language Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 06c401a8-aacf-4f6c-8e64-a06402f842f2 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Decoupled Weight Decay Regularization
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 91a599a7-5ba8-4feb-a32c-3c19314911bd · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Open-MAGVIT2: An Open-Source Project Toward Democratizing Auto-regressive Visual Generation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bf7e34d4-748f-450b-bd7f-e26c7d52d446 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Unitok: A unified tokenizer for visual generation and understanding.arXiv preprint arXiv:2502.20321, 2025a
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 80798ca6-a46a-486a-895f-f7cd2c7c48be · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 68859522-0407-4f53-af23-8ed249757b50 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Chartqa: A benchmark for question answering about charts with visual and logical reasoning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b5aea5d-75da-4423-9297-ff48dbb72219 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Docvqa: Adatasetforvqaondocumentimages
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d135103-6b18-490a-b525-88df7750b719 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Finite Scalar Quantization: VQ-VAE Made Simple
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c9a3ae1c-d32b-4e07-883c-3941afa5c276 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Dinov2: Learning robust visual features without supervision
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ef515e8-27f2-414e-a768-5e56e8441a70 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Scalable diffusion models with transformers
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 645d9226-af78-4eab-8101-401f697c3ac0 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Learning Transferable Visual Models From Natural Language Supervision
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 85ae22e2-031f-448e-ac44-1642b3a56903 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Generating diverse high-fidelity images with vq-vae-2, 2019
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27012b5f-b880-4d38-acb4-d2ab42fb7377 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Beyond Next-Token: Next-X Prediction for Autoregressive Visual Generation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 04d5aa3a-92b2-4fb7-ba7a-5d72ac9ca1e0 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder High-Resolution Image Synthesis with Latent Diffusion Models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f574d2f9-70b3-400d-924f-340fea066395 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Improved techniques for training gans
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d2b0276-6fbd-476a-b91f-72f1a8142944 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder A-okvqa: A benchmark for visual question answering using world knowledge
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f377163-2b12-469a-aab1-b8df0b0e9e2d · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Latent diffusion model without variational autoencoder
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c34d26b8-8024-4de0-92a9-4a6ca2db5991 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder DINOv3
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dc92ee85-2095-4ae6-957a-2a24db50dc8b · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Towards VQA Models That Can Read
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 710f038f-f07f-482a-9292-0d1a144b9ba3 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Dualtoken: Towards unifying visual understanding and generation with dual visual vocabularies
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2057130f-78dd-48f7-9fb8-d10e9f279a6c · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder RoFormer: Enhanced Transformer with Rotary Position Embedding
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 16b262ae-4a66-49ad-961a-7f3be5a36dd0 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c9faba58-3a7f-4f41-bc13-2a93e00bb89d · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Visual autoregressive modeling: Scalable image generation via next-scale prediction
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 933766f4-ca43-44d4-bb30-019b59d10094 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Cambrian-1: A fully open, vision-centric exploration of multimodal llms
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9479af7c-bb79-4a33-b089-80f4764bba12 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 206f7527-d29f-4e23-9b64-878c060dcb67 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Neural discrete representation learning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef620b3b-092e-4edb-b4a7-b18f0869d23e · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Neural Discrete Representation Learning
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 983ba479-7068-436a-8a45-247638c713ed · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Omnitokenizer: A joint image-video tokenizer for visual generation
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2b1eae2-d31c-4c63-8d71-7af4e4f2e51f · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3a590ae4-e225-4455-94a5-95dc1de28959 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Omnigen-ar: Autoregressive any-to-image generation
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fae378a-170a-40fc-a943-38bf7090da2e · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Representation entanglement for generation: Training diffusion transformers is much easier than you think
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0ff29571-d063-4e14-9be4-5b8da79761cb · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Grok-1.5 vision preview, 2024
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f37b1219-676d-41bf-930f-686a8fbc47ef · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Vision transformer with deformable attention,
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b4e737f-8506-4735-a904-a7407279b68c · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Vision Transformer with Deformable Attention
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f893c78b-399f-4961-91cf-efbd2244dc8c · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Videogpt: Video generation using vq-vae and transformers, 2021
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71145cbf-138e-4362-bb5f-bd5b55f768a1 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder FasterDiT: Towards Faster Diffusion Transformers Training without Architecture Modification
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 93018c37-be11-4763-8fa7-f072f62d6b3b · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Reconstruction vs
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01cc8e07-66e7-4cf2-bfab-aba86488ab76 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Vector-quantized Image Modeling with Improved VQGAN
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5a89d216-e243-4d74-a8de-d66eca737997 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Scaling Autoregressive Models for Content-Rich Text-to-Image Generation
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7aeacc02-a825-47a9-a700-c0469cee457a · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Language model beats diffusion–tokenizer is key to visual generation
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dc2bc2f-ae86-47ed-a132-8bdeb4826b94 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder An image is worth 32 tokens for reconstruction and generation
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3b63dc6-6667-448c-85d4-d60709bea98e · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Representation alignment for generation: Training diffusion transformers is easier than you think
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 559fc6a7-bbc5-414e-82a8-f9eaf0a68718 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Sigmoid loss for language image pre-training, 2023
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e869353-1d06-49d3-88e4-c37ff811eea3 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder The unreasonable effectiveness of deep features as a perceptual metric
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c4e9324-6488-4379-adaf-e8e7fc994412 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Spherical leech quantization for visual tokenization and generation, 2025
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8e48652e-9415-4f0e-ac53-12c6a299112f · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder arXiv preprint arXiv:2507.08441 , year=
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation da2f3a28-03e6-4da7-963d-580fe207e836 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Diffusion Transformers with Representation Autoencoders
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 69277d24-07ac-4c3c-a60c-59bdb0c99a53 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Fast Training of Diffusion Models with Masked Transformers
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 80afaa1a-edb8-48a5-bfd4-a95221c22ac1 · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Scaling the Codebook Size of VQGAN to 100,000 with a Utilization Rate of 99%
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9325f93c-ee1f-4110-b324-ed3a0dd9fb6a · outbound
IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder A Proofs for Section 3 A.1 Dense RD-AE flow We derive (6)–(7) using the notation of Section 3.1
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
No inbound Pith citation observations are available.