Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T04:16:07.560939Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2608.13556.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T04:16:07.560939Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
32 of 32 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e1ff04b0-f62e-4edf-8228-4f445fd41fea · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Cosmos World Foundation Model Platform for Physical AI
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4447f9fa-a3c3-431d-8097-f730f0dfeb96 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdcfce0e-4edc-4c09-a229-1410ffb3817a · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 822e9cf7-499b-4321-8c06-45437b1d7d2c · outbound
V-RAE: Rethinking Video Latent Spaces for Generation The latent perturbation is used only as a reconstruction-training augmentation; evaluation, latent-statistics estimation, and latent video generation all use clean latents
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7be6db4f-9549-4388-8706-111c5887fa99 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation The Kinetics Human Action Video Dataset
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c640daf3-f0a5-4252-9f5c-4bb9c82b55b5 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation VideoPoet: A Large Language Model for Zero-Shot Video Generation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b88751b-242a-4767-b78d-6090785898a2 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation HunyuanVideo: A Systematic Framework For Large Video Generative Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32c377c0-0683-4177-a23f-c8d74756ed76 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Atoken: A unified tokenizer for vision.arXiv preprint arXiv:2509.14476,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1715a996-a8eb-4e7c-93d8-7855af826bc1 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Open-MAGVIT2: An Open-Source Project Toward Democratizing Auto-regressive Visual Generation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24aedaf7-f34f-43c1-988b-43d5371d167e · outbound
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bce3e1be-f1a3-454b-b373-e8c803aa75d7 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Improved Baselines with Representation Autoencoders
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c927a75-48a4-446e-a8db-1c66ec61b5ed · outbound
V-RAE: Rethinking Video Latent Spaces for Generation UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a1d4876-fec9-4f39-b268-392b5243f904 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Scaling text-to-image diffusion transformers with representation autoencoders.arXiv preprint arXiv:2601.16208,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dea9f47f-51ca-4dea-94e4-9db511d75c7a · outbound
V-RAE: Rethinking Video Latent Spaces for Generation SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ed848e9-08e3-4044-99f2-b9f7936e713e · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Towards Accurate Generative Models of Video: A New Metric & Challenges
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d46f0a2-d88e-4157-87d3-ebf0d5b62093 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Larp: Tokenizing videos with a learned autoregressive generative prior
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c0fc0f33-5340-4164-8dc2-b8294f5720c0 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation VidTwin: Video VAE with Decoupled Structure and Dynamics
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c801366-1054-466b-b96e-bbc3f9721aa7 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Making Reconstruction FID Predictive of Diffusion Generation FID
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e19e2ac3-fd39-4026-88f7-8aa24b5b4913 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Cogvideox: Text-to-video diffusion models with an expert transformer
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 07ba8789-d4c7-4792-84dd-94679d813625 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ce2d571-5fb6-4834-853f-f71141afcf86 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d574185f-1288-415a-aa4d-085b3074756b · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Diffusion Transformers with Representation Autoencoders
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 668e9aa1-a0a4-4e1e-bc87-0ea517977c57 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Efficient universal perception encoder.arXiv preprint arXiv:2603.22387,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f503f54a-b211-4ffd-9eeb-50438244f92e · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Only the input channel count and corresponding time shift change for EUPE-B
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4a60b787-5796-4cda-9d61-5ab9045089ec · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Representation entanglement for generation: Training diffusion transformers is much easier than you think.Advances in Neural Information Processing Systems, 38:7714–7743, 2025a
Reference 2004
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caceaa4c-1ffa-4ba9-be4c-a72b7725b96e · outbound
V-RAE: Rethinking Video Latent Spaces for Generation RoFormer: Enhanced Transformer with Rotary Position Embedding
Reference 2012
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0fe687e-725d-44d6-b5c2-0b3ce4674630 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Dera: Decoupled representation alignment for video tokenization.arXiv preprint arXiv:2512.04483,
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddcbd4b4-db9c-40cf-b329-2bc79b53a10b · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Wan: Open and Advanced Large-Scale Video Generative Models
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d972246-bfff-44ac-bf89-e1863b9d3ed9 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation A Short Note about Kinetics-600
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a781aca-ad0e-49de-a977-ae39dad674cd · outbound
V-RAE: Rethinking Video Latent Spaces for Generation V-JEPA 2.1: Unlocking Dense Features in Video Self-Supervised Learning
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff700b73-2f1e-4f9f-8c12-40d2fd0a3c5c · outbound
V-RAE: Rethinking Video Latent Spaces for Generation CoVLA: Comprehensive vision-language-action dataset for autonomous driving
Reference 2025
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 83a764d7-a9ff-4c8b-a812-46ac2f97e315 · outbound
V-RAE: Rethinking Video Latent Spaces for Generation Improving the Diffusability of Autoencoders
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.