Pith. sign in

Paper Citation Record · LEDGER

Latent-Compressed Variational Autoencoder for Video Diffusion Models

As of 4 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 1 inbound Pith citation observation for arXiv:2604.16479.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.16479 v1

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T15:23:19.583713Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T05:01:24.112587Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

56 of 56 outbound references displayed

  • verified exact29
  • verified fuzzy26
  • unresolved0
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dfed5bca-6fb6-40ee-834b-290c87435230 · outbound

This paper cites Cosmos World Foundation Model Platform for Physical AI.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.061762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:75329a923fdbce69d521253eb3d982e92bf4c9be024347913528a606f9740135

Observation 299ca983-65ea-478c-a38d-a9abea0c6237 · outbound

This paper cites Frozen in time: A joint video and image encoder for end-to- end retrieval.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Frozen in time: A joint video and image encoder for end-to- end retrieval

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.016766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:b019d227c1ed889d7102e22438407295882268724873b76eb8521c36fa6ed4aa

Observation 62de98d4-5b97-4e57-a82c-1cc21d72cc0f · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.273409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:0be0a5d5e74740d61381af9b2ddd1b4f70a6ec78d9dee07d79b4caddf92f36d0

Observation 62664a9d-9ae2-49f5-b8a0-7593a75b9d30 · outbound

This paper cites Video generation models as world simulators.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Video generation models as world simulators

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.022196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:019058a32a5cdefed8438565c94a61db5dc93d1ce8e6bbca372038505d7c2c07

Observation 20901210-6fd1-41c5-b864-9ae3cf45ed07 · outbound

This paper cites Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.226989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:8552b904855e7eb340a64bfd4b07e7ef78880f085fc4026f980a7b922b633025

Observation 9ef849ae-98c7-451d-abc4-c3364675db64 · outbound

This paper cites Dc-videogen: Efficient video gen- eration with deep compression video autoencoder.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Dc-videogen: Efficient video gen- eration with deep compression video autoencoder

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.083637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:ce69a824254311bbe9fd68ece7e5b4a9604e7df22a06349b38c7022349d32a50

Observation bd28493f-3edb-4436-b076-4de25ed501f4 · outbound

This paper cites Dc-ae 1.5: Accelerating dif- fusion model convergence with structured latent space.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Dc-ae 1.5: Accelerating dif- fusion model convergence with structured latent space

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.024781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:ba720141e9448ae08962a44643d9adc96fc2cafdbd33ee05de33e431fda8ef96

Observation a95f81c1-9a04-4660-8bdb-7c43a13f9e54 · outbound

This paper cites Od- vae: An omni-dimensional video compressor for improving latent video diffusion model.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Od- vae: An omni-dimensional video compressor for improving latent video diffusion model

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.027418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:80e5a81f690cfd5f84c0de0b217fa5bdee0896da66b90c4ddb0e924e8faab7cc

Observation 6d4affe2-e4f6-4c8b-9a5f-4db35d603316 · outbound

This paper cites Panda-70m: Captioning 70m videos with multiple cross- modality teachers.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Panda-70m: Captioning 70m videos with multiple cross- modality teachers

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.029737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:4be79bd9b73f6f5fead61e17c8e84ef5a1ebbc9b622831f2eccb3d49f5fc709f

Observation 6d9acdb6-ad50-4c36-8fac-137715dc5cf9 · outbound

This paper cites Taming transformers for high-resolution image synthesis.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Taming transformers for high-resolution image synthesis

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.031995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:b6a0f6c280a013fa9662e3ddbff75b27812f3629599a8db8c434c0b92a4fd75f

Observation b718b1df-7e54-4312-935b-527521294bf7 · outbound

This paper cites Video generation arena leader- board.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Video generation arena leader- board

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.064237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:d21129f53659f060a3ecede93f466588a0f65ede2779df2cf120ed34e0d29fca

Observation 5688cdf8-fe5f-463c-b86c-14fdcb0ff23d · outbound

This paper cites Seedance 1.0: Exploring the Boundaries of Video Generation Models.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Seedance 1.0: Exploring the Boundaries of Video Generation Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:09:57.662354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:93958b309aea61479b0787d9565a312523880571e6399197b1ceaa377892e7ed

Observation fe0fc68e-b065-4cca-bed9-cdce1f5c7ed0 · outbound

This paper cites Generative adversarial nets.Advances in Neural Information Processing Systems, 27.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Generative adversarial nets.Advances in Neural Information Processing Systems, 27

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.061886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:1d57341944b483254a9ffc541d3e31ca0555077b46fdcace5312c0de81118a66

Observation 019c6939-0c20-44b8-ab20-4cf7bd6e40e6 · outbound

This paper cites An introduction to wavelets.IEEE computa- tional science and engineering, 2(2):50–61.

Latent-Compressed Variational Autoencoder for Video Diffusion Models An introduction to wavelets.IEEE computa- tional science and engineering, 2(2):50–61

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.066750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:144c8c6d9a69cb0dc627c8a71a360ca5cb3bd52a9fa83741647ddab50be5c54c

Observation 379c8280-c1fa-4db6-8b19-b58afae57b6f · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

Latent-Compressed Variational Autoencoder for Video Diffusion Models AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.030009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:1e0a15f5608b28bfc23a81dc4241f7619abc768ad24a9541f77206b06d022c34

Observation 1b6693c7-a193-487d-a603-9726cd377731 · outbound

This paper cites LTX-Video: Realtime Video Latent Diffusion.

Latent-Compressed Variational Autoencoder for Video Diffusion Models LTX-Video: Realtime Video Latent Diffusion

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.364378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:c239f55d57576df82da460c90abf05ebed4f534c6c4c7b07bfa6dc03267f266d

Observation ac0c1624-78af-40ce-9719-35b212adf4b2 · outbound

This paper cites Learnings from Scaling Visual Tokenizers for Reconstruction and Generation.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Learnings from Scaling Visual Tokenizers for Reconstruction and Generation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.077326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:09e53cd3671470ba2e363d513af4d1fbaf70ae07977554ff04cac9b29b61dd80

Observation 1f1b6ae9-bb4e-4dc1-b469-8e31eb10843c · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:27:43.523852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:1829247301a26b09396bebad6b9663c3e994ec6593dc98b6b3d9557020e2ec99

Observation 2e8e02ed-5c3f-4c03-9a86-dba9533f73c5 · outbound

This paper cites simple diffusion: End-to-end diffusion for high resolution images.

Latent-Compressed Variational Autoencoder for Video Diffusion Models simple diffusion: End-to-end diffusion for high resolution images

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.070341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:a39ea3c2ed70413dc176adf3d547dc47efaf7e52ef84664afd967a1bb7e24281

Observation 2c9a178d-1bb3-46f4-add7-f34f44d93e5c · outbound

This paper cites Image quality metrics: Psnr vs.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Image quality metrics: Psnr vs

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.079767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:59e665787796e7e4f915ef85a8e0605b0d8b434cc979c7bb7ff4e94f1476894d

Observation f3c9203b-3fe4-4318-93c9-3c3577e834cd · outbound

This paper cites The Kinetics Human Action Video Dataset.

Latent-Compressed Variational Autoencoder for Video Diffusion Models The Kinetics Human Action Video Dataset

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.295921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:d5a52c6573538324ceb01f75ea5c5de88fa086e1f81d2ea2082f954b848da94f

Observation 7325a249-a7d9-48c0-a616-17cfd1c12dd2 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Adam: A Method for Stochastic Optimization

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.177123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:2c2793b3960b856d69eac8e8a1ebfb9ea476bc9ccc2df57e6a6f0e72816529e6

Observation b8adbe17-00c9-478a-890c-41f63f0d2cc7 · outbound

This paper cites Auto-Encoding Variational Bayes.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Auto-Encoding Variational Bayes

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.044934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:da7669e1debfe2edee3eaaed9a0e7dc1f03fba26f4324ab1c6e2b63058983d3f

Observation d8789e48-97b0-4e7a-b65d-b281bf12358b · outbound

This paper cites VideoPoet: A Large Language Model for Zero-Shot Video Generation.

Latent-Compressed Variational Autoencoder for Video Diffusion Models VideoPoet: A Large Language Model for Zero-Shot Video Generation

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:51:05.830483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:58ed10c9f44a45a178a455b55cac3ea0a2bc7ad9f61c566a853acda870be6bc3

Observation 2bca163d-c335-4628-b5b3-037118c5bc6a · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Latent-Compressed Variational Autoencoder for Video Diffusion Models HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.018392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:d6795e4d437cb45427c6b96c6d304482e59275c7fe95889e5312018549b69c9a

Observation 45ba7cd5-1331-4cd5-8f01-3c90311fa8b0 · outbound

This paper cites Video autoencoder: self-supervised disentanglement of static 3d structure and motion.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Video autoencoder: self-supervised disentanglement of static 3d structure and motion

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.056940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:041d7ba6c6b8a5c7b06965288d23646e176ab05a6423fcade5e418fb1e5eceba

Observation d5a3f91a-7680-48b3-8914-5bd17281410f · outbound

This paper cites Wf-vae: Enhancing video vae by wavelet-driven energy flow for latent video diffusion model.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Wf-vae: Enhancing video vae by wavelet-driven energy flow for latent video diffusion model

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.044026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:9508340a85a26ec2465ec296690647aade501ab5d942fb0b0acc559e628891c3

Observation 8353aabe-6aa1-4242-8a43-366ddb86beb0 · outbound

This paper cites Open-Sora Plan: Open-Source Large Video Generation Model.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Open-Sora Plan: Open-Source Large Video Generation Model

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.345396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:533df12a40c19a428fd9ffba6cafb5206112b429c81523aa73d719acf472052c

Observation 59b3d51f-4e21-4d0a-9cba-4fba4daa9806 · outbound

This paper cites Hi-VAE: Efficient Video Autoencoding with Global and Detailed Motion.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Hi-VAE: Efficient Video Autoencoding with Global and Detailed Motion

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.093142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:f7f2277376fccf4c75d7f01f1ace3457b3c17f0b9ed4f9e48fba515c5c014ece

Observation 78409137-931e-4603-9148-3cdba8c041a7 · outbound

This paper cites Decoupled Weight Decay Regularization.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Decoupled Weight Decay Regularization

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.316506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:01fe19162dcb6fe055c009cea5bf26b808d8bbb3146f051dec3819a829b2c651

Observation 4c32aa39-611f-420a-95df-449cd8b044ca · outbound

This paper cites Latte: Latent diffusion transformer for video generation.Transactions on Machine Learning Research.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Latte: Latent diffusion transformer for video generation.Transactions on Machine Learning Research

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.051289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:d69118098c0f8d18bd1125ee6a9e896b41ef81d943c96c1e53a375da1a395078

Observation ba2dd2a9-6785-41a6-a0ce-e476e6b03df6 · outbound

This paper cites OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation.

Latent-Compressed Variational Autoencoder for Video Diffusion Models OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:34:53.267725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:43e9be44de467cedae793327655fead2638f7caf95c6d1fec2c69527e1c61c52

Observation 8114c507-9fcb-43bf-835d-136e00231656 · outbound

This paper cites Movie Gen: A Cast of Media Foundation Models.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Movie Gen: A Cast of Media Foundation Models

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:26.778393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:550734316f2b726c2ad03aac78b539cd268f6987089ffe2c580cb4a9dcca8627

Observation 5c6fce5f-89e4-4c23-b771-5806db933590 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Latent-Compressed Variational Autoencoder for Video Diffusion Models High-resolution image synthesis with latent diffusion models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.074653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:d05429f0dd1c76207647e500d2fb2430446684ae2eac45f33c7a4e828b596e37

Observation 4010fa6e-836f-40a9-a49b-d47def46bd88 · outbound

This paper cites Temporal generative adversarial nets with singular value clipping.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Temporal generative adversarial nets with singular value clipping

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.046459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:fb521c3f879241afac3941ec1686a8b918d0d5380ee2bc184ca80d4719405a2d

Observation d4c993ac-391b-4521-bcc5-67f7c7a4aca3 · outbound

This paper cites The JPEG 2000 still image compression standard.

Latent-Compressed Variational Autoencoder for Video Diffusion Models The JPEG 2000 still image compression standard

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.054174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:d5723f3472341732d308600d2a37af37b964b7f83e4cafeb94636292999aade2

Observation e3fbc601-646c-47c6-a677-8c997a906602 · outbound

This paper cites Stylegan-v: A continuous video generator with the price, image quality and perks of stylegan2.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Stylegan-v: A continuous video generator with the price, image quality and perks of stylegan2

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.048897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:416e8e5eaab66481f5acc8a3e04dddd4bf03d70636054e8e048aef997bb5a9d1

Observation 22b7c1fa-f686-4732-a324-d4e83df66dfe · outbound

This paper cites Improving the Diffusability of Autoencoders.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Improving the Diffusability of Autoencoders

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.024402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:b3fc7d26622495161a300aa60617eede875cad684f3e2fa354861d2261fb94c3

Observation 14da305a-1bd1-4404-9fc6-e7b623cbfa84 · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

Latent-Compressed Variational Autoencoder for Video Diffusion Models UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.251740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:268c6e95a07ff50f6637c8a83bae4bd1e31711a031ad8a37c8b6275b735f243f

Observation b5459e31-b2f8-4ed5-a603-69b9b53627d4 · outbound

This paper cites Adapting LLMs to Time Series Forecasting via Temporal Heterogeneity Modeling and Semantic Alignment.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Adapting LLMs to Time Series Forecasting via Temporal Heterogeneity Modeling and Semantic Alignment

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.307568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:eef8fd176a4e6c9b7276120a08424d5c6093ebc346ffe4e61cb28eb58c2726d7

Observation e5292539-05bd-4aab-a127-72e04750685d · outbound

This paper cites Haar Wavelet Based Approach for Image Compression and Quality Assessment of Compressed Image.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Haar Wavelet Based Approach for Image Compression and Quality Assessment of Compressed Image

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-04T18:11:21.620230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:7a4cbadcb4321af0299b04c5fafa1f3c94cbf0986d1b3df5005f5a5c34a033ce

Observation 75ce7ae5-0377-4ea9-b494-9c7b4ea5a487 · outbound

This paper cites FVD: A new metric for video generation.

Latent-Compressed Variational Autoencoder for Video Diffusion Models FVD: A new metric for video generation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.014436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:fdeb35444be3381a358ebc6e6ded3e6e245f7df20e7b17945d0dd12b3a57ddbd

Observation 7a42e7d1-a736-4ea2-b3b7-2ef1d992341c · outbound

This paper cites Attention is all you need.Advances in Neural Information Processing Systems, 30.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Attention is all you need.Advances in Neural Information Processing Systems, 30

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.018956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:9a92cbdece7768fc033dd47b74bca7a97982c852567aa36bafe3d3c7b97221b8

Observation 1d8fd16f-127a-4aa7-89ac-9dd60d02c1db · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Wan: Open and Advanced Large-Scale Video Generative Models

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.215355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:baa05d5eba0d417a2554365c9beb2cc67dfa72a1e1c736bdb2df5fe36db5fac4

Observation 718de5d1-f640-48e9-b5ef-ad0c8738b35f · outbound

This paper cites Image quality assessment: from error visibility to structural similarity.IEEE Transactions on Image Processing, 13(4):600–612.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Image quality assessment: from error visibility to structural similarity.IEEE Transactions on Image Processing, 13(4):600–612

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.036895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:1cfd6aa3e79d250bb6115beecf827bff41db3e2ab62c5066cdbf0991736ecb1e

Observation f00a44cd-fde9-42b1-8786-25d8e2eae49f · outbound

This paper cites Improved video V AE for latent video diffusion model.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Improved video V AE for latent video diffusion model

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.041693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:775a0a66bef2006c8b59b2550c25e03baa6c0a99a26a62cbaa34e975c83112cd

Observation c279d6c0-affa-4d50-842d-6f4cecb259e9 · outbound

This paper cites H3ae: High compression, high speed, and high quality autoencoder for video diffusion models.

Latent-Compressed Variational Autoencoder for Video Diffusion Models H3ae: High compression, high speed, and high quality autoencoder for video diffusion models

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.162685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:432a0590781cb78beeef60b9782dbe14e7ee474a91459bcd723bff4d7f79b74e

Observation 3f1f18fe-4366-4385-8537-b9c8b50e29db · outbound

This paper cites Learning to generate time-lapse videos using multi-stage dy- namic generative adversarial networks.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Learning to generate time-lapse videos using multi-stage dy- namic generative adversarial networks

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.034424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:9fd8303e1c9deec6eee139452dbb3cd36975254fa0ee1a6718989d97f3513fd7

Observation 5636f674-a34d-423c-93c8-18822b8e2238 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

Latent-Compressed Variational Autoencoder for Video Diffusion Models CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.053184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:2d7a1eed7a7ecab23d8aee3c5fe89eed401ceeff2303a39a8eccba66b42f1664

Observation 77bafbe1-84b6-488a-acff-59a4c5739655 · outbound

This paper cites Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:06:45.090605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:f0a03929ecc94af49e85b2e38ebd5a2a805312347c71c91681efbc6133887f89

Observation 1c000e58-fda9-4a7a-a9b9-0108e4c15c48 · outbound

This paper cites Efficient Video Diffusion Models via Content-Frame Motion-Latent Decomposition.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Efficient Video Diffusion Models via Content-Frame Motion-Latent Decomposition

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.353140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:16c710d07749acc78ba5f79a9c48bd5db1babbd22f7fd066f85e3dca3d9877d3

Observation 8c7ce93b-2f00-4337-a11b-2cfeeed89b72 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

Latent-Compressed Variational Autoencoder for Video Diffusion Models The unreasonable effectiveness of deep features as a perceptual metric

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.039061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:56b9060a42586f9e23d50fee5e971a61fb9120ed92d9b38e57a0393cc1cf46ef

Observation b0cd959c-cd32-423a-819f-a2579642c43c · outbound

This paper cites A survey on perceptually optimized video coding.

Latent-Compressed Variational Autoencoder for Video Diffusion Models A survey on perceptually optimized video coding

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.059288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:7e21b14acc2450f0495a0ecba5a600e12297ee967c92847e40a71f215c1ef31b

Observation 5a609579-f837-4fd6-9024-cd913e77c140 · outbound

This paper cites Cv- vae: A compatible video vae for latent generative video mod- els.Advances in Neural Information Processing Systems, 37: 12847–12871.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Cv- vae: A compatible video vae for latent generative video mod- els.Advances in Neural Information Processing Systems, 37: 12847–12871

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.077188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:bf38d0ee83a2eddaf7ae14f179c2f8707ee14e186cb1543e27b87b9fc98dccf9

Observation 192448e4-2953-4271-bbb2-8cd6066d7af5 · outbound

This paper cites Open-Sora: Democratizing Efficient Video Production for All.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Open-Sora: Democratizing Efficient Video Production for All

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:01:52.219350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:82838606481b0a97a86f7b4b9d0477ee0422f22f80def05cb470b16b685ea48e

Observation f16c48cd-0194-4afb-9b9a-68c00bc42e9e · outbound

This paper cites Allegro: Open the Black Box of Commercial-Level Video Generation Model.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Allegro: Open the Black Box of Commercial-Level Video Generation Model

Reference 56

Resolution
malformed identifier
arxiv_id, observed 2026-05-11T10:41:03.125071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:0351cf0f72e3bfc7711ae3dd57ebf2f7f0a639d46c6529295f81b6376d1aced0

Pith citing papers

Observation 00d907ee-09eb-41de-a468-c6ec1b203ffc · inbound

Kepler-Encoder-v0.1: Towards a Multimodal Embedding Model for Robots cites this paper.

Kepler-Encoder-v0.1: Towards a Multimodal Embedding Model for Robots Latent-Compressed Variational Autoencoder for Video Diffusion Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T05:01:24.112587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:01:24.112587Z digest=sha256:e19093e5b7c065f9c8327a2a63edccf4b98e5c16a021284bb2be8f1821783606