Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-20T14:53:15.718359Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 1 inbound Pith citation observation for arXiv:2605.17085.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-20T14:53:15.718359Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-20T14:53:15.718359Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-05-20T14:53:23.198073Z
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2801ff57-1e74-4225-b0f4-afd3a9f4636a · outbound
Taming Audio VAEs via Target-KL Regularization Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 771bc962-b180-478b-a295-4c2b60751b29 · outbound
Taming Audio VAEs via Target-KL Regularization Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 37ab0a92-c010-47af-95f0-b251a54e9137 · outbound
Taming Audio VAEs via Target-KL Regularization Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 87d74ee6-96e8-40ac-bede-460fd3eb4ead · outbound
Taming Audio VAEs via Target-KL Regularization Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 08e5bc35-139f-4b17-88df-1dfb9801d680 · outbound
Taming Audio VAEs via Target-KL Regularization Taming Audio VAEs via Target-KL Regularization
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 684f0f92-7476-44c4-8ded-005f1e3a7e21 · outbound
Taming Audio VAEs via Target-KL Regularization Model architecture Our model is built on the same framework of neural audio codec models, except we replace the quantization bottleneck with a gaus- sian regularization
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2611bae7-9537-42a6-a7ee-8479da6ff6ae · outbound
Taming Audio VAEs via Target-KL Regularization Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 91155a2c-46dd-4c89-8208-5490d10a79d8 · outbound
Taming Audio VAEs via Target-KL Regularization This allows for direct comparison to discrete neural audio codecs and enables systematic study of the rate-distortion trade-off for continuous audio compres- sion models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f3f913b4-64b3-4bad-affe-0c2bd62ec039 · outbound
Taming Audio VAEs via Target-KL Regularization Neural discrete representation learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1ed826fb-88ba-4124-8c89-3ed2c5b9db55 · outbound
Taming Audio VAEs via Target-KL Regularization High-resolution image synthesis with latent diffusion models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 92576b4d-5b78-447a-ad0d-bee40344e3ca · outbound
Taming Audio VAEs via Target-KL Regularization Audiolm: a language modeling approach to audio generation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4770f34f-9e9f-4ada-8db9-fd99cd8978ec · outbound
Taming Audio VAEs via Target-KL Regularization Stable audio open
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 77cf3431-ac9d-4fce-b064-c5ca517ad311 · outbound
Taming Audio VAEs via Target-KL Regularization Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fd56e338-9c70-4e3b-b60e-b98286b6bad7 · outbound
Taming Audio VAEs via Target-KL Regularization VampNet: Music Generation via Masked Acoustic Token Modeling
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b8583be2-aef0-4be5-98d7-8b637f38461b · outbound
Taming Audio VAEs via Target-KL Regularization MaskGCT: Zero-Shot Text-to-Speech with Masked Generative Codec Transformer
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 76880868-1f65-483d-a52e-60e57fb1c33d · outbound
Taming Audio VAEs via Target-KL Regularization SoundStorm: Efficient Parallel Audio Generation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7d7aca55-741a-471d-add7-dd127d086ea2 · outbound
Taming Audio VAEs via Target-KL Regularization Auto-Encoding Variational Bayes
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7a56c53d-6cf0-4c59-af0d-26f022923b05 · outbound
Taming Audio VAEs via Target-KL Regularization Denoising dif- fusion probabilistic models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 933421a4-e9a7-4c06-9768-f425aac0e273 · outbound
Taming Audio VAEs via Target-KL Regularization Au- dioLDM: Text-to-audio generation with latent diffusion mod- els
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2d4bb04f-ccf6-4b91-85a7-7c2c06ce171b · outbound
Taming Audio VAEs via Target-KL Regularization Scaling rectified flow transformers for high-resolution image synthesis
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 83b4282f-b643-445d-bc11-3de7a17fbed1 · outbound
Taming Audio VAEs via Target-KL Regularization Soundstream: An end-to- end neural audio codec
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9d243bcb-1239-48e8-81cd-3fc20255b94c · outbound
Taming Audio VAEs via Target-KL Regularization High-fidelity audio compression with improved rvqgan
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3a2cb772-818c-41ff-aa15-7eadf2cf56a9 · outbound
Taming Audio VAEs via Target-KL Regularization In- terpreting rate-distortion of variational autoencoder and using model uncertainty for anomaly detection
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 88592c11-9e50-43de-ad61-c73f6bf38523 · outbound
Taming Audio VAEs via Target-KL Regularization Practical Lossless Compression with Latent Variables using Bits Back Coding
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e3159072-c9d8-4185-88e6-4e2bac5f83f3 · outbound
Taming Audio VAEs via Target-KL Regularization Fixing a Broken ELBO
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bdb16ce2-de3e-455c-9a8a-64ca6abb24a0 · outbound
Taming Audio VAEs via Target-KL Regularization Improved variational in- ference with inverse autoregressive flow
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c62bb11d-3d35-44f0-bea4-db2b52abbf64 · outbound
Taming Audio VAEs via Target-KL Regularization An introduction to variational autoencoders
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 636c0cda-1aa7-491a-803c-a712e014276f · outbound
Taming Audio VAEs via Target-KL Regularization BigVGAN: A Universal Neural Vocoder with Large-Scale Training
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c54fba4e-1e70-448e-97f0-4089dc6858d8 · outbound
Taming Audio VAEs via Target-KL Regularization Moshi: a speech-text foundation model for real-time dialogue
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9cf33edb-f31d-4617-8a83-2def3c1439e0 · outbound
Taming Audio VAEs via Target-KL Regularization Audio set: An ontology and human-labeled dataset for audio events
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 20135b68-3616-4b79-bada-563d7092f5ab · outbound
Taming Audio VAEs via Target-KL Regularization SpectroStream: A Versatile Neural Codec for General Audio
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6b273d94-6d2f-4367-8629-6fd27bb371f1 · outbound
Taming Audio VAEs via Target-KL Regularization High Fidelity Neural Audio Compression
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a1ea9b96-7ea7-4f91-a008-fa96c690987a · outbound
Taming Audio VAEs via Target-KL Regularization Progressive Distillation for Fast Sampling of Diffusion Models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c30fea26-bd78-456d-8ad6-14f4103c5604 · outbound
Taming Audio VAEs via Target-KL Regularization sim- ple diffusion: End-to-end diffusion for high resolution im- ages
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7262eaa6-04e7-4c77-96c1-8d0a3fb68ae6 · outbound
Taming Audio VAEs via Target-KL Regularization Simple-tts: End-to-end text-to-speech synthesis with latent diffusion
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9c9e3d41-425b-40ee-987f-4911387a26f6 · outbound
Taming Audio VAEs via Target-KL Regularization Scalable diffusion models with transformers
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 69d5631a-61d2-4e30-9e6a-6acd53f2c2cf · outbound
Taming Audio VAEs via Target-KL Regularization Ditto-tts: Efficient and scalable zero-shot text-to-speech with diffusion transformer
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 054760b5-1092-498f-a46d-067d866ca8d2 · outbound
Taming Audio VAEs via Target-KL Regularization Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e0306326-c226-411b-a171-bd5f9c80c4fb · outbound
Taming Audio VAEs via Target-KL Regularization Byt5: Towards a token-free future with pre-trained byte-to- byte models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9dd8e9e3-17a6-43e0-b1c3-47de3dde233d · outbound
Taming Audio VAEs via Target-KL Regularization Exploring the limits of transfer learning with a unified text-to-text transformer
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6c3d5fca-e480-47fb-8de3-1a6bfa14f44e · outbound
Taming Audio VAEs via Target-KL Regularization Phonemizer: Text to phones transcription for multiple languages in python
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3c53ee44-f5df-41db-84f7-c92726aefef5 · outbound
Taming Audio VAEs via Target-KL Regularization Emilia: A large-scale, extensive, multilin- gual, and diverse dataset for speech generation
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3e4bfec1-d8dd-4d70-8845-cb42683db895 · outbound
Taming Audio VAEs via Target-KL Regularization SILA: Signal-to-Language Augmentation for Enhanced Control in Text-to-Audio Generation
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4492068a-99c8-41aa-b65c-f503dcfe9086 · outbound
Taming Audio VAEs via Target-KL Regularization Sketch2sound: Controllable audio generation via time-varying signals and sonic imitations
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2deef7fe-17c5-41e7-86e9-fd1420012cc4 · outbound
Taming Audio VAEs via Target-KL Regularization FLAM: Frame-wise language-audio model- ing
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f9db439e-e46e-4f1f-ace9-3ae10376f717 · outbound
Taming Audio VAEs via Target-KL Regularization Scaling instruction- finetuned language models
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 59e97aa2-3ac4-4d1f-b1a7-bde25cad6181 · outbound
Taming Audio VAEs via Target-KL Regularization Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 76bbbf21-7af8-4adf-9c7d-07e30845c50b · outbound
Taming Audio VAEs via Target-KL Regularization Finite Scalar Quantization: VQ-VAE Made Simple
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6d3b93c9-7adf-4e9d-ae65-8fa20a0d5050 · outbound
Taming Audio VAEs via Target-KL Regularization Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 08e5bc35-139f-4b17-88df-1dfb9801d680 · inbound
Taming Audio VAEs via Target-KL Regularization Taming Audio VAEs via Target-KL Regularization
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.