Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T04:32:48.158348Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 0 inbound Pith citation observations for arXiv:2608.08667.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T04:32:48.158348Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
78 of 78 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d2b59e89-d15c-4fee-bf0d-ad21db2c58f6 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies SoundStream: An End-to-End Neural Audio Codec
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation cfb1e325-496f-4b72-ac7c-cf182fc1f6fb · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies High Fidelity Neural Audio Compression
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eedf5e07-1a5f-4115-ba13-f91b73f997d6 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1452ac3a-e68b-402c-a1ba-a499c7043767 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Simple and Controllable Music Generation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e491632-8991-4cc2-ad95-7b3ec1c2ca39 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8ddbbf9-7549-4140-81f6-76fd7ee31168 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Moshi: a speech-text foundation model for real-time dialogue
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6e4639d-ccf9-4e20-b76d-7e15b40a3572 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies AudioLM: a Language Modeling Approach to Audio Generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 680aedd0-803b-4fae-9647-1e979396ba00 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies DiTAR: Diffusion Transformer Autoregressive Modeling for Speech Generation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bba90d26-910c-41ad-b9f9-90b46d3d7161 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Continuous Audio Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afa56b14-ca97-4c73-934b-d52173cce56f · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Cover and Joy A
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7c51330-02ac-4a8d-b75f-af6363632c44 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies UniAudio 1.5: Large Language Model-driven Audio Codec is A Few-shot Audio Task Learner
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa80f50f-65e4-460b-88ca-a6960e769793 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies The Information Bottleneck Method
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1eead16d-e552-4c8d-8eb4-92aaee82e972 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies High-Fidelity Audio Compression with Improved RVQGAN
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3fbad04c-68e5-4345-b72c-69ed12b0b7cb · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Discrete Audio Tokens: More Than a Survey! https://arxiv.org/abs/2506.10274, 2025
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8b297a1-93e5-4c46-bbc7-961cfe1ca123 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f76e66a9-2689-4c01-92e5-6c5640ddb264 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 418150a3-dc47-4c54-890e-6f658235067d · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies BigCodec: Pushing the Limits of Low-Bitrate Neural Speech Codec
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e241e6a3-07c9-4bc7-a7d6-21f9e446c9f6 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Finite Scalar Quantization: VQ-VAE Made Simple
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6133a4c6-3614-43f3-8784-4bcd3ebf1084 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies SoundStorm: Efficient Parallel Audio Generation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de8e945f-6794-4d3c-8279-d13c0ea79703 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Masked Audio Generation using a Single Non-Autoregressive Transformer
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce1932c2-443a-457c-b121-0823992a3ee5 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies W2v-BERT: Combining Contrastive Learning and Masked Language Modeling for Self-Supervised Speech Pre-Training
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce5fab92-d7ab-452b-8292-797e3f8cf7d9 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fab2696-b580-4295-94d8-0f055aad8f63 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies SoCodec: A Semantic-Ordered Multi-Stream Speech Codec for Efficient Language Model Based Text-to-Speech Synthesis
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a1c8901e-0f0d-4fb5-821c-77d2bf1fb045 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies SemantiCodec: An Ultra Low Bitrate Semantic Audio Codec for General Sound
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe5adde9-4640-45d3-b736-5a4b7939f957 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Codec Does Matter: Exploring the Semantic Shortcoming of Codec for Audio Language Model
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d49c1e6b-732d-4b39-9099-322895f7ec7b · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bde1496-b9f9-4427-b04e-13ff8dd373a2 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies WavLM: Large-Scale Self-Supervised Pre-Training for Full Stack Speech Processing
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc6cf0c3-0130-4677-bd7b-856b8021eaa1 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies MOSS-Audio-Tokenizer: Scaling Audio Tokenizers for Future Audio Foundation Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c749d68-1735-49ff-8fa5-439b9cdb195d · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies ALMTokenizer: A Low-bitrate and Semantic-rich Audio Codec Tokenizer for Audio Language Modeling
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cc305a7-31ff-49e2-aa1c-8c4aac5e2a48 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Fewer-token Neural Speech Codec with Time-invariant Codes
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 710342e8-7526-40e8-b051-35247d38dce7 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Learning Source Disentanglement in Neural Audio Codec
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d93c519-337d-4f78-aefc-0f3d941c7728 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b7d3346-a2e6-4982-a0b7-84125d828276 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c7dffe6-19b0-40ab-ab93-6df36700500a · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 033e1daf-d353-4952-8593-68c7cffb7fd9 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Stable Audio Open
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74619dc8-a967-49db-b4d2-5365e11688ff · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies NaturalSpeech 2: Latent Diffusion Models are Natural and Zero-Shot Speech and Singing Synthesizers
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02d37584-975e-4631-9d6e-7f81ba866280 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Make-An-Audio: Text-To-Audio Generation with Prompt-Enhanced Diffusion Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26ae9252-f018-422c-9de0-40b23401c289 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies DashengTokenizer: One layer is enough for unified audio understanding and generation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aed3a8ba-eff8-4564-9320-224800260784 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies WavCube: Unifying Speech Representation for Understanding and Generation via Semantic-Acoustic Joint Modeling
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2eecf9e9-4ac7-4739-aef3-27d356176657 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Ming-UniAudio: Speech LLM for Joint Understanding, Generation and Editing with Unified Representation
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40e40133-7778-4238-a3be-cf5ab78ed5ca · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies LoSATok: Low-dimensional Semantic-Acoustic Tokenizer for Cross-Domain Audio Understanding and Generation
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7b5a52d9-2c0b-48f0-9dff-202af3ada8be · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Flow Matching for Generative Modeling
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbde2eef-b8a9-41e2-9698-44f2f76c5e36 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee5b1124-7fb4-4fe3-b88b-661b823849cb · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies GIVT: Generative Infinite-V ocabulary Transformers
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 68a66f64-d066-4465-8ea1-a405c4fcd8c5 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Autoregressive Image Generation without Vector Quantization
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 91e01e7c-9497-4c35-9fc3-4ad627ebc00b · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Hyperspherical Latents Improve Continuous-Token Autoregressive Generation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0b217a0-0fd8-4229-af2f-fb3244eee52d · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies MaskGIT: Masked Generative Image Transformer
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cb337cd-8f82-4908-8d8b-8a0ec5dd4e6e · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Denoising Diffusion Probabilistic Models.Advances in Neural Information Processing Systems, 33:6840–6851, 2020
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2e3b1db-8c91-4c9d-b78e-28678519f5bb · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Generative Spoken Language Modeling from Raw Audio
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c6e7a43-219d-442a-a9be-318988926923 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Textually Pretrained Speech Language Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7af9cea4-6af8-4ac6-8bbd-8c18142e4867 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies MusicLM: Generating Music From Text
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 919576c7-bfbd-4e15-811f-af3c294d908c · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Diffsound: Discrete Diffusion Model for Text-to-sound Generation
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 992b07fb-5507-4222-ac29-6a397499279c · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies InstructTTS: Modelling Expressive TTS in Discrete Latent Space with Natural Language Style Prompt
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e6decfd-b6b7-4b0d-9cd1-16704eff6bf9 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies E2 TTS: Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0bde76f-c503-4a1f-9568-0732554a0d40 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3592ff85-663b-4f84-b3af-417b3b73eb0b · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Autoregressive Speech Synthesis without Vector Quantization
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation afc55e3d-a429-488d-a0ce-a9232926b084 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies VibeVoice Technical Report
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14f27a22-bbf6-415a-801c-83a19178a63f · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies DASB - Discrete Audio and Speech Benchmark
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 695ee055-b283-4777-965a-4ecb5687dd40 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies TADA: A Generative Framework for Speech Modeling via Text-Acoustic Dual Alignment
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 586f54b3-7fb0-4dbc-9409-73fcce1625d8 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies FlexiCodec: A Dynamic Neural Audio Codec for Low Frame Rates
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dac3314-cccf-4292-8650-377336324588 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Signal Estimation from Modified Short-Time Fourier Transform, 1984
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 90ad8a82-fbf4-4465-9083-21d17a99227b · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies WaveNet: A Generative Model for Raw Audio
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fedec146-001e-4fd3-a522-7f5592eb1cef · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12333eaf-8c9b-4265-8eb5-e5a2924e9724 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Natural TTS Synthesis by Conditioning WaveNet on Mel Spectrogram Predictions
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3082628-f791-459d-a9f3-4393e90dbd7d · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ebdef5a-ff7f-4971-8a60-b3fcdc44870e · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies A Scale for the Measurement of the Psychological Magnitude Pitch
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a6fa749-8f53-4560-bb0e-e2523e919cb4 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Comparison of Parametric Representations for Monosyllabic Word Recognition in Continuously Spoken Sentences
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8d7505e-0465-4e84-a12b-19c646290737 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies WaveGrad: Estimating Gradients for Waveform Generation
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5b0935a2-50dc-4572-a1b2-e851afb857f5 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies DiffWave: A Versatile Diffusion Model for Audio Synthesis
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a72bbb14-e7c8-4620-a0b5-4e100feb541b · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies dMel: Speech Tokenization made Simple
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3fbabab-de1d-447a-af47-08297b6e5b09 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Grad-TTS: A Diffusion Probabilistic Model for Text-to-Speech
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9781d73b-e82d-432b-877d-83190c2d1675 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e800c94e-5a30-402a-bb39-9e4aaf69d7ca · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Sound Texture Perception via Statistics of the Auditory Periphery: Evidence from Sound Synthesis.https://doi.org/10.1016/j.neuron.2011.06.032, 2011
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 454ad516-7a76-4550-b4c9-12049203e106 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Neural Discrete Representation Learning
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bc48978-6e4c-4923-b962-f8b1e1fb50c9 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies SNAC: Multi-Scale Neural Audio Codec
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78417bdc-e7b6-4707-afd5-1877e1372cdd · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies An Image is Worth 32 Tokens for Reconstruction and Generation
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9600fbcc-9f63-44c7-a359-952eb896998d · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies FlexTok: Resampling Images into 1D Token Sequences of Flexible Length
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7282f635-559e-47c4-891b-d17a25385c20 · outbound
A Unifying Perspective on Audio Generative Modeling: Latent Representations and Modeling Strategies Variable-rate discrete representation learning
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.