Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-15T12:26:37.300599Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 100 inbound Pith citation observations for arXiv:2406.02430.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-15T12:26:37.300599Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:23:15.641196Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
45 of 45 outbound references displayed
External citation measurements
5
pith, observed 2026-08-05T02:28:24.338817Z
Observation 4756c9b5-0052-4463-b6c7-7f5322bf9b4c · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Streaming voice conversion via intermediate bottleneck features and non-streaming teacher guidance
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7a3c3035-377d-4058-b211-d959de06cbb1 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models StreamVoice: Streamable Context-Aware Language Modeling for Real-time Zero-Shot Voice Conversion
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e2f60dc4-b155-4d77-ad62-d3f221336643 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models BASE TTS: Lessons from building a billion-parameter Text-to-Speech model on 100K hours of data
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 53c5908d-2c1c-4b39-8fe2-588f9940ae35 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 006c7288-10a0-4add-ba1c-330ff6c2651a · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Deep Reinforcement Learning: An Overview
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 979bebc8-0d3e-45fd-8ca7-22bb17eac15d · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 644271db-4c5d-4826-9649-6a0f66576348 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models ResGrad: Residual Denoising Diffusion Probabilistic Models for Text to Speech
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1d4b87f3-9a05-4801-bdb4-14a32e18b6dc · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models LLaMA: Open and Efficient Foundation Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 36c79e26-f01d-4b1e-954a-a6c8070fc732 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Better speech synthesis through scaling
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d23726c2-113b-44f4-aa51-28e4166d5cf8 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models BigVGAN: A Universal Neural Vocoder with Large-Scale Training
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0441f097-3029-4902-b262-3fe63dad3c05 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Glow-WaveGAN: Learning Speech Representations from GAN-based Variational Auto-Encoder For High Fidelity Flow-based Speech Synthesis
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dd3e66ab-0ea7-41fc-af35-223fa74ac991 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Basis-MelGAN: Efficient Neural Vocoder Based on Audio Decomposition
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d7ebdee7-0e41-4c7c-89d6-d2cb89f158a1 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Common Voice: A Massively-Multilingual Speech Corpus
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e025bd0-40a1-40b1-89aa-44c5492b1da1 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models DiDiSpeech: A large scale mandarin speech corpus
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 029a82d4-079b-478f-abb3-40d13ae9a8f8 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models FunASR: A Fundamental End-to-End Speech Recognition Toolkit
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e3ced301-f751-4fa9-a4b5-f08d66b409ff · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models FastSpeech 2: Fast and High-Quality End-to-End Text to Speech
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f7720dba-20cb-429d-9c36-7acfce416ee8 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Controllable and Lossless Non-Autoregressive End-to-End Text-to-Speech
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4888ea06-1982-4030-9480-89c7d4842d21 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models LibriSpeech: an ASR corpus based on public domain audio books
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3bf8aecc-c6a2-43e1-bde2-2c0212de1e5e · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models WeNet 2.0: More Productive End-to-End Speech Recognition Toolkit
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 244c1ec1-8efe-40cb-824b-9ec48aa3bb21 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Developing far-field speaker system via teacher-student learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation be268926-e6b8-4033-9530-b496693524b1 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models VoxCeleb: a large-scale speaker identification dataset
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation feb995b5-9758-4d62-b15d-8bf198e9a477 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Singing Voice Synthesis Using Deep Autoregressive Neural Networks for Acoustic Modeling
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1a553839-b497-4fc3-8507-673b51d844c1 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models LiteSing: Towards fast, lightweight and expressive singing voice synthesis
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eaab47d9-a256-4d2c-bb7c-82f7902d3886 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Prosody-aware SpeechT5 for expressive neural TTS
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8d25e85c-fb2c-4627-a062-0503dff6d454 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 75228981-875d-439e-84e7-34cba46e75a8 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5126f41a-d88e-41e7-b8cc-29a36906a6a9 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2416902b-a921-482c-bde9-5c4c5826eb96 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Consistency Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f8a34ae0-e1a1-436d-acc6-d9a3c263153c · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Scaling Rectified Flow Transformers for High-Resolution Image Synthesis
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8fa63e6b-41c5-4885-a1e7-7d0509da4b5a · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e4d6af0-fd01-47db-b303-3273d6c9719c · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models A White Paper on Neural Network Quantization
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2435136e-38f8-4e99-84cf-e3400f2314c3 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models decoupleQ: Towards 2-bit Post-Training Uniform Quantization via decoupling Parameters into Integer and Floating Points
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6092f1ee-f92a-4a94-96ec-7c9a045bf1c2 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1b46cae5-38b9-4e85-b7a8-66bdc50c31c2 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 09a8622b-3123-44d5-bc89-b24fd7ab793b · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Zero-Shot Accent Conversion using Pseudo Siamese Disentanglement Network
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c7ee5465-f4e3-43ff-a7f0-10f80a53ac2a · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Diffusion-Based Voice Conversion with Fast Maximum Likelihood Sampling Scheme
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5e45a6f5-b77a-4f83-867a-09990480a4b3 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Proximal Policy Optimization Algorithms
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 985d286e-bda1-4c57-81d3-b0dcdb611313 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Diffusion Model Alignment Using Direct Preference Optimization
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2cae3087-dbe4-41ee-a48a-66159963c655 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models MusicRL: Aligning Music Generation to Human Preferences
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2b1927a3-db00-49e5-9f38-c32ea6a04dea · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models SpeechAlign: Aligning Speech Generation to Human Preferences
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aa06ca70-0c3d-4002-80f4-1a3e4a9dc51a · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b1a582d4-f2e3-43b8-a5b4-83a6915af56d · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Minimum word error rate training for attention-based sequence-to-sequence models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3fffca85-55af-4cda-a00a-bed68e84db18 · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Transforming and Combining Rewards for Aligning Large Language Models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8269caef-9368-47f3-809b-a6a66bde15fd · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb7da3c8-9b5c-419a-aa7b-70d4dd5b084d · outbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models SpeechX: Neural Codec Language Model as a Versatile Speech Transformer
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 041024c9-5f47-488e-9eb4-24126add3958 · inbound
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6000b29f-4584-435c-9497-184926710ea2 · inbound
F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2287c955-3603-42e4-99a2-762e4db69d33 · inbound
CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0024d1c8-e6d0-4bcc-9ca5-e0bd5e3fc405 · inbound
Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e91e50a7-28a0-41b1-a4ec-af7cc51c30ab · inbound
Qwen2.5-Omni Technical Report Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 66210f73-2ac1-4062-9482-daa2511608c5 · inbound
SeamlessEdit: Background Noise Aware Zero-Shot Speech Editing with in-Context Enhancement Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ca8bfa3e-ec9a-44ad-b667-1584f0572553 · inbound
Accelerating Autoregressive Speech Synthesis Inference With Speech Speculative Decoding Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d33002b-6b3c-47b2-89c1-3fcdaa12fcef · inbound
UniTTS: An end-to-end TTS system without decoupling of acoustic and semantic information Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0938d39-b2e4-4570-afee-d5c45c76504f · inbound
CosyVoice 3: Towards In-the-wild Speech Generation via Scaling-up and Post-training Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8d541a03-a710-477c-a4db-f2ce42905748 · inbound
VoiceStar: Robust Zero-Shot Autoregressive TTS with Duration Control and Extrapolation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d4d6018-6c07-457a-8bcd-920d8e4b9f3c · inbound
Accelerating Diffusion-based Text-to-Speech Model Training with Dual Modality Alignment Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d788c02-3a57-4d3a-9989-b5a4490215d9 · inbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 061b2b2b-82f4-4c5a-84d2-8a4e4e9d4b26 · inbound
Phir Hera Fairy: An English Fairytaler is a Strong Faker of Fluent Speech in Low-Resource Indian Languages Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 914a2250-af2d-4a02-841e-30a570b8fdb0 · inbound
MagiCodec: Simple Masked Gaussian-Injected Codec for High-Fidelity Reconstruction and Generation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c068f11-9a6e-4a59-b3d1-ebd2798c923b · inbound
NTPP: Generative Speech Language Modeling for Dual-Channel Spoken Dialogue via Next-Token-Pair Prediction Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd8141c1-7133-4193-a7d0-aac7c6faba2a · inbound
RoboEgo System Card: An Omnimodal Model with Native Full Duplexity Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97c7ba52-f118-4652-bb40-c2f3d27a5a33 · inbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42ac2bef-7eda-46a8-93e0-beeed2e49c55 · inbound
A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 811b01f7-0773-4610-a07c-7bf31ea74375 · inbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0d2d73e-b959-474c-abb0-a24333169bfc · inbound
SimuPanel: A Novel Immersive Multi-Agent System to Simulate Interactive Expert Panel Discussion Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba6482a5-3d62-479f-b595-93150211c630 · inbound
IndexTTS2: A Breakthrough in Emotionally Expressive and Duration-Controlled Auto-Regressive Zero-Shot Text-to-Speech Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fff0af9-2b2b-4938-aa93-41cff71acb69 · inbound
Robust and Efficient Autoregressive Speech Synthesis with Dynamic Chunk-wise Prediction Policy Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b63ce8a-c9bf-4023-b1e3-db6397667678 · inbound
Differentiable Reward Optimization for LLM based TTS system Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae0691d8-4a74-473f-aada-22d1c6418655 · inbound
Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c215fa50-64f0-4d60-b724-468a474cd059 · inbound
ZipVoice-Dialog: Non-Autoregressive Spoken Dialogue Generation with Flow Matching Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 971e9b2c-bccb-40d3-ac1b-bd6a2a5cb146 · inbound
Quantize More, Lose Less: Autoregressive Generation from Residually Quantized Speech Representations Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49409248-3195-4e3d-b17a-5e25c7bb2263 · inbound
DMOSpeech 2: Reinforcement Learning for Duration Prediction in Metric-Optimized Speech Synthesis Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f313ab2b-dc9a-4109-a0a2-f5ea84ba8cf5 · inbound
Step-Audio 2 Technical Report Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e0ae2f24-26f0-4756-b95f-8a1031fa1b01 · inbound
SplitMeanFlow: Interval Splitting Consistency in Few-Step Generative Modeling Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4245843-32f7-45af-b82e-a5839e43af1a · inbound
Seed LiveInterpret 2.0: End-to-end Simultaneous Speech-to-speech Translation with Your Voice Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be7f1afa-59f8-436c-94e5-b0b99ba1b117 · inbound
BoSS: Beyond-Semantic Speech Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aed9d272-a437-42ff-bec8-c77046dd3603 · inbound
TTS-1 Technical Report Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 368df377-e60e-4603-b996-af6b64354af5 · inbound
Adaptive Duration Model for Text Speech Alignment Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c770a0db-f760-470e-a2bc-9da33e913861 · inbound
AudioGen-Omni: A Unified Multimodal Diffusion Transformer for Video-Synchronized Audio, Speech, and Song Generation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54bdc28c-fedc-4056-8ee1-8e20b8959378 · inbound
Towards Hallucination-Free Music: A Reinforcement Learning Preference Optimization Framework for Reliable Song Generation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 653e739d-550c-4d36-abff-8a836d6e1634 · inbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65d4b235-9ce0-4f6d-bdf3-10d2af920dff · inbound
NoteIt: A System Converting Instructional Videos to Interactable Notes Through Multimodal Video Understanding Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 958bcb74-b0d2-4b6d-b750-76a12a654377 · inbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8d55cdd-b927-4d12-b893-ac9fee5a77e7 · inbound
MPO: Multidimensional Preference Optimization for Language Model-based Text-to-Speech Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78740931-f52f-485a-baba-662d9142bc05 · inbound
FireRedTTS-2: Towards Long Conversational Speech Generation for Podcast and Chatbot Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 413d3af7-a553-4508-a4ab-f335efe03ca0 · inbound
DiTReducio: A Training-Free Acceleration for DiT-Based TTS via Progressive Calibration Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 127fb2f9-b8ac-4151-a00b-fa8dde61b793 · inbound
Qwen3-Omni Technical Report Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d7bd7831-89ca-4f66-824a-53913e906523 · inbound
StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a9d33b68-1469-43b4-b45e-0df35f32b1cc · inbound
Ovi: Twin Backbone Cross-Modal Fusion for Audio-Video Generation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ad9c0fc2-96f3-42d0-8681-5988cbee1899 · inbound
UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ad9b056-fa54-43bd-9269-f8e379043065 · inbound
Two-Dimensional Quantization for Geometry-Aware Audio Coding Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4e4ec9b8-24c5-477d-98a4-25da48619be5 · inbound
Qwen3-TTS Technical Report Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 762d14bd-1ad1-4767-9a1c-6ca2513c7d76 · inbound
Edit Content, Preserve Acoustics: Imperceptible Text-Based Speech Editing via Self-Consistency Rewards Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e099125-5d51-42d8-9179-f8b163ee6863 · inbound
Voxtral TTS Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ddf0ac01-a325-4c9b-bd3d-9f2f54c30453 · inbound
OmniVoice: Towards Omnilingual Zero-Shot Text-to-Speech with Diffusion Language Models Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 344fe637-5357-4c17-9763-3bb213d23851 · inbound
Controllable Singing Style Conversion with Boundary-Aware Information Bottleneck Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5923032c-c62d-4621-90d4-f0fac40f1c00 · inbound
CapTalk: Unified Voice Design for Single-Utterance and Dialogue Speech Generation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 93d3680e-622e-4165-aa15-bb465142f1f3 · inbound
WAND: Windowed Attention and Knowledge Distillation for Efficient Autoregressive Text-to-Speech Models Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b45ebf96-f36d-4ff1-8d55-f2885248ae0b · inbound
Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a6a1f961-7ef8-4cc2-ac74-94e3e8085089 · inbound
Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 368dbd12-4a7c-4adc-85a2-480ae8d16224 · inbound
Bridging What the Model Thinks and How It Speaks: Self-Aware Speech Language Models for Expressive Speech Generation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0babbef2-c98d-4d4c-b65d-27c612265560 · inbound
X-VC: Zero-shot Streaming Voice Conversion in Codec Space Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0f264bc0-690c-42dc-bfd4-a656e57854fb · inbound
From Reactive to Proactive: Assessing the Proactivity of Voice Agents via ProVoice-Bench Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5cd30102-480c-45e5-8003-39ec429f79ef · inbound
Qwen3.5-Omni Technical Report Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b3ef62b0-6daf-46df-a116-19ff86208f6b · inbound
MINT-Bench: A Comprehensive Multilingual Benchmark for Instruction-Following Text-to-Speech Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bc83a086-21fb-4ef8-b518-a539e471edfb · inbound
Text-To-Speech with Chain-of-Details: modeling temporal dynamics in speech generation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1d6cf6c1-a7ea-4697-acb7-4c5191f789ed · inbound
MMControl: Unified Multi-Modal Control for Joint Audio-Video Generation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 973c30f0-f12b-4bdd-b61f-d359795cec47 · inbound
JaiTTS: A Thai Voice Cloning Model Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f7a73220-847e-4795-90c8-943f60284c24 · inbound
JaiTTS: A Thai Voice Cloning Model Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 92f9b56f-4056-4c15-a7a8-528731586257 · inbound
Kinetic-Optimal Scheduling with Moment Correction for Metric-Induced Discrete Flow Matching in Zero-Shot Text-to-Speech Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation beb68d03-d8a0-4a6b-83c4-21e8f1661d43 · inbound
Beyond Content: A Comprehensive Speech Toxicity Dataset and Detection Framework Incorporating Paralinguistic Cues Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6c05834e-168f-4850-8991-eee0d38ab236 · inbound
SemaVoice: Semantic-Aware Continuous Autoregressive Speech Synthesis Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 59e97aa2-3ac4-4d1f-b1a7-bde25cad6181 · inbound
Taming Audio VAEs via Target-KL Regularization Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 52ccb11b-2904-498b-b600-24adc1f572da · inbound
Raon-OpenTTS: Open Models and Data for Robust Text-to-Speech Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 283516ce-deb9-4eb0-92d9-253a24097c53 · inbound
RobustSpeechFlow: Learning Robust Text-to-Speech Trajectories via Augmentation-based Contrastive Flow Matching Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 205d5100-80f6-4707-acfb-43c77163227e · inbound
RobustSpeechFlow: Learning Robust Text-to-Speech Trajectories via Augmentation-based Contrastive Flow Matching Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3db25c6d-0be8-49f4-a71a-530a3de1d0b9 · inbound
Raon-Speech Technical Report Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a795d270-d0da-4fde-a878-3e70b2b63207 · inbound
PilotTTS: A Disciplined Modular Recipe for Competitive Speech Synthesis Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e85a2a67-8566-44eb-ba9a-0b7ffd97609e · inbound
Native Audio-Visual Alignment for Generation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d2dc4718-1a32-4d2a-952e-3628a94ac764 · inbound
SwanVoice: Expressive Long-Form Zero-Shot Speech Synthesis for Both Monologue and Dialogue Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cfd54f65-456c-433f-bae3-90776aa463a9 · inbound
UNISON: A Unified Sound Generation and Editing Framework via Deep LLM Fusion Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 27dcc097-dae4-42d1-b873-027a5d2b0953 · inbound
LaSR: Context-Aware Speech Recognition via Latent Reasoning Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 791d3621-378a-4a24-b29b-8342366a8cc7 · inbound
PolySpeech-100: A Large-Scale Benchmark for Speech Understanding Across 100+ Languages and Dialects Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d81efb37-1666-4930-9056-de1de2e10669 · inbound
UniVocal: Unified Speech-Singing Code-Switching Synthesis Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 697c4449-917b-41c8-bb3a-db42a690e70d · inbound
EntangleCodec: A Unified Discrete Audio Tokenizer via Semantic-Acoustic Entanglement Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2ce5ec35-c2ca-49f5-a2e9-a3785716a6e5 · inbound
WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 99a214d0-0985-4774-a150-9dcfd89a5d19 · inbound
Foley-Omni: A Unified Multimodal Generation Model from Task-Level Audio Synthesis to Complete Video Soundtrack Generation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 56cca99c-ac62-4935-844d-3307dc3e73b2 · inbound
CleanCodec: Efficient and Robust Speech Tokenization via Perceptually Guided Encoding Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63bb06b3-a176-42a3-a002-abe3735fdc0a · inbound
GLASS: GRPO-Trained LoRA for Acoustic Style Steering in Zero-Shot Text-to-Speech Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 03a510e2-8b82-4fb0-801f-762f3e0dd610 · inbound
HybridCodec: Fast Dual-Stream, Semantically Enhanced Neural Audio Codec Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation de8cf5ce-c91f-405e-b045-5ea72d02aea4 · inbound
VoxCPM2 Technical Report Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5826c81b-871c-4474-8f24-be0c20aeceb9 · inbound
dots.tts Technical Report Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6a2f00b2-9121-43f4-a782-855895f90337 · inbound
TLDR: Compressing Audio Tokens for Efficient Autoregressive Text-to-Speech Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 08f970a2-ee8b-4e8f-adbe-114e3cb3ea4e · inbound
MeanVC 2: Robust Low-Latency Streaming Zero-Shot Voice Conversion Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5ce2933d-a971-4f6e-84f2-e9b06c8422e7 · inbound
FlashTTS: Fast Streaming TTS with MTP Acceleration and X-pred Mean Flow Distillation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aff23cce-2c94-4fff-a242-33255dcde729 · inbound
End-to-End Training for Discrete Token LLM based TTS System Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 71d79e18-eb6a-4310-a9ef-82b96131562b · inbound
M*: A Modular, Extensible, Serving System for Multimodal Models Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d9dd68d5-dbbb-48fa-8d26-52246c014db0 · inbound
Investigating Human-Model Discrepancies in Speech Quality Assessment via Acoustic and Prosodic Perturbations Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9535b3f4-9a54-4aa9-aef8-edd9b2d215f9 · inbound
Zero-VC: Zero-Lookahead Streaming Voice Conversion via Speaker Anonymization Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5e24eb89-96f4-4e72-a11c-44b7b17e5bc6 · inbound
Transcript-Free Flow-Matching Text-to-Speech via Speech Feature Conditioning Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 503ee912-f86a-42e1-811e-e14b1b17ffde · inbound
Imitation Learning for Elder-Facing Speech Synthesis Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d3913959-4d88-4849-840a-12893bb77b8e · inbound
Bagpiper-Edit: Zero-Shot Open-Ended Audio Editing via Rich-Caption Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4e8f3d24-7753-47da-9c58-309bf989a0a0 · inbound
ProsoCodec: Prosody-Oriented Speech Codec for Voice Conversion Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c90182f8-5fa8-48f0-9058-3fbfa6ba711d · inbound
Bagpiper-TTS: Natural Language Guided Universal Speech Synthesis Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a6aa5307-c326-4407-96aa-9bb43418adbb · inbound
AudioCALM: Continuous Autoregressive Language Modeling for Universal Audio Generation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.