Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:45:17.467537Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 16 inbound Pith citation observations for arXiv:2506.13053.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:45:17.467537Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T07:39:23.947687Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T03:27:34.896349Z
55 of 55 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 46dd6f12-388d-47f2-b4d5-1b58801c4773 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Neural codec language models are zero-shot text to speech synthesizers,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2cbfe618-9751-4ea9-90ad-f4aac7513486 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching V oicebox: Text-guided multilingual universal speech generation at scale,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4840c43-7c3b-4257-980b-e308b9d53508 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching E2 tts: Embarrassingly easy fully non- autoregressive zero-shot tts,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 19e19724-5546-4746-8948-d7be6e8427d4 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4ee4589-42be-40e8-b106-1220f2f921ad · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching MaskGCT: Zero-shot text-to-speech with masked generative codec transformer,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 47614f9a-f3b0-4357-b723-32f725028c9f · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching FireRedTTS: A Foundation Text-To-Speech Framework for Industry-Level Generative Speech Applications
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 811b01f7-0773-4610-a07c-7bf31ea74375 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48f1086b-b0ef-4d23-aeb1-644adc8484d4 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Libritts: A corpus derived from librispeech for text-to-speech,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1a7c5959-0e17-4458-b945-020459b91c96 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Libriheavy: A 50,000 hours asr corpus with punctuation casing and context,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e190dbb3-1fbf-4a1f-9e04-e730369071e3 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Emilia: An extensive, multilingual, and diverse speech dataset for large-scale speech generation,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c4b312cb-f7a2-492f-b2c2-3ea22dbe1a6a · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Naturalspeech 2: Latent diffusion models are natural and zero- shot speech and singing synthesizers,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0648c2e9-ec80-40e7-b40c-854a16fed321 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Flow matching for generative modeling,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c726225b-2ddf-465f-9cf5-1cc4fdbf401a · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Sf- speech: Straightened flow for zero-shot voice clone,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f7b3f264-2b8f-494d-a920-a01aff7949ae · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching P-flow: A fast and data-efficient zero-shot tts through speech prompting,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e3f0334-850f-41da-9b83-dbe729364cf0 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Attention is all you need,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e0b8f32-b682-4c29-ba91-3741fc249547 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Zipformer: A faster and better encoder for automatic speech recognition,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2110306a-b124-47bc-be6c-8bd87d96343e · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Classifier-free diffusion guidance,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 170f12b1-dcbb-4192-bdf7-bbe89d575b1c · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Freeu: Free lunch in diffusion u-net,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4a2bb12a-be56-4f7b-9cf2-5cc4415addbf · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching U-dits: Downsample tokens in u-shaped diffusion transformers,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1fad3369-6cbe-4ed7-be2d-5ae2ffe0c242 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Fastspeech: Fast, robust and controllable text to speech,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94ec5498-4dba-42f3-ae84-7ae06ad8cc7c · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Conformer: Convolution-augmented transformer for speech recognition,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b9194bd8-2d98-4212-8bfe-fd2d1567bdf4 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Glow-tts: A generative flow for text-to-speech via monotonic alignment search,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64bcbf28-fe5b-4eff-a317-75f4b2b627ca · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Flow-tts: A non-autoregressive network for text to speech based on flow,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a84aa383-803f-4188-93d9-7e3c6359d2ab · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Simple- speech: Towards simple and efficient text-to-speech with scalar latent transformer diffusion models,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e5c11d3b-3dc1-4568-b894-f926fe03ec62 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching DiTTo-TTS: Diffusion transformers for scalable text-to-speech without domain-specific factors,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9954c653-8243-48f0-8f7c-bee76df3b892 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Convnext v2: Co-designing and scaling convnets with masked autoencoders,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38ae1325-a8b2-453f-a1ba-4bc4e81cc555 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching On distillation of guided diffusion models,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8a9aeb93-b8ac-45a2-9e03-00166a481c25 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Tacotron: Towards end-to- end speech synthesis,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c94e5b87-c82e-415e-b725-0232d8a39a93 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Natural tts synthesis by conditioning wavenet on mel spectrogram predictions,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40a0df60-3838-450b-ae44-fdff04008690 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Revisiting Over-Smoothness in Text to Speech
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e61555de-6d6c-4486-9be1-8c877b4ac97c · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Grad- tts: A diffusion probabilistic model for text-to-speech,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 52d9923c-d046-424b-b4af-68dee322933b · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Matcha-tts: A fast tts architecture with conditional flow matching,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e5a4718e-8a2d-4a26-be94-da876beb5e8e · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Consistency models,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b152212b-1d78-43b0-bb85-10d3a4ece214 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Flow straight and fast: Learning to generate and transfer data with rectified flow,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f858a989-c3cc-448b-a3f4-7e1ede8df9f3 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Comospeech: One-step speech and singing voice synthesis via consistency model,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e52ae0af-4104-4462-99cb-d18419bc7542 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Reflow- tts: A rectified flow model for high-fidelity text-to-speech,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation da6a2ad3-611e-4da0-9196-71b71c128ca1 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching V oiceflow: Efficient text- to-speech with rectified flow matching,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6e52ee7d-d6cf-49f8-a13f-9e2be46f25db · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Flashspeech: Efficient zero-shot speech synthesis,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fd3bdbbb-af80-4d3d-beef-ca010a89d28b · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Slimspeech: Lightweight and efficient text-to-speech with slim rectified flow,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 92523cae-b210-41cc-ae14-2d9e03326ba3 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Lightspeech: Lightweight and fast text to speech with neural architecture search,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ceb84728-7a54-48fa-8c99-e7d41459232b · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Librispeech-pc: Benchmark for evaluation of punctuation and capitalization capabilities of end-to-end asr models,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a08e7b59-6a11-42a9-b2db-21d3c47cefe3 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Common voice: A massively-multilingual speech corpus,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4bf2354c-ea56-44df-9c1f-b1f82f44ebda · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Didispeech: A large scale mandarin speech corpus,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b922f287-09de-418f-95d3-a394b95e554c · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching V ocos: Closing the gap between time-domain and fourier- based neural vocoders for high-quality audio synthesis,
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66d38499-0d1c-4dba-9295-38bca924f20c · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Robust speech recognition via large-scale weak supervision,
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28641d09-7533-4f30-9ff5-06211cc5fb0e · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Paraformer: Fast and accurate parallel transformer for non-autoregressive end-to-end speech recognition,
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87dc4532-6b2a-4054-85d9-5caaf380e904 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Hubert: Self-supervised speech representation learning by masked prediction of hidden units,
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf4eb6b6-3f38-412d-a71e-47c23e715ee4 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Wavlm: Large-scale self-supervised pre- training for full stack speech processing,
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 523f1dfa-d805-481b-a534-f640a4ae254a · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Ecapa-tdnn: Em- phasized channel attention, propagation and aggregation in tdnn based speaker verification,
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31e12026-c525-468c-a783-f51f3b05bb96 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Utmos: Utokyo-sarulab system for voicemos challenge 2022,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 263e86b0-e18e-426c-8237-e6dcd3010d8e · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80c5d825-271b-4e00-9722-bc71a0a9a635 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45102f47-a413-4735-b698-7e1c2858eaa9 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Spark-TTS: An Efficient LLM-Based Text-to-Speech Model with Single-Stream Decoupled Speech Tokens
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11b0aff7-79ae-4cc5-9a05-3e065b821938 · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Yourtts: Towards zero-shot multi-speaker tts and zero-shot voice conversion for everyone,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation df6b7d1e-4c64-4752-ab52-b75bfd80e94d · outbound
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching Amphion: an open-source audio, music, and speech generation toolkit,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1f256cb3-22a5-4e2f-be92-73228fa1566d · inbound
ZipVoice-Dialog: Non-Autoregressive Spoken Dialogue Generation with Flow Matching ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ab2e20d5-7d2d-4ad0-b454-c474d228d099 · inbound
Universal Speech Content Factorization ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef4f0b28-14ef-48b4-aad7-45836011a635 · inbound
OmniVoice: Towards Omnilingual Zero-Shot Text-to-Speech with Diffusion Language Models ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a7b0fc38-7d06-4288-9379-6e9f4c2ce3d3 · inbound
MMControl: Unified Multi-Modal Control for Joint Audio-Video Generation ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 93f45acc-75fe-4625-954f-7155ac151c43 · inbound
Unison: Harmonizing Motion, Speech, and Sound for Human-Centric Audio-Video Generation ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a148795a-e901-425f-921a-39a80ed99b83 · inbound
Unison: Harmonizing Motion, Speech, and Sound for Human-Centric Audio-Video Generation ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fce86089-99f4-4aab-9876-524fa05de4cd · inbound
From Flat Language Labels to Typological Priors: Structured Language Conditioning for Multilingual Speech-to-Speech Translation ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 456e7331-2449-4af2-a6c5-fa6a9ce4ec35 · inbound
SwanVoice: Expressive Long-Form Zero-Shot Speech Synthesis for Both Monologue and Dialogue ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 35be436f-b14d-462d-9f24-c9059434f62d · inbound
WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 108
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d30edfa3-c12d-4673-9a20-d17c3e8d3216 · inbound
VoxCPM2 Technical Report ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d1c78e43-7078-41ac-acf0-683a37a18d96 · inbound
Towards Unified Song Generation and Singing Voice Conversion with Accompaniment Co-Generation ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a83fb087-1b43-4394-b06b-ed907f6a3bd5 · inbound
End-to-End Training for Discrete Token LLM based TTS System ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7a947619-156c-491b-a449-a4279363051c · inbound
ReGen: Hierarchical Multi-Prompt Representation Generation for Efficient Waveform Diffusion Models ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0297f999-eccb-423d-876a-5296308efd40 · inbound
FreyaTTS: A Compact Tokenizer-Free Flow-Matching Transformer for Turkish-First Speech Synthesis ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e07bb246-f5cc-41a3-9d83-cc58e8eec40d · inbound
FreyaTTS: A Compact Tokenizer-Free Flow-Matching Transformer for Turkish-First Speech Synthesis ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51c0f031-9d9e-4949-9112-f3042e5cf35e · inbound
Zero-Shot Face-to-Speech Synthesis via Latent Space Adaptation of a Style-Diffusion TTS Model ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.