Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T21:10:25.911203Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 1 inbound Pith citation observation for arXiv:2606.07080.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T21:10:25.911203Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T08:35:51.971819Z
A source-named dated measurement, never combined with another source.
Source: cited_works
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a227bf2f-6592-466b-9663-aec08fa98698 · outbound
dots.tts Technical Report OmniVoice: Towards Omnilingual Zero-Shot Text-to-Speech with Diffusion Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d8150420-e9be-4467-9b82-fef24412903f · outbound
dots.tts Technical Report Longcat-audiodit: High-fidelity diffusion text-to-speech in the waveform latent space
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f77e995f-57c2-4b82-baba-14104bf02e86 · outbound
dots.tts Technical Report CosyVoice 3: Towards In-the-wild Speech Generation via Scaling-up and Post-training
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 225c645c-38b0-424d-9aa8-3d8c79502e02 · outbound
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation acec61fb-0e67-4519-b0db-1c0c0d69e4f3 · outbound
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3209b1e8-4ec5-4a6a-975c-a756751fd2ed · outbound
dots.tts Technical Report V oxcpm: Tokenizer-free tts for context-aware speech generation and true-to-life voice cloning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation efba0c9f-6502-4976-9253-af1178854f41 · outbound
dots.tts Technical Report Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 34453ff6-37d0-474a-ba2a-82c01b6763b0 · outbound
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0fa92e1a-041f-4bd0-a519-4c4aa1fe2098 · outbound
dots.tts Technical Report HoliTok:A Coutinuous Holistic Tokenization with Robust Dual Capabilities of Speech Generation and Understanding
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a4cabc65-01f3-4afe-a009-3bd74711a548 · outbound
dots.tts Technical Report SOAR: Self-Correction for Optimal Alignment and Refinement in Diffusion Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5826c81b-871c-4474-8f24-be0c20aeceb9 · outbound
dots.tts Technical Report Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d2ae9d75-0bc9-4f8c-b763-dce5fdfe8bc9 · outbound
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dd263aa1-7927-49e5-9d92-a2a16a60a19e · outbound
dots.tts Technical Report CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4af40350-8541-4072-9114-7128ce4677d3 · outbound
dots.tts Technical Report FireRedTTS-2: Towards Long Conversational Speech Generation for Podcast and Chatbot
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 50b6787f-db3d-4ff6-a9c9-ceeef55efee3 · outbound
dots.tts Technical Report Fish audio s2 technical report
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d62cfc91-7398-45ee-acda-3dc9648f1de1 · outbound
dots.tts Technical Report MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation da31b60e-9d12-4c5a-ab33-74f6270c59ac · outbound
dots.tts Technical Report XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 951413b5-1f55-4636-91a8-be5539f7f333 · outbound
dots.tts Technical Report WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f1567a9c-9f66-4bd0-be6c-4c6bf270299a · outbound
dots.tts Technical Report Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a2908299-c8be-4a81-99a4-116b80418312 · outbound
dots.tts Technical Report Ming-uniaudio: Speech llm for joint understanding, generation and editing with unified representation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 293b4a13-3d26-4be9-aec4-7e6fc43bdeff · outbound
dots.tts Technical Report vllm-omni: Fully disaggregated serving for any-to-any multimodal models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7e2e26bd-231f-4cdb-bcc0-08895295bebb · inbound
Stable Autoregressive Speech Generation with Low-Frame-Rate High-Dimensional Continuous Tokens dots.tts Technical Report
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.