Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2305.07243.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T22:30:26.989835Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T20:10:07.252358Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation fbece7d4-c200-461f-a8ae-be65655fed70 · inbound
MLAAD: The Multi-Language Audio Anti-Spoofing Dataset Better speech synthesis through scaling
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 36c79e26-f01d-4b1e-954a-a6c8070fc732 · inbound
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models Better speech synthesis through scaling
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4e19c775-5283-49ee-b401-6287a4341f94 · inbound
GenVC: Self-Supervised Zero-Shot Voice Conversion Better speech synthesis through scaling
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecf8f860-9809-4c0e-8dfd-c0805f94c1a2 · inbound
IndexTTS: An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System Better speech synthesis through scaling
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54dc202b-f995-44b5-9425-fc2300fe4d46 · inbound
Vevo: Controllable Zero-Shot Voice Imitation with Self-Supervised Disentanglement Better speech synthesis through scaling
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c64258b-9b03-4a75-b2c2-c5ce3c09c76a · inbound
DeePen: Penetration Testing for Audio Deepfake Detection Better speech synthesis through scaling
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4e51cc1d-d6a3-47a1-9995-9a7b2bd0e0f2 · inbound
MPE-TTS: Customized Emotion Zero-Shot Text-To-Speech Using Multi-Modal Prompt Better speech synthesis through scaling
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9951754c-b7b0-4e2a-9983-da0f89676296 · inbound
Semantics-Aware Human Motion Generation from Audio Instructions Better speech synthesis through scaling
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40e12eec-8569-4de8-89b0-2d584479c018 · inbound
StreamFlow: Streaming Flow Matching with Block-wise Guided Attention Mask for Speech Token Decoding Better speech synthesis through scaling
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efa6ef73-0ed5-4a80-8573-fa8b4f8c2b25 · inbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Better speech synthesis through scaling
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b03c3b1c-9fbf-4f63-b925-c5162c052d5f · inbound
De-AntiFake: Rethinking the Protective Perturbations Against Voice Cloning Attacks Better speech synthesis through scaling
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edcddf70-53f9-47ca-b41a-b5fb3e02e6dd · inbound
Quantize More, Lose Less: Autoregressive Generation from Residually Quantized Speech Representations Better speech synthesis through scaling
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b830277d-a5b9-48e4-8861-901a9974e525 · inbound
Step-Audio 2 Technical Report Better speech synthesis through scaling
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9f8eb8eb-c408-4fb3-bb91-927714cba40b · inbound
JWB-DH-V1: Benchmark for Joint Whole-Body Talking Avatar and Speech Generation Version 1 Better speech synthesis through scaling
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 448ecfc7-dea8-42b5-b8e6-c5ee136542be · inbound
TTS-1 Technical Report Better speech synthesis through scaling
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42bed4f2-64cc-40a0-803e-8e9e71bf4a9b · inbound
SpeechFake: A Large-Scale Multilingual Speech Deepfake Dataset Incorporating Cutting-Edge Generation Methods Better speech synthesis through scaling
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6229081c-1426-49c8-90a8-da36704ba759 · inbound
Next Tokens Denoising for Speech Synthesis Better speech synthesis through scaling
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 980f1468-0f6e-48d7-bd95-59f812a4ea20 · inbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Better speech synthesis through scaling
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5bf894b-6301-462a-a9bb-ff8ace21d634 · inbound
Large Language Model Data Generation for Enhanced Intent Recognition in German Speech Better speech synthesis through scaling
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 314b1853-2332-4c49-a51b-96a09c774be3 · inbound
Making Separation-First Multi-Stream Audio Watermarking Feasible via Joint Training Better speech synthesis through scaling
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1011716-3d8a-44e4-a33f-e9b35a272e03 · inbound
AT-ADD: All-Type Audio Deepfake Detection Challenge Evaluation Plan Better speech synthesis through scaling
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a759e3db-93c3-416b-b1be-c4cac265d7d1 · inbound
Enhancing Conversational TTS with Cascaded Prompting and ICL-Based Online Reinforcement Learning Better speech synthesis through scaling
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d5d29c86-d4cd-4850-963e-10ca7a84a21b · inbound
X-Voice: Enabling Everyone to Speak 30 Languages via Zero-Shot Cross-Lingual Voice Cloning Better speech synthesis through scaling
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ecca9889-85c2-48b9-b1e4-0094a650bd7a · inbound
X-Voice: Enabling Everyone to Speak 30 Languages via Zero-Shot Cross-Lingual Voice Cloning Better speech synthesis through scaling
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 994559ed-963d-4afc-a376-75312646c751 · inbound
FlashTTS: Fast Streaming TTS with MTP Acceleration and X-pred Mean Flow Distillation Better speech synthesis through scaling
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 39f7ca8a-ae34-4a16-aefd-8bbc381cc9bd · inbound
An Evaluation Framework for Text-to-Speech Voice Reconstruction Better speech synthesis through scaling
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dfa2c6d2-65a8-4f9f-90d2-750168104c71 · inbound
Sarashina2.2-TTS: Tackling Kanji Polyphony in Japanese Speech Generation via Data Scaling and Targeted Data Synthesis Better speech synthesis through scaling
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d9687248-44c7-475a-9473-379c1881c1da · inbound
FlexiSLM: A Dynamic and Controllable Frame Rate Spoken Language Model Better speech synthesis through scaling
Reference 113
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.