Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:51:37.449634Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 1 inbound Pith citation observation for arXiv:2509.21597.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:51:37.449634Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-09T19:27:59.124425Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T15:41:30.302351Z
58 of 58 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3c9df119-d0da-46cd-9b1f-c3eb531b61da · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Audio Deepfake Detection: A Survey
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7803ce74-c139-4053-b630-e7c77203e743 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Audio deepfake detection: What has been achieved and what lies ahead,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation daeacacb-2cc2-4614-975f-5f4d40dbb71f · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Deepfake-Eval-2024: A Multi-Modal In-the-Wild Benchmark of Deepfakes Circulated in 2024
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b5dfd74-a37f-4645-b841-e6705d88297d · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Spoofceleb: Speech deepfake detection and sasv in the wild,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 959177e2-5a88-4ebb-949c-8cdab8606487 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Asvspoof 2019: Spoofing countermeasures for the detection of synthesized, converted and replayed speech,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 689b33e3-af25-4e09-9f48-ed9f5621adc3 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors ASVspoof 2021: accelerating progress in spoofed and deepfake speech detection
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5394366-8729-4453-a7cf-1fec96a86e48 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Asvspoof 5: Design, collection and validation of resources for spoofing, deepfake, and adversarial attack detection using crowdsourced speech,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6aa475e6-1940-411b-9759-3efdb5da4073 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Does audio deepfake detection generalize?
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2abf36bc-24b1-4814-9291-8e5ed8651151 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Mlaad: The multi- language audio anti-spoofing dataset,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 20d4f0b2-f689-4445-9e1e-7607c4ab999f · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors CodecFake: Enhancing Anti-Spoofing Models Against Deepfake Audios from Codec-Based Speech Synthesis Systems
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80c37585-68ca-4b7f-9bca-babb7a807a39 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors CodecFake+: Codec-Based Resynthesized Data as a Proxy for Detecting CodecFake Speech
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a40b0a1-ae04-482f-a161-188f55999bd6 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Diffssd: A diffusion-based dataset for speech forensics,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation db553c61-84cc-4a41-9887-edd3f42a4fe6 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Diffuse or confuse: A diffusion deepfake speech dataset,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c7921486-6ef8-4697-9821-463a7b2e947a · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Habla: A dataset of latin american spanish accents for voice anti-spoofing,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ae27e744-b168-47f9-aba9-fd5ba83d3ff8 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Replay Attacks Against Audio Deepfake Detection
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cd32478-3282-47a3-a036-faf95705d651 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors The codecfake dataset and countermeasures for the universally detection of deepfake audio,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 190b633b-ee2d-4bdf-b3f0-30a3331aede6 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors WaveFake: A Data Set to Facilitate Audio Deepfake Detection
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ddab8a6-2d52-4f41-9a13-0bfbe04de403 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Safeear: Content privacy-preserving audio deepfake detection,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7176dc27-c6ad-4d45-a56c-bc547b9f5a3f · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Ditse: High-fidelity generative speech enhancement via latent diffusion trans- formers,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d71bcd26-412b-422f-9fa1-38c10417c075 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Generalized end-to-end loss for speaker verification,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f48bc8e-52c2-44c5-be0d-f3a8162e9817 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Tacotron: Towards End-to-End Speech Synthesis
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01e687d7-9627-4e0e-a149-cae8520a6312 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Towards end-to-end prosody transfer for expressive speech synthesis with tacotron,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 43cd36db-8f20-456f-ac34-67eab62b2160 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Fastspeech: Fast, robust and controllable text to speech,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80ba89bc-1bef-46a9-97a6-e5657b0b5713 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors FastSpeech 2: Fast and High-Quality End-to-End Text to Speech
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69604e31-cdde-4e76-86d2-65fdb18466d1 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors For: A dataset for synthetic speech detection,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6977245a-bab6-4d85-b7af-d14001b7d912 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Yourtts: Towards zero-shot multi-speaker tts and zero- shot voice conversion for everyone,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 25d40685-ba5c-4719-8b70-deced384f34b · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Naturalspeech: End-to-end text-to-speech synthesis with human-level quality,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e0e62d68-ae4f-454e-ae2f-1f2c1f514cfd · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e6337bc-773f-4dfa-8355-23c7b47772a9 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors SRC4VC: Smartphone-Recorded Corpus for Voice Conversion Benchmark
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb91aa34-5c25-4e61-b291-2d576eb2a0f1 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Timit- tts: A text-to-speech dataset for multimodal synthetic media detection,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f5e2585a-e1b2-4ff9-93dc-985f5b2e598b · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04b0ec63-b22c-40ef-8119-78b31c7fbdfa · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1310b110-fb69-4d15-97c3-6f560284ff91 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Dfadd: The diffusion and flow-matching based audio deepfake dataset,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7021fb36-88db-41d8-98da-7a9eea181cdb · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1e8ed99-865a-46d0-b157-dc3c692119e8 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Discrete audio tokens: More than a survey!
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2720178-eab5-43be-8c9c-a2ed7f31df33 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Ai-synthesized voice detection using neural vocoder artifacts,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e841972b-d1d4-4742-a653-641994ef2381 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Post-training for deepfake speech detection,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6025f3e4-a62e-4f81-a2d5-3a49818cc163 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors CtrSVDD: A Benchmark Dataset and Baseline Analysis for Controlled Singing Voice Deepfake Detection
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f28a755-311c-43a2-9ced-246b0f1e4cc4 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Deepfake cross-lingual evaluation dataset (decro),
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 61e8c6c1-64a9-4b08-b3f5-ccf71e95c799 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors MSceneSpeech: A Multi-Scene Speech Dataset For Expressive Speech Synthesis
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 354a086e-c37b-4353-a736-343d1471398e · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors An open dataset of synthetic speech,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3dd39851-7c36-44dd-a2ad-f86a738dedb7 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Audio recordings dataset of genuine and replayed speech at both ends of a telecommunication channel,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2954ff89-68d4-435a-9000-0081ebccd861 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Jvnv: A corpus of japanese emotional speech with verbal content and nonverbal expressions,
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c865ca0-694a-4acd-b7e7-71d5530be07d · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Speech enhancement and dereverberation with diffusion-based generative models,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e5cf11cd-a2f9-4153-869e-b30dca7a113a · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Genhancer: High-fidelity speech enhancement via generative modeling on discrete codec tokens,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1a55586c-d0a7-4478-a5ef-f492be13aad4 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Miipher: A robust speech restora- tion model integrating self-supervised speech and text representations,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b79cc746-2c38-40b5-af86-bcc9eaf3337f · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Hifi-gan-2: Studio-quality speech en- hancement via generative adversarial networks conditioned on acoustic features,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation cb6ef1cb-1b67-401f-aae5-53ec15948686 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors The lj speech dataset,
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2aaabab3-dfa5-4341-9f5c-253eff4f6d21 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Automatic speaker verification spoofing and deepfake detection using wav2vec 2.0 and data augmentation,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation fde5178a-8192-484f-bfbb-65f028905928 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors wav2vec 2.0: A framework for self-supervised learning of speech representations,
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0b3d28a-3653-4618-82c8-d13853caf0bf · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors End-to-end spectro-temporal graph attention networks for speaker verification anti-spoofing and speech deepfake detection,
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 17cd747c-aeb8-4357-8892-6e3bd4ff3a21 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Speech enhancement—a review of modern meth- ods,
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d7e482e8-561b-499d-adaf-844fa79f4b39 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Advances in Speech Separation: Techniques, Challenges, and Future Trends
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c88548e1-a921-4747-a9c7-3097766cac45 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Cstr vctk corpus: English multi-speaker corpus for cstr voice cloning toolkit (version 0.92),
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7345fffa-9fe2-4ad4-b07f-1afae1aa4b15 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Available: https://api.semanticscholar.org/CorpusID: 213060286 10 VOLUME ,
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation cee372bc-1581-4c38-be1a-50762573c827 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1b0f6f4-4e2e-4053-bfad-138d0e188ad0 · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Available: https://doi.org/10.5281/zenodo.7603208
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d06eb128-b672-428e-91b5-472e7890b23c · outbound
AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors The Codecfake Dataset and Countermeasures for the Universally Detection of Deepfake Audio
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 443d723c-b3cb-413d-8d98-66e5d4368898 · inbound
Alethia: A Foundational Encoder for Voice Deepfakes AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.