Pith. sign in

Paper Citation Record · LEDGER

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors

As of 21 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 1 inbound Pith citation observation for arXiv:2509.21597.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.21597 v2

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:51:37.449634Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-09T19:27:59.124425Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T15:41:30.302351Z

Reference resolution

58 of 58 outbound references displayed

  • verified exact1
  • verified fuzzy27
  • unresolved29
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3c9df119-d0da-46cd-9b1f-c3eb531b61da · outbound

This paper cites Audio Deepfake Detection: A Survey.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Audio Deepfake Detection: A Survey

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.252307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.252307Z digest=sha256:b6b9bb164424757b8c9425168b5e2517b3f5ca7f16b7e9d13307be14649331ee

Observation 7803ce74-c139-4053-b630-e7c77203e743 · outbound

This paper cites Audio deepfake detection: What has been achieved and what lies ahead,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Audio deepfake detection: What has been achieved and what lies ahead,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.323615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.256670Z digest=sha256:bd23cb3e748e8887d4858819a75fed7e4685896be9c7cda4fbf25f324ca03c07

Observation daeacacb-2cc2-4614-975f-5f4d40dbb71f · outbound

This paper cites Deepfake-Eval-2024: A Multi-Modal In-the-Wild Benchmark of Deepfakes Circulated in 2024.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Deepfake-Eval-2024: A Multi-Modal In-the-Wild Benchmark of Deepfakes Circulated in 2024

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.260136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.260136Z digest=sha256:3c3c06dd8429e5ce656470b89e062ed8faec70441abef5a74ffc2138a087906a

Observation 5b5dfd74-a37f-4645-b841-e6705d88297d · outbound

This paper cites Spoofceleb: Speech deepfake detection and sasv in the wild,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Spoofceleb: Speech deepfake detection and sasv in the wild,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.313456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.264346Z digest=sha256:b4921d31b71299706698a5a6f539738e35d1c2a50737f6a42f321d1d2912546f

Observation 959177e2-5a88-4ebb-949c-8cdab8606487 · outbound

This paper cites Asvspoof 2019: Spoofing countermeasures for the detection of synthesized, converted and replayed speech,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Asvspoof 2019: Spoofing countermeasures for the detection of synthesized, converted and replayed speech,

Reference 5

Resolution
malformed identifier
no resolver link, observed 2026-08-15T15:51:37.268241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.268241Z digest=sha256:7547ac29a378ffaa63ed2bc11a25e0f1d08939e1ebf47b1ddbd4a3c3a0990f02

Observation 689b33e3-af25-4e09-9f48-ed9f5621adc3 · outbound

This paper cites ASVspoof 2021: accelerating progress in spoofed and deepfake speech detection.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors ASVspoof 2021: accelerating progress in spoofed and deepfake speech detection

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.271752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.271752Z digest=sha256:8645e4268ad72fe4ec0d022dd0d7c5a166b8c2604568fb35c48a659266b0756f

Observation e5394366-8729-4453-a7cf-1fec96a86e48 · outbound

This paper cites Asvspoof 5: Design, collection and validation of resources for spoofing, deepfake, and adversarial attack detection using crowdsourced speech,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Asvspoof 5: Design, collection and validation of resources for spoofing, deepfake, and adversarial attack detection using crowdsourced speech,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.301896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.275797Z digest=sha256:07cf338bc6967776eb0b007d818c2cc7e6b91d478b7cbfe92e62208991cfa73a

Observation 6aa475e6-1940-411b-9759-3efdb5da4073 · outbound

This paper cites Does audio deepfake detection generalize?.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Does audio deepfake detection generalize?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.279230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.279230Z digest=sha256:494273d868b06a6bba0c044f7b8a4c3206b01128d731adba69de0ab9b0dbe9b6

Observation 2abf36bc-24b1-4814-9291-8e5ed8651151 · outbound

This paper cites Mlaad: The multi- language audio anti-spoofing dataset,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Mlaad: The multi- language audio anti-spoofing dataset,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.290535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.282571Z digest=sha256:e7edb6cce805882dc4ce3e1648f9ed323cb0bfde8a73a140e053ac7d4f13db61

Observation 20d4f0b2-f689-4445-9e1e-7607c4ab999f · outbound

This paper cites CodecFake: Enhancing Anti-Spoofing Models Against Deepfake Audios from Codec-Based Speech Synthesis Systems.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors CodecFake: Enhancing Anti-Spoofing Models Against Deepfake Audios from Codec-Based Speech Synthesis Systems

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.285668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.285668Z digest=sha256:13e6c531ac25d5cd835665d69b9cc0a2f190d70f134c211d9d68cd07f63c342e

Observation 80c37585-68ca-4b7f-9bca-babb7a807a39 · outbound

This paper cites CodecFake+: Codec-Based Resynthesized Data as a Proxy for Detecting CodecFake Speech.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors CodecFake+: Codec-Based Resynthesized Data as a Proxy for Detecting CodecFake Speech

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.289233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.289233Z digest=sha256:ced2480917a0affda01ace0a0f9fae192864258234d765d55dd323e80898468c

Observation 9a40b0a1-ae04-482f-a161-188f55999bd6 · outbound

This paper cites Diffssd: A diffusion-based dataset for speech forensics,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Diffssd: A diffusion-based dataset for speech forensics,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.277902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.292888Z digest=sha256:99ef8449199a776b42bc078bc885df001cfbd3660c5ff49037eeef0a3a1115bd

Observation db553c61-84cc-4a41-9887-edd3f42a4fe6 · outbound

This paper cites Diffuse or confuse: A diffusion deepfake speech dataset,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Diffuse or confuse: A diffusion deepfake speech dataset,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.266654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.296218Z digest=sha256:ed71c4be2167b4680945075e235af22d15d6ee95e611a63ce6177421bf2f15c7

Observation c7921486-6ef8-4697-9821-463a7b2e947a · outbound

This paper cites Habla: A dataset of latin american spanish accents for voice anti-spoofing,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Habla: A dataset of latin american spanish accents for voice anti-spoofing,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.256100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.299457Z digest=sha256:2ff12831a2848aab5faef2032b0067fb521330389c3cf09dfe98f55acc39d2b4

Observation ae27e744-b168-47f9-aba9-fd5ba83d3ff8 · outbound

This paper cites Replay Attacks Against Audio Deepfake Detection.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Replay Attacks Against Audio Deepfake Detection

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.302670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.302670Z digest=sha256:8cd59a9af6db46494bf331ea9557d9d63948a2ea5b1c8500728d2bea8a28e609

Observation 8cd32478-3282-47a3-a036-faf95705d651 · outbound

This paper cites The codecfake dataset and countermeasures for the universally detection of deepfake audio,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors The codecfake dataset and countermeasures for the universally detection of deepfake audio,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.245059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.306798Z digest=sha256:05711bd3eeb8518fda2f67c1d6f06b9c49e73e2d8e9519bedfe39d539b3be3c1

Observation 190b633b-ee2d-4bdf-b3f0-30a3331aede6 · outbound

This paper cites WaveFake: A Data Set to Facilitate Audio Deepfake Detection.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors WaveFake: A Data Set to Facilitate Audio Deepfake Detection

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.314175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.314175Z digest=sha256:b6e2e5e53218f0d5b83114ffbf2aafc0ea549b6c839520d54a6d83f0ca2fb8dd

Observation 0ddab8a6-2d52-4f41-9a13-0bfbe04de403 · outbound

This paper cites Safeear: Content privacy-preserving audio deepfake detection,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Safeear: Content privacy-preserving audio deepfake detection,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.233601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.317810Z digest=sha256:252d8fecce310441d45b3e6c80f345decc9145bcb3e603b7c992b12b083c45ed

Observation 7176dc27-c6ad-4d45-a56c-bc547b9f5a3f · outbound

This paper cites Ditse: High-fidelity generative speech enhancement via latent diffusion trans- formers,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Ditse: High-fidelity generative speech enhancement via latent diffusion trans- formers,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.321547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.321547Z digest=sha256:22c77d743cc284f01a31c52777b5f07367c25018fad96431bcc78fbfbed98b73

Observation d71bcd26-412b-422f-9fa1-38c10417c075 · outbound

This paper cites Generalized end-to-end loss for speaker verification,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Generalized end-to-end loss for speaker verification,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.324962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.324962Z digest=sha256:49a1f8a121de2bc081131397697eff54217a6c845f436f5fe4ce0b39e4c9679b

Observation 2f48bc8e-52c2-44c5-be0d-f3a8162e9817 · outbound

This paper cites Tacotron: Towards End-to-End Speech Synthesis.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Tacotron: Towards End-to-End Speech Synthesis

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.328425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.328425Z digest=sha256:3f14540dde15e80b7695b3558025755710c3fc85e24d8114a34eaeedaf565b3a

Observation 01e687d7-9627-4e0e-a149-cae8520a6312 · outbound

This paper cites Towards end-to-end prosody transfer for expressive speech synthesis with tacotron,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Towards end-to-end prosody transfer for expressive speech synthesis with tacotron,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.215024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.332018Z digest=sha256:811b8094eb157f4fb808eae2b2a69d7ec9577a01a1dbdf3fa3bc27a2cfe6a7fd

Observation 43cd36db-8f20-456f-ac34-67eab62b2160 · outbound

This paper cites Fastspeech: Fast, robust and controllable text to speech,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Fastspeech: Fast, robust and controllable text to speech,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.335449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.335449Z digest=sha256:4abe8a09ed0b2a1568c1fbf85113f6c56e42a5a27dbd173a4b4d2d4ac67ed785

Observation 80ba89bc-1bef-46a9-97a6-e5657b0b5713 · outbound

This paper cites FastSpeech 2: Fast and High-Quality End-to-End Text to Speech.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors FastSpeech 2: Fast and High-Quality End-to-End Text to Speech

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.338770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.338770Z digest=sha256:21bf4ec934c1413a1b50ba10ce8094655de51b438e5573b60cf7a426e18ba000

Observation 69604e31-cdde-4e76-86d2-65fdb18466d1 · outbound

This paper cites For: A dataset for synthetic speech detection,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors For: A dataset for synthetic speech detection,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.196154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.342135Z digest=sha256:581bbd3df2464e4a5f7369ca59d2cd22a8dd8f468177ae0d8e76fda5685a641f

Observation 6977245a-bab6-4d85-b7af-d14001b7d912 · outbound

This paper cites Yourtts: Towards zero-shot multi-speaker tts and zero- shot voice conversion for everyone,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Yourtts: Towards zero-shot multi-speaker tts and zero- shot voice conversion for everyone,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.185128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.345585Z digest=sha256:ae3dff01e4196d5ee3d8e12ba21940a85f5fd190d440b19627ee54495eed16d2

Observation 25d40685-ba5c-4719-8b70-deced384f34b · outbound

This paper cites Naturalspeech: End-to-end text-to-speech synthesis with human-level quality,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Naturalspeech: End-to-end text-to-speech synthesis with human-level quality,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.174149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.349111Z digest=sha256:64ab578f93fad383bdf736276c8d446ea2875bbe57d463fa81aa059783f1985b

Observation e0e62d68-ae4f-454e-ae2f-1f2c1f514cfd · outbound

This paper cites Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.352429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.352429Z digest=sha256:7c2d9dea9ec10f406cb9a9dfd8eccc9a66b46ac72a0b0168310c5acc2925af7f

Observation 5e6337bc-773f-4dfa-8355-23c7b47772a9 · outbound

This paper cites SRC4VC: Smartphone-Recorded Corpus for Voice Conversion Benchmark.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors SRC4VC: Smartphone-Recorded Corpus for Voice Conversion Benchmark

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.356183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.356183Z digest=sha256:deec120c8c326a3ec43522e06934b3789605418b346062893d03ecf874d7aeeb

Observation eb91aa34-5c25-4e61-b291-2d576eb2a0f1 · outbound

This paper cites Timit- tts: A text-to-speech dataset for multimodal synthetic media detection,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Timit- tts: A text-to-speech dataset for multimodal synthetic media detection,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.162599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.359658Z digest=sha256:b840f557403361732bd581a7b6ffe77114ce29bae01f2ed6e38fc310b8bc4a0a

Observation f5e2585a-e1b2-4ff9-93dc-985f5b2e598b · outbound

This paper cites NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.363039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.363039Z digest=sha256:b207ace9065a9cd1f0c5615facfe33134a9c479a25507940964e8cd19acdcab5

Observation 04b0ec63-b22c-40ef-8119-78b31c7fbdfa · outbound

This paper cites CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.366532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.366532Z digest=sha256:ac59acbdef216e066a816411331d5e0f1603c52b8d2ca3c3fcc0c30965bbd2f8

Observation 1310b110-fb69-4d15-97c3-6f560284ff91 · outbound

This paper cites Dfadd: The diffusion and flow-matching based audio deepfake dataset,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Dfadd: The diffusion and flow-matching based audio deepfake dataset,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.150973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.369880Z digest=sha256:6b42d6faf97bb1c75a61f92bb3fa76919d2081061785f95418e9d9284965abb6

Observation 7021fb36-88db-41d8-98da-7a9eea181cdb · outbound

This paper cites Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.373108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.373108Z digest=sha256:d58b9377fcc14e7e4200348c7aba400d73af5aee5cb7de18444c00c5047e9c75

Observation d1e8ed99-865a-46d0-b157-dc3c692119e8 · outbound

This paper cites Discrete audio tokens: More than a survey!.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Discrete audio tokens: More than a survey!

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.376740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.376740Z digest=sha256:ed1b6954506877bc2842cfe7cea4764200422587b4cec8c8eb44ad62b366bdd9

Observation c2720178-eab5-43be-8c9c-a2ed7f31df33 · outbound

This paper cites Ai-synthesized voice detection using neural vocoder artifacts,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Ai-synthesized voice detection using neural vocoder artifacts,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.140583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.379881Z digest=sha256:3824cbf2fa491308b24448de4f9b531ce3521aa900221f7775db9e766c88f5a5

Observation e841972b-d1d4-4742-a653-641994ef2381 · outbound

This paper cites Post-training for deepfake speech detection,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Post-training for deepfake speech detection,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.382958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.382958Z digest=sha256:79c8dd1b2ce8fb5718dceb389ba65b0c924176b8584e711f9bd15112f5e6e327

Observation 6025f3e4-a62e-4f81-a2d5-3a49818cc163 · outbound

This paper cites CtrSVDD: A Benchmark Dataset and Baseline Analysis for Controlled Singing Voice Deepfake Detection.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors CtrSVDD: A Benchmark Dataset and Baseline Analysis for Controlled Singing Voice Deepfake Detection

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.385915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.385915Z digest=sha256:58fbf59db602e885b56d29603d34629eba038ea30e2b255291977c5b966de915

Observation 8f28a755-311c-43a2-9ced-246b0f1e4cc4 · outbound

This paper cites Deepfake cross-lingual evaluation dataset (decro),.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Deepfake cross-lingual evaluation dataset (decro),

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.129069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.389412Z digest=sha256:fb71d628e8a140f951ff86fb545af15038d59769c02420c6d258022a5836bc2a

Observation 61e8c6c1-64a9-4b08-b3f5-ccf71e95c799 · outbound

This paper cites MSceneSpeech: A Multi-Scene Speech Dataset For Expressive Speech Synthesis.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors MSceneSpeech: A Multi-Scene Speech Dataset For Expressive Speech Synthesis

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.396304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.396304Z digest=sha256:2bbcd4736f596c31e3fca5d2d9e81092bd695b40b7a6815d230f77677235a5ad

Observation 354a086e-c37b-4353-a736-343d1471398e · outbound

This paper cites An open dataset of synthetic speech,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors An open dataset of synthetic speech,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.118245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.399755Z digest=sha256:28f097914f346371c83ed58ac917370e0ee84e5c59f9c9ceb51cc8ce4244760d

Observation 3dd39851-7c36-44dd-a2ad-f86a738dedb7 · outbound

This paper cites Audio recordings dataset of genuine and replayed speech at both ends of a telecommunication channel,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Audio recordings dataset of genuine and replayed speech at both ends of a telecommunication channel,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.107058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.403319Z digest=sha256:6f17936b190187e0282195b619734469e843278c2540fdcb2106a749a87ba779

Observation 2954ff89-68d4-435a-9000-0081ebccd861 · outbound

This paper cites Jvnv: A corpus of japanese emotional speech with verbal content and nonverbal expressions,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Jvnv: A corpus of japanese emotional speech with verbal content and nonverbal expressions,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.406574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.406574Z digest=sha256:6eea5a7b34d029e894dfabb59a1cd7e5d0a61174f4de01bc9ca494aba954b2f7

Observation 7c865ca0-694a-4acd-b7e7-71d5530be07d · outbound

This paper cites Speech enhancement and dereverberation with diffusion-based generative models,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Speech enhancement and dereverberation with diffusion-based generative models,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.088162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.409819Z digest=sha256:1b1ef83ba247e7fd8f207919deec0e5623e8e7145573f395aa28f9056d2c088a

Observation e5cf11cd-a2f9-4153-869e-b30dca7a113a · outbound

This paper cites Genhancer: High-fidelity speech enhancement via generative modeling on discrete codec tokens,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Genhancer: High-fidelity speech enhancement via generative modeling on discrete codec tokens,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.076965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.413049Z digest=sha256:e228f40918fe8fe9fe421660080c1518b6ca9b34158e3e2d6b69618449fecefd

Observation 1a55586c-d0a7-4478-a5ef-f492be13aad4 · outbound

This paper cites Miipher: A robust speech restora- tion model integrating self-supervised speech and text representations,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Miipher: A robust speech restora- tion model integrating self-supervised speech and text representations,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.065384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.416107Z digest=sha256:6989b3e2ae0152f59366c1d3431240a3b8ed0302a407f30ca1af9338df41a375

Observation b79cc746-2c38-40b5-af86-bcc9eaf3337f · outbound

This paper cites Hifi-gan-2: Studio-quality speech en- hancement via generative adversarial networks conditioned on acoustic features,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Hifi-gan-2: Studio-quality speech en- hancement via generative adversarial networks conditioned on acoustic features,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.053765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.419401Z digest=sha256:086e4f1df638b615fb6de06cec129e463f1b35d58a9cfc8826d6efac0ae146e4

Observation cb6ef1cb-1b67-401f-aae5-53ec15948686 · outbound

This paper cites The lj speech dataset,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors The lj speech dataset,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.422757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.422757Z digest=sha256:c235a7849f67ac02a3f5958122aa7259f213525c0b3881415f85b86dccc9c217

Observation 2aaabab3-dfa5-4341-9f5c-253eff4f6d21 · outbound

This paper cites Automatic speaker verification spoofing and deepfake detection using wav2vec 2.0 and data augmentation,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Automatic speaker verification spoofing and deepfake detection using wav2vec 2.0 and data augmentation,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.035337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.426100Z digest=sha256:93f88bfd8daff51ae04fdf8ef97b773b36c1d5ca0649ec39b511f97e69e6278f

Observation fde5178a-8192-484f-bfbb-65f028905928 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors wav2vec 2.0: A framework for self-supervised learning of speech representations,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.429384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.429384Z digest=sha256:180fb6fd4a073d3ae5adedec6f9331eb87778786cdfae19359c7ca6a65a7fa80

Observation e0b3d28a-3653-4618-82c8-d13853caf0bf · outbound

This paper cites End-to-end spectro-temporal graph attention networks for speaker verification anti-spoofing and speech deepfake detection,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors End-to-end spectro-temporal graph attention networks for speaker verification anti-spoofing and speech deepfake detection,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.015961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.436227Z digest=sha256:51d938bc60f74e68cf8ad6f52b9e4f13b51c26390e3f584180debe69dff71e02

Observation 17cd747c-aeb8-4357-8892-6e3bd4ff3a21 · outbound

This paper cites Speech enhancement—a review of modern meth- ods,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Speech enhancement—a review of modern meth- ods,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:38.003891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.439491Z digest=sha256:0cc40329152c1014f9857c403bdc9a6ec099e4fa84a20117e86da12d0afd486a

Observation d7e482e8-561b-499d-adaf-844fa79f4b39 · outbound

This paper cites Advances in Speech Separation: Techniques, Challenges, and Future Trends.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Advances in Speech Separation: Techniques, Challenges, and Future Trends

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.442667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.442667Z digest=sha256:bc649b7e0f69cf2497087d212f146c36109e9cf2c5c6972ae950a0faa415874b

Observation c88548e1-a921-4747-a9c7-3097766cac45 · outbound

This paper cites Cstr vctk corpus: English multi-speaker corpus for cstr voice cloning toolkit (version 0.92),.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Cstr vctk corpus: English multi-speaker corpus for cstr voice cloning toolkit (version 0.92),

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.445977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.445977Z digest=sha256:5faebe52821fc56c74531882694ab65d8dd1ad51c50bd51f4eb8c227eb931981

Observation 7345fffa-9fe2-4ad4-b07f-1afae1aa4b15 · outbound

This paper cites Available: https://api.semanticscholar.org/CorpusID: 213060286 10 VOLUME ,.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Available: https://api.semanticscholar.org/CorpusID: 213060286 10 VOLUME ,

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:37.983918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.449634Z digest=sha256:3ecfe3d6c6a4dc6dcddaca9fb2e2c956a6545906a0ae75b28a6a1c4d857e7dcb

Observation cee372bc-1581-4c38-be1a-50762573c827 · outbound

This paper cites wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.432808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.432808Z digest=sha256:35e09e124a355628e83cf2849262c563fe6ff62d3eabf3bb0f2ec1f167524722

Observation e1b0f6f4-4e2e-4053-bfad-138d0e188ad0 · outbound

This paper cites Available: https://doi.org/10.5281/zenodo.7603208.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors Available: https://doi.org/10.5281/zenodo.7603208

Reference 2023

Resolution
verified exact
doi, observed 2026-08-15T15:51:37.481241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T15:51:37.392955Z digest=sha256:8d49656d49d631addfe33ff7b10512b25c1ed386a96462d6e910bcf988f90097

Observation d06eb128-b672-428e-91b5-472e7890b23c · outbound

This paper cites The Codecfake Dataset and Countermeasures for the Universally Detection of Deepfake Audio.

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors The Codecfake Dataset and Countermeasures for the Universally Detection of Deepfake Audio

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:37.310253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:37.310253Z digest=sha256:55550627c77a1ad7d8ab55b08f910672872afea2cdbdee35ee178b6d94b03882

Pith citing papers

Observation 443d723c-b3cb-413d-8d98-66e5d4368898 · inbound

Alethia: A Foundational Encoder for Voice Deepfakes cites this paper.

Alethia: A Foundational Encoder for Voice Deepfakes AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-06-04T02:07:39.638614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:c75c8448272ed4f34425b5c07457b1741b28a006ae918b2eb40ecfef4a62de8a