Pith. sign in

Paper Citation Record · LEDGER

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation

As of 19 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2411.09167.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.09167 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T21:03:50.616649Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact1
  • verified fuzzy46
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation aebfcad1-9433-4e40-a47d-b7152ce54428 · outbound

This paper cites Deepfake detection: Current challenges and next steps,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Deepfake detection: Current challenges and next steps,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.357406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.398544Z digest=sha256:d8947d04476a884b6d58ea13e4d4733fe75146da6ffa93397317294fe6e72bd9

Observation 04d81a3e-1812-4ca5-a923-c132e258baa5 · outbound

This paper cites Deepfake generation, detection and datasets: a rapid-review,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Deepfake generation, detection and datasets: a rapid-review,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.345615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.403616Z digest=sha256:79e625a7968e798c26bd954dc5225f14ce1efdd8b2dc3abc0883d96b7b160a44

Observation fe50a0c5-b082-48ba-bb03-de58ac6e9cf2 · outbound

This paper cites Listen to This Deepfake Audio Impersonating a CEO in Brazen Fraud Attempt,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Listen to This Deepfake Audio Impersonating a CEO in Brazen Fraud Attempt,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.330605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.407637Z digest=sha256:3c535458f3809609403384355a4c5670fecd60946e5bd2f6b3db91ef597e2f45

Observation ea2b6ddd-123f-4f30-8fff-2f9b12a22a5f · outbound

This paper cites Telegram Still Hasn’t Removed an AI Bot That’s Abusing Women,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Telegram Still Hasn’t Removed an AI Bot That’s Abusing Women,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.308724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.415634Z digest=sha256:05398ccabfee7f7427e8509ec224b66c5b2f55d15f2c80a8c6a1e25717253941

Observation 621e5e9a-dc45-4d07-a193-8b4b647e461b · outbound

This paper cites WaveFake: A Data Set to Facilitate Audio Deepfake Detection,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation WaveFake: A Data Set to Facilitate Audio Deepfake Detection,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T21:03:50.419374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:03:50.419374Z digest=sha256:3ad83443026f176eeaea8b922f23e24a8c62ad5f18fe45c9a77442d711806cdb

Observation e2525813-e802-404b-883e-b07e413f2542 · outbound

This paper cites P-flow: A fast and data-efficient zero-shot tts through speech prompting,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation P-flow: A fast and data-efficient zero-shot tts through speech prompting,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.291366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.423466Z digest=sha256:ac6c3d69abc8f278caa6d36a1ca4e97d94cca5cab2c0aa486261b1790bd41041

Observation 48931689-c0c4-4b8c-979b-a394767daa44 · outbound

This paper cites Phoneme hallucinator: One-shot voice conversion via set expansion,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Phoneme hallucinator: One-shot voice conversion via set expansion,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.281101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.427580Z digest=sha256:b8498dcb281ec072e658cd5a73e942773fb23b32011f6c1dc95acf5932d075e7

Observation 20c01067-5c2c-43c7-8102-78cab9e1b8db · outbound

This paper cites A review of modern audio deepfake de- tection methods: challenges and future directions,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation A review of modern audio deepfake de- tection methods: challenges and future directions,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.269265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.430772Z digest=sha256:719e5ffacf707cef3482d9f11c82e5d5883e4ac3a5b341f6628f39040e2a5d81

Observation b7875bd8-8415-4c32-81c4-2c5575b9a0ee · outbound

This paper cites AI-Synthesized V oice Detection Using Neural V ocoder Artifacts,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation AI-Synthesized V oice Detection Using Neural V ocoder Artifacts,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.258527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.434038Z digest=sha256:47d3f76d456a60db760078a53b2b4f38bdd90c6152fefa359d8eecd044aaf284

Observation b28150f5-6399-4659-bec9-a4a786d59fd6 · outbound

This paper cites Spoofing speech detection using modified relative phase information,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Spoofing speech detection using modified relative phase information,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.246833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.437550Z digest=sha256:94ab05fbebc588df38b410090729950572c75e83bc47b8c27782674a1891ebfa

Observation bbf5cfda-d9e3-4104-a3b2-6ea9b950bacd · outbound

This paper cites Robustness of speech spoofing detectors against adversarial post-processing of voice conversion,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Robustness of speech spoofing detectors against adversarial post-processing of voice conversion,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.235653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.440645Z digest=sha256:17495f55340d4b2656d7af3cf4ae4a55a43a04bf74b58ba7bfb4700ab4bb52ee

Observation 8ca4d97b-15e5-40ec-89f5-c4af54fdeb7c · outbound

This paper cites Detecting spoofed speeches via segment-based word cqcc and average zcr for embedded systems,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Detecting spoofed speeches via segment-based word cqcc and average zcr for embedded systems,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.223857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.443997Z digest=sha256:e0e1f9aa8e52b3487b6ceafecc827e8b668bc85cda44c72d0f1b4656539f05c8

Observation b61c2bb1-17dc-4007-aa1c-be2b67282926 · outbound

This paper cites Detecting ai-synthesized speech using bispectral analysis.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Detecting ai-synthesized speech using bispectral analysis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T21:03:50.447559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:03:50.447559Z digest=sha256:66c071b592a25fd9a21ec8882b60422b3b2f902f825220dd8439707d987a95bc

Observation 2855a35f-82c0-468a-8c0f-4196f934212f · outbound

This paper cites Fake Audio Detection Based On Unsupervised Pretraining Models,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Fake Audio Detection Based On Unsupervised Pretraining Models,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.205798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.450945Z digest=sha256:dfcf70b957a4c9056a920f167da6c5a36bef008ad089a53d0884c8ad44ccb7bc

Observation 34627ef9-0ad9-4600-8c1b-379c65f4cf92 · outbound

This paper cites Aasist: Audio anti-spoofing using integrated spectro-temporal graph attention networks,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Aasist: Audio anti-spoofing using integrated spectro-temporal graph attention networks,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.195020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.454275Z digest=sha256:0a0e0b05c02a1b441aaf079971eba8ccebc4a3165c703ec5c96aa990c99840f5

Observation bb285b13-9a30-4218-9c41-aa37295ca1dd · outbound

This paper cites A robust audio deepfake detection system via multi-view feature,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation A robust audio deepfake detection system via multi-view feature,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.183775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.457640Z digest=sha256:97a656f7b016f26943d6ae75fa9650239d60a0b3c5c22f3b3bb2eab1b1680ca9

Observation e9f1a0d2-f3dc-4c58-b75e-5ebe507e76b2 · outbound

This paper cites Towards end-to-end synthetic speech detection,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Towards end-to-end synthetic speech detection,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T21:03:50.460964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:03:50.460964Z digest=sha256:dc8357aa0ab7cbf51a27b9ee2524973f62d25d3cb7a6ae977460f95541df4f46

Observation fec4fbec-7fc5-43dd-9bb7-5266f27a305e · outbound

This paper cites Does audio deepfake detection generalize?.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Does audio deepfake detection generalize?

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.160755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.464261Z digest=sha256:5e93c5c942f2af80ab7e91a46acdabedfaf448d52ee527247433a9b6d6981fe4

Observation c730db81-5258-42b6-b145-1d26e928cb4f · outbound

This paper cites Sedeptts: Enhancing the naturalness via semantic dependency and local convolution for text-to-speech synthesis,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Sedeptts: Enhancing the naturalness via semantic dependency and local convolution for text-to-speech synthesis,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.149500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.467755Z digest=sha256:58a2534d8d722f2e53ee5868780e4b9ba1d0019fedb58961a8c0ef5a85db2f9a

Observation a6d52a78-cce8-47db-bc24-91d5e3264148 · outbound

This paper cites Drvc: A frame- work of any-to-any voice conversion with self-supervised learning,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Drvc: A frame- work of any-to-any voice conversion with self-supervised learning,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.137040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.471568Z digest=sha256:40e4c94a8339051a538cebb874a09f803ddcf5388588d8a8ce13aacc9503afc1

Observation c605c15d-bb5f-4882-b3bf-987edebc7e80 · outbound

This paper cites WaveNet: A Generative Model for Raw Audio.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation WaveNet: A Generative Model for Raw Audio

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T21:03:50.475224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:03:50.475224Z digest=sha256:361302c18cab0a8a7d54eb62eded356b0b65c3a0cb58dc33533b5428cad4dd4e

Observation 5c6374b2-b10b-4a39-b5af-0ac885e673ce · outbound

This paper cites Efficient neural audio synthesis,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Efficient neural audio synthesis,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.125490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.479062Z digest=sha256:a6349ca40ce6eee9d6bb829b82f4a3d851d9fd1a9661c70f0608c71dfb2ae2c3

Observation 877aff25-41f8-4692-8b4e-b5d06b7450bc · outbound

This paper cites Guided-tts: A diffusion model for text-to- speech via classifier guidance,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Guided-tts: A diffusion model for text-to- speech via classifier guidance,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.112746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.482408Z digest=sha256:3e509079de27a6d46d66416c890f21f26ee5be52c2d23a92e1d40b6ee5fc2be3

Observation 7d8b0804-acdd-4b4b-b658-8c73c0a7fa2f · outbound

This paper cites Grad- tts: A diffusion probabilistic model for text-to-speech,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Grad- tts: A diffusion probabilistic model for text-to-speech,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.100061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.486123Z digest=sha256:d0cacf8f2643acd6e152c6760268a19220e5375d14827013265b15ed60e81cb4

Observation f0bde8c4-fef3-4144-b34e-2e4bc366190e · outbound

This paper cites Melgan: Generative adversarial networks for conditional waveform synthesis,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Melgan: Generative adversarial networks for conditional waveform synthesis,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T21:03:50.489360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:03:50.489360Z digest=sha256:bb5a796ad64216f157c26bee473854f6cedb2e2ce4edda2c9298be6fdef9fefd

Observation 3ef7e56b-ef17-4b49-ad37-4dcdae36f867 · outbound

This paper cites Dspgan: a gan-based universal vocoder for high-fidelity tts by time- frequency domain supervision from dsp,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Dspgan: a gan-based universal vocoder for high-fidelity tts by time- frequency domain supervision from dsp,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.081850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.492571Z digest=sha256:c2262e26e40f151c2b6f95d5bc480065ddcfc58dbd660c1249c9555d4ee9f0d3

Observation dbb75dca-e490-49b1-ae4d-b0733b569a0d · outbound

This paper cites Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.070170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.495713Z digest=sha256:241282ae83b15417d9bd6c0195015b0531783539e8415e6211eabb25a10aa1c5

Observation 539ff319-10f4-419f-8349-91413cac04d6 · outbound

This paper cites Hierspeech: Bridging the gap between text and speech by hierarchical variational inference using self-supervised representations for speech synthesis,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Hierspeech: Bridging the gap between text and speech by hierarchical variational inference using self-supervised representations for speech synthesis,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.060076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.499139Z digest=sha256:b01523070c2e85fa771d2c3004e00a992329b76752e327471cf3cb1da7b7698b

Observation f6711423-c16d-4e9d-b2a8-b25daa92b9fe · outbound

This paper cites Unisyn: an end-to-end unified model for text-to-speech and singing voice synthesis,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Unisyn: an end-to-end unified model for text-to-speech and singing voice synthesis,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.048360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.501911Z digest=sha256:fc588700eaccf5f01c758e0e4b0c2955b31119cb15651236a28c2a3852030840

Observation 8f7e01f4-cf42-4151-a032-e722445fb7b1 · outbound

This paper cites ” hello, it’s me.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation ” hello, it’s me

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.037503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.504759Z digest=sha256:23601e3347efac4f9c74649e5d8ceb31821e2b88060ad8533e9f790754f39791

Observation 4f32f520-36e3-4ed5-93b5-13d6fd41ef7d · outbound

This paper cites Audio replay attack detection with deep learning frameworks.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Audio replay attack detection with deep learning frameworks

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.027326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.508430Z digest=sha256:45ad8d853214f50afd0fdf34f0e81d99f32ac7bb8e9134de9a69626afa75bffb

Observation 059b43dc-4cc9-4475-bc22-495de91d6c6d · outbound

This paper cites Improved rawnet with feature map scaling for text-independent speaker verification using raw waveforms,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Improved rawnet with feature map scaling for text-independent speaker verification using raw waveforms,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.016564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.511197Z digest=sha256:5b12926e858617441e368235421ee73d0a88bc525d764e65f82caa60d8b2ab57

Observation 11621a76-62e0-4285-aec2-e2707a5b9b2c · outbound

This paper cites Asvspoof 2021: Towards spoofed and deepfake speech detection in the wild,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Asvspoof 2021: Towards spoofed and deepfake speech detection in the wild,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.005572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.514308Z digest=sha256:aa7e981304526079f9475eda5de2e657bd3b04d71876b92aa3c63c3b4a896122

Observation 89eea66c-b226-4055-9d75-087b47d7221f · outbound

This paper cites Stc antispoofing systems for the asvspoof2019 challenge,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Stc antispoofing systems for the asvspoof2019 challenge,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.995027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.517624Z digest=sha256:870946e0af63b2a79c4148d96f3f3bcbb2ac5023ff85f141dbefbd31ca2f8bed

Observation 3b9e7a7d-82fc-4071-bbd0-7e57ef2dc2fd · outbound

This paper cites Domain Generalization via Aggregation and Separation for Audio Deepfake Detection,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Domain Generalization via Aggregation and Separation for Audio Deepfake Detection,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.983555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.521033Z digest=sha256:c67b5579df5bb08d5883aa63eac1110d2274cca35c488637b77e003b954da78a

Observation 000d6ec5-0b0c-4ea5-a0b3-b8a869968d20 · outbound

This paper cites Fully automated end-to-end fake audio detection,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Fully automated end-to-end fake audio detection,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.971950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.525118Z digest=sha256:b065b320ca768c04d934c822da4893b3fff662e34be3075d6c7bea243a72deae

Observation b04acc32-b85b-4168-9e60-e02ee0f9b99e · outbound

This paper cites Deepfake audio detection with vision transformer based method,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Deepfake audio detection with vision transformer based method,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.959988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.528720Z digest=sha256:f1040d517efa961fb8763153a95d43f069e523e33f28b30ab95ec5b07cc74c52

Observation 538aa9f7-c70e-46c3-9c78-c4c12eee19e3 · outbound

This paper cites Audio spectrogram transformer for synthetic speech detection via speech formant analysis,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Audio spectrogram transformer for synthetic speech detection via speech formant analysis,

Reference 38

Resolution
verified exact
raw_fallback, observed 2026-08-12T21:03:50.735106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.532078Z digest=sha256:5f4bce1dcc842dd3b05551dc8e08fa2e7f00b240267eb7ab3c301745299c446c

Observation daebac64-bb1e-488b-b90d-d872590f7ade · outbound

This paper cites Audio transformer for synthetic speech detection via multi-formant analysis,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Audio transformer for synthetic speech detection via multi-formant analysis,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.947979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.536006Z digest=sha256:622631297f010c415d7c05659d37918a9476c9ed0596586a82549967cedce2e8

Observation 87e51d2a-19b3-4bfd-a435-cc249074bfe6 · outbound

This paper cites CLAD: Robust Audio Deepfake Detection Against Manipulation Attacks with Contrastive Learning.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation CLAD: Robust Audio Deepfake Detection Against Manipulation Attacks with Contrastive Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T21:03:50.539535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:03:50.539535Z digest=sha256:ef38881be364e4346ebf23590f86dc106034faf69823f700e9c04e701d62a7b1

Observation 4018a5d5-aa32-4b34-a67a-0bdf73e356fd · outbound

This paper cites Towards attention-based contrastive learning for audio spoof detection,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Towards attention-based contrastive learning for audio spoof detection,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.934632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.543619Z digest=sha256:552bff6138ae11e9c9884fbd55428c0314d24b7b0392bc056acf7e5130e4e1f9

Observation 40d5e325-2705-4440-a6a8-5b24fce36ad6 · outbound

This paper cites Deep residual learning for image recognition,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Deep residual learning for image recognition,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T21:03:50.547154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:03:50.547154Z digest=sha256:e1fe776257b1913737becd1966bd1760265d2e321373b15f071904960568a5e9

Observation c0ec93cc-7c56-40ab-ba2e-2d413d495fa8 · outbound

This paper cites How does batch normalization help optimization?.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation How does batch normalization help optimization?

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.915811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.550648Z digest=sha256:62d56cd98839807295c676d12b3b71ee9dee140491fad1a3cb36c518d61f19cb

Observation f69dd6e7-bceb-45cc-ab44-910e29567ec0 · outbound

This paper cites Do You Really Mean That? Content Driven Audio-Visual Deepfake Dataset and Multimodal Method for Temporal Forgery Localization,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Do You Really Mean That? Content Driven Audio-Visual Deepfake Dataset and Multimodal Method for Temporal Forgery Localization,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.905595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.554355Z digest=sha256:8dcf1e43fe8c061077831b1d879205e44c9f0ff4d9be0944f6ecbc9fb900a335

Observation 1536f258-f88c-422b-a485-f31d1bd884f0 · outbound

This paper cites Focal loss for dense object detection,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Focal loss for dense object detection,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T21:03:50.557953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:03:50.557953Z digest=sha256:88ae13646cd7896743768d3170afb56585e8056992c7f6d5e5ba34dc4bc54bee

Observation 1ecdfb60-8d23-4ae0-84eb-036d7ee22b78 · outbound

This paper cites Transferring Audio Deepfake Detection Capability across Languages,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Transferring Audio Deepfake Detection Capability across Languages,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.889044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.561677Z digest=sha256:4b06e780465592662a85ec2705086c3396ee0bcde485866a0cfba3b2d609e441

Observation db382c65-5051-4ae2-ae4f-7d76bc31b311 · outbound

This paper cites Parallel Wavegan: A Fast Waveform Generation Model Based on Generative Adversarial Networks with Multi-Resolution Spectrogram,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Parallel Wavegan: A Fast Waveform Generation Model Based on Generative Adversarial Networks with Multi-Resolution Spectrogram,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.877856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.565276Z digest=sha256:3825b8deabfa8aa0af2f06a797680f1e8c9fd2560f41d9f87a00486db1f20b6f

Observation 3710c738-201f-41e2-89f1-ce19c1a245c8 · outbound

This paper cites Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T21:03:50.569452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:03:50.569452Z digest=sha256:3cef271c14e15aa339fae387b57ac3bd1e5379dbdebc10849a83e0c0eb8f4e90

Observation b4105ee4-2270-468f-add7-554cb6ea21a0 · outbound

This paper cites The lj speech dataset,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation The lj speech dataset,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T21:03:50.573598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:03:50.573598Z digest=sha256:e92541b4bd0f4763ee79f054df8a9d93b5f17811a32177acf69e11bfcf8af94b

Observation 010e184a-b160-43cd-9268-97607cd3389b · outbound

This paper cites JSUT corpus: free large-scale Japanese speech corpus for end-to-end speech synthesis.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation JSUT corpus: free large-scale Japanese speech corpus for end-to-end speech synthesis

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T21:03:50.577466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:03:50.577466Z digest=sha256:cab5bf2f61c552d4ae3886a5827c8d9e846861c5c3d14c4db7a823cdf3496a59

Observation 1bd94059-e8f2-4a82-8a32-9aaf21ecda34 · outbound

This paper cites Conformer: Convolution- augmented Transformer for Speech Recognition,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Conformer: Convolution- augmented Transformer for Speech Recognition,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.853819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.581155Z digest=sha256:7e972669621b7be0dcb9a58cc97d52d30456532646620fd8c57de304e52beb41

Observation 75385680-2562-4535-a18b-495763690df2 · outbound

This paper cites Libritts: A corpus derived from librispeech for text-to-speech,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Libritts: A corpus derived from librispeech for text-to-speech,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.841895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.584565Z digest=sha256:d0e479a4e03658cf62af584b2dfc5c0286f178ccfc35e81ca417a2ddb1ce0fc7

Observation 80007aa0-991e-4878-8652-46158991b821 · outbound

This paper cites End-to-end spectro-temporal graph attention networks for speaker ver- ification anti-spoofing and speech deepfake detection,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation End-to-end spectro-temporal graph attention networks for speaker ver- ification anti-spoofing and speech deepfake detection,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T21:03:50.587445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:03:50.587445Z digest=sha256:18260c2081b4d2ab8d97592e21ff5261a69a6a0fb4546a649de0bfc0834a0ca0

Observation 19753f59-73a5-47be-9155-d0eb26188e62 · outbound

This paper cites wav2vec 2.0: A frame- work for self-supervised learning of speech representations,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation wav2vec 2.0: A frame- work for self-supervised learning of speech representations,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T21:03:50.590424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:03:50.590424Z digest=sha256:57453017702c9631a260c32a69bd3d3830c9a67fb81e291715ce039f74676231

Observation 7964a61d-0d1f-42df-ba87-5741f83702e5 · outbound

This paper cites WavLM: Large-Scale Self-Supervised Pre- Training for Full Stack Speech Processing,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation WavLM: Large-Scale Self-Supervised Pre- Training for Full Stack Speech Processing,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.816007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.593722Z digest=sha256:47b31e2115b4cdf3c85bd195282a71de969e2ccd7f44f9fcde6805b51fad0127

Observation d9fa4ad3-51e8-4cd6-8287-c0e4eaefdfb3 · outbound

This paper cites Wav2CLIP: Learning Robust Audio Representations from Clip,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Wav2CLIP: Learning Robust Audio Representations from Clip,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.805081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.597387Z digest=sha256:b581db07037df5e41cb3a1d9b04122b6a9b3f5953a6af726dcb9cd2f98b08adb

Observation 34979645-32d9-4f95-a81d-6c342e6e74ea · outbound

This paper cites Audioclip: Extending Clip to Image, Text and Audio,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Audioclip: Extending Clip to Image, Text and Audio,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.794013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.600871Z digest=sha256:a74d6bf5f522e8a3124cb637165ad5f10b8a9580e944fb0c6ede5cf040bbb0a0

Observation 79ebbbf1-12b5-4bc1-8c13-3fb05833543c · outbound

This paper cites ESResNe(X)t-fbsp: Learning Robust Time-Frequency Transformation of Audio,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation ESResNe(X)t-fbsp: Learning Robust Time-Frequency Transformation of Audio,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.782764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.604498Z digest=sha256:903aa8e591bee0d145b8531f5584aa4d1a3e900b49aa76fe6ec19aead9946800

Observation 84f2265d-9fb6-4361-a936-c78e38d90854 · outbound

This paper cites Gradient centralization: A new optimization technique for deep neural networks,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Gradient centralization: A new optimization technique for deep neural networks,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:50.771551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.608862Z digest=sha256:acb06432b8746fda14a32f8a77901f4c705e820a09306e9cfafc61f99c53cbb4

Observation 0a26ebd8-7da7-4f63-86dd-ccbbaef3b541 · outbound

This paper cites Visualizing data using t-SNE.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Visualizing data using t-SNE

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-12T21:03:50.613089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:03:50.613089Z digest=sha256:69eb7405da93d648028c86057aa80d591617946f26acc72485fe83c9def202d2

Observation 315c7e29-26e8-4ed9-87d3-528985034b0a · outbound

This paper cites Grad-cam: Visual explanations from deep networks via gradient-based localization,.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Grad-cam: Visual explanations from deep networks via gradient-based localization,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T21:03:50.616649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:03:50.616649Z digest=sha256:c8aa70280af4be8d4f6d34e7c96fbeb62c16d8e8c3541fa1e32ffde266ce5e5f

Observation e6a772e7-4641-421c-8504-51d2aae22ed4 · outbound

This paper cites Available: https://www.vice.com/en/article/pkyqvb/ deepfake-audio-impersonating-ceo-fraud-attempt.

Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation Available: https://www.vice.com/en/article/pkyqvb/ deepfake-audio-impersonating-ceo-fraud-attempt

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:03:51.319342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T21:03:50.411459Z digest=sha256:84c7a7aeb91495eb20f15b8e1a586c0c9bb9b4365358ed5fc89aa5022699fc8e

Pith citing papers

No inbound Pith citation observations are available.