Pith. sign in

Paper Citation Record · LEDGER

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts

As of 18 August 2026, this Paper Citation Record lists 84 of 84 outbound references and 3 inbound Pith citation observations for arXiv:2502.05674.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.05674 v4

Coverage vector

measured 84 of 84 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T18:27:07.170869Z

measured 87 of 87 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T19:08:14.134784Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T15:41:31.056003Z

Reference resolution

84 of 84 outbound references displayed

  • verified exact0
  • verified fuzzy72
  • unresolved11
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1345af91-b6c5-40b7-8c8f-62d21a6fe2e0 · outbound

This paper cites Does Audio Deepfake Detection Generalize? In Interspeech, pages 2783–2787, 2022.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Does Audio Deepfake Detection Generalize? In Interspeech, pages 2783–2787, 2022

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T18:27:06.713383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:27:06.713383Z digest=sha256:032241c1d91cc2d9fe7305009ce3a9538646f6fd1eba22f76006c1e2d1f17c49

Observation 5aa98b5c-0e3b-4095-b82d-6103aea8899c · outbound

This paper cites Generalization of Audio Deepfake Detection.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Generalization of Audio Deepfake Detection

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T18:27:06.718488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:27:06.718488Z digest=sha256:f44ef5f76b9748f7b1a6480f253bb66d2058035d6d25029425ecf5fcc4ef1bf9

Observation 2ec0dd75-763b-472d-bf07-14ffe6576022 · outbound

This paper cites Breaking Security-Critical V oice Authentication.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Breaking Security-Critical V oice Authentication

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T18:27:06.723009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:27:06.723009Z digest=sha256:d58764dd15f63611f81750d5a12a94f510b14547baf02362aba1a9faad4fe6aa

Observation 4e7bbe8f-e7e7-4b94-af7d-dd9b4dc49b66 · outbound

This paper cites StreamVC: Real-Time Low-Latency V oice Conversion.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts StreamVC: Real-Time Low-Latency V oice Conversion

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T18:27:06.727581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:27:06.727581Z digest=sha256:a277fd6a293c40eb864c12fbb479449ede0cae50c6cc34c35d8c91991f7b7405

Observation d1c42ac1-0e18-4477-a83e-3f8619838e1a · outbound

This paper cites Kameoka, T.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Kameoka, T

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.327936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.732126Z digest=sha256:90d7e0b4ec2c2677e78309dd6c4db1f1d52262423bfe4d5d30ca689e612e8bf2

Observation 5b34009f-8279-4f79-8baf-0c7707b3e168 · outbound

This paper cites DDDM-VC: Decoupled Denoising Diffusion Models with Disentangled Representation and Prior Mixup for Verified Robust V oice Conversion.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts DDDM-VC: Decoupled Denoising Diffusion Models with Disentangled Representation and Prior Mixup for Verified Robust V oice Conversion

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.314227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.736432Z digest=sha256:2041691c378e6de196622bfacd297b983d28c541b62a4129e0c364bf2dc44aff

Observation 68789fef-2273-4f39-b86e-4bab35ed109e · outbound

This paper cites YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot V oice Conversion for Everyone.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot V oice Conversion for Everyone

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.300183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.741411Z digest=sha256:1ed6bf2e508ae63f511df56f550c9e63ced37f0245304e44962e3ba23ded18b7

Observation d7e51e87-adb6-4cee-b819-821ae3e68175 · outbound

This paper cites AutoVC: Zero-Shot V oice Style Transfer with Only Autoencoder Loss.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts AutoVC: Zero-Shot V oice Style Transfer with Only Autoencoder Loss

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.286387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.745606Z digest=sha256:bd010765e5471b8d77920d16e725ca4da1f542278b787770d3cb8b911b81fbb8

Observation 438210dc-985b-4f9e-9c79-af13982fcf81 · outbound

This paper cites CycleGAN-VC2: Improved CycleGAN-Based Non-Parallel V oice Conversion.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts CycleGAN-VC2: Improved CycleGAN-Based Non-Parallel V oice Conversion

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.271819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.749802Z digest=sha256:f753fec50152f1d9fcf084e20d4af81907b310b93681e50386a8b7da92bc6ba8

Observation 89327487-9164-4549-b262-72d051965acb · outbound

This paper cites Reimagining Speech: A Scoping Review of Deep Learning-Powered Voice Conversion.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Reimagining Speech: A Scoping Review of Deep Learning-Powered Voice Conversion

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T18:27:06.753977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:27:06.753977Z digest=sha256:eb2c31fb0babcb6f101dfa5cab34f2fdb9035f483ac222edb600104031ddd61d

Observation 33da8299-bd80-48c2-9815-0ea12e08c3b9 · outbound

This paper cites Sahidullah, Héctor Delgado, Andreas Nautsch, Junichi Yamagishi, Nicholas Evans, Tomi H.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Sahidullah, Héctor Delgado, Andreas Nautsch, Junichi Yamagishi, Nicholas Evans, Tomi H

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.258275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.758746Z digest=sha256:db81fa8eb4f067cf50998be32d52ed0a3e5428faf5774882deffff1b3a6a5e46

Observation 8a242eeb-1f1d-490b-8392-0c07f37e45ed · outbound

This paper cites ASVspoof 2021: Accelerating Progress in Spoofed and Deepfake Speech Detection.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts ASVspoof 2021: Accelerating Progress in Spoofed and Deepfake Speech Detection

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.244067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.763065Z digest=sha256:3fea1dfdc14105d4bc77eaa746405e1415accb80a82d05f133b323022382f88f

Observation da961fce-171f-40f9-bb59-9b2b52fcdd62 · outbound

This paper cites WaveFake: A Data Set to Facilitate Audio Deepfake Detection.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts WaveFake: A Data Set to Facilitate Audio Deepfake Detection

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.229961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.767159Z digest=sha256:8ddea474d20e1004908579203b3b079be8a89c86fafeaf10ce788ffea00cdf4b

Observation 5fed6d72-f097-4e6b-972e-a79dc8c1caf4 · outbound

This paper cites Multi-Dataset Co-Training with Sharpness-Aware Optimization for Audio Anti-spoofing, 2023.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Multi-Dataset Co-Training with Sharpness-Aware Optimization for Audio Anti-spoofing, 2023

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.216131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.771246Z digest=sha256:679d874be0ec477170082a03d10fee4a22d483b2b1cae9da4c5041580ab572dd

Observation 3d2e0939-3276-4995-a14c-c6ea22f75f6d · outbound

This paper cites Deep Feature Engineering for Noise Robust Spoofing Detection.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Deep Feature Engineering for Noise Robust Spoofing Detection

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.202004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.775335Z digest=sha256:342f703223b2fc4927bcf6421bc6bf35e42ee920189ba70d1d834f3367563655

Observation 2b73daea-cc3c-4b18-9994-afd474119037 · outbound

This paper cites Investigating Raw Wave Deep Neural Networks for End-to-End Speaker Spoofing Detection.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Investigating Raw Wave Deep Neural Networks for End-to-End Speaker Spoofing Detection

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.188625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.779704Z digest=sha256:cb953d9d51a97cc6c0148857ea891eec4ccc6143a382c68e04c839cdd3f2ec0c

Observation 22319550-6841-4364-9f12-919c41527995 · outbound

This paper cites Peinado, Jose A.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Peinado, Jose A

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.174699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.783970Z digest=sha256:f280bdcf1f6811eb7de5d6164148863cb01b07de8cae64e31ce9ada230ea8c6d

Observation 71fe2bad-5434-4755-93fe-81d53575a250 · outbound

This paper cites ResNet and Model Fusion for Automatic Spoofing Detection.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts ResNet and Model Fusion for Automatic Spoofing Detection

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.160439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.788478Z digest=sha256:89c35e4467e792d48042b81fbc3ad58c5d99b60da4cd80bfd300ccae2af117d6

Observation ec435d53-bd00-425d-a0d7-83259e43b3b2 · outbound

This paper cites Investigating Self-Supervised Front Ends for Speech Spoofing Coun- termeasures.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Investigating Self-Supervised Front Ends for Speech Spoofing Coun- termeasures

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.145658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.792682Z digest=sha256:301355996b626caa16eb528c8c1815c20d55ea0a388775753f1d816570bff806

Observation c7fe8b46-67ce-4376-a120-665746a844b2 · outbound

This paper cites Exploring generalization to unseen audio data for spoofing: insights from SSL models.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Exploring generalization to unseen audio data for spoofing: insights from SSL models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.130837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.797189Z digest=sha256:3fd628a45135ebad777d7c2dc11d3e3d0b4a1a0771aa891451716d95088936a2

Observation f5aab420-040c-443e-b8a7-e11dc5d697c4 · outbound

This paper cites WavLM: Large-Scale Self-Supervised Pre-Training for Full Stack Speech Processing.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts WavLM: Large-Scale Self-Supervised Pre-Training for Full Stack Speech Processing

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.116468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.801782Z digest=sha256:07cc815627ee2ccd54ca7cd1b859137fe45ac2b09970bab439e5fe03d1e8db33

Observation e12868fc-1548-4327-aea1-5e492a6e1ad7 · outbound

This paper cites HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.102205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.806128Z digest=sha256:3f5f834bef076a542412765f1885408dd5ff7b9fea8e009fa300b71cc2e9e53a

Observation c4916d82-3a41-4222-860a-dba7d0e4b7b6 · outbound

This paper cites Wav2vec 2.0: A Framework for Self-supervised Learning of Speech Representations.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Wav2vec 2.0: A Framework for Self-supervised Learning of Speech Representations

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.088028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.810638Z digest=sha256:3cf339724d9428600ef8b224425dcb593dfe00df21f2dd08244fe315396219b1

Observation a7804705-f259-476b-a929-850c128f8cad · outbound

This paper cites Müller, Nicholas Evans, Hemlata Tak, Philip Sperl, and Konstantin Böttinger.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Müller, Nicholas Evans, Hemlata Tak, Philip Sperl, and Konstantin Böttinger

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.074206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.814597Z digest=sha256:8d0666b147f4bcede9b39532e362971ba10f7d852dcca237304fd32a3e0edf70

Observation 8086f2ed-5eaf-4a43-9cb7-9e8d0fb53d54 · outbound

This paper cites Audio Deepfake Detection: A Survey.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Audio Deepfake Detection: A Survey

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T18:27:06.818969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:27:06.818969Z digest=sha256:2b10e900c3a4268771ece6c6f01ceea0a1f22867b72c3a5133b1e5ec2167d31d

Observation 9996e138-e564-4345-a097-47eb8dd80e43 · outbound

This paper cites Open Challenges in Synthetic Speech Detec- tion.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Open Challenges in Synthetic Speech Detec- tion

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.059824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.823447Z digest=sha256:a8145e398e699853229b333c0a14129a4400ce94a3d0cf887522f72d2473cd7c

Observation d1084edd-f653-43d9-931b-6993270cd109 · outbound

This paper cites SpoofCeleb: Speech Deepfake Detection and SASV In The Wild.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts SpoofCeleb: Speech Deepfake Detection and SASV In The Wild

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.044930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.827452Z digest=sha256:4390fcf77aad283878edfb6f3ff050bc87e0d02681f581f46a39814a2ae3b903

Observation 381b3710-a182-4c9c-ad7c-61b1a0fe4224 · outbound

This paper cites Generalisation in Humans and Deep Neural Networks, 2018.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Generalisation in Humans and Deep Neural Networks, 2018

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.030883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.831872Z digest=sha256:ff6403de4163b7f6077b68fe76070787475790ea160567ef0575047552ef557c

Observation 56a9e245-3569-4774-bbea-6dde02b84408 · outbound

This paper cites Covariate Shift Adaptation by Impor- tance Weighted Cross Validation.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Covariate Shift Adaptation by Impor- tance Weighted Cross Validation

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.016801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.836170Z digest=sha256:f7e3a123af01bb62d1619c37c1e1b2d1004facf5e683b9e7535be20e3e989750

Observation f67d27c8-c3ee-4002-86db-70498e52b84e · outbound

This paper cites Analysis of Representations for Domain Adaptation.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Analysis of Representations for Domain Adaptation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:08.002201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.840459Z digest=sha256:2bfb0ec4f1845083b0165ae9b548fca9f5c080e5bafc712dffa803d1720859f0

Observation ab6d245c-deea-4072-ac52-147eb0aed620 · outbound

This paper cites A Survey on Concept Drift Adaptation.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts A Survey on Concept Drift Adaptation

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.987898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.844654Z digest=sha256:514718411a073d3000ec4f53d6768c7becf04267dcd60af5e3a7e63defa3836e

Observation ab0187ce-bba3-4e4a-9894-dc6417627c79 · outbound

This paper cites Kinnunen, Nicholas Evans, Kong Aik Lee, and Junichi Yamagishi.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Kinnunen, Nicholas Evans, Kong Aik Lee, and Junichi Yamagishi

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.973874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.848753Z digest=sha256:ae4fa2fee3ca49d51fc04c4fc5eb6f735ea43f08334214a0c4b80aec30e4c2da

Observation 137367a7-9122-4c42-878e-98048f7c7ba0 · outbound

This paper cites MLS: A Large- Scale Multilingual Dataset for Speech Research.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts MLS: A Large- Scale Multilingual Dataset for Speech Research

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.958969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.853072Z digest=sha256:833bb110f4d239c8297061f9527d9d805317909d3c2b37383ca6ffb3dc846ab2

Observation bdfbd036-9bc6-4ecf-b519-a289bf6c6b89 · outbound

This paper cites an unresolved cited work.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-08T18:27:07.944425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.857451Z digest=sha256:55937ef9f87a3c2ba5667ca58beb666671ed9ee383e0be59324ec288210c2c86

Observation 48a1d464-c985-45ff-bad4-4c226e1d4fdb · outbound

This paper cites Spoofed Training Data for Speech Spoofing Countermeasure Can Be Efficiently Created Using Neural V ocoders.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Spoofed Training Data for Speech Spoofing Countermeasure Can Be Efficiently Created Using Neural V ocoders

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.929337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.861758Z digest=sha256:ef40e36145db029bd075790660eaa04518c0ac1e7580235412b6438243de8195

Observation 0b1fa958-fbb7-4aaa-ae2e-a0b081ea2f89 · outbound

This paper cites AI-Synthesized V oice Detection Using Neural V ocoder Artifacts.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts AI-Synthesized V oice Detection Using Neural V ocoder Artifacts

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.915079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.866391Z digest=sha256:9fedf87fcf01d175c5337aeb5191bd2d35b2c40b59c44de174022e62ec2f6b42

Observation d56d4eff-816c-44f3-84a1-01657bf6739d · outbound

This paper cites A Cross-V ocoder Study of Speaker Independent Synthetic Speech Detection using Phase Information.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts A Cross-V ocoder Study of Speaker Independent Synthetic Speech Detection using Phase Information

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.900879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.870551Z digest=sha256:0b779ccc0e08d1c5e2ef867151fa36f07aefdd4f9b23a061a8e46a9045b0b4ee

Observation 50fb15cc-2749-4f36-a176-433e191ef292 · outbound

This paper cites Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.886765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.874684Z digest=sha256:af158b878f00d277afad8ac720b6a735d20b68f676f6b71c983f78ed2cee301f

Observation 13960e45-a3bf-45a4-9551-6de246cf8bd9 · outbound

This paper cites Anomaly Detection of Deepfake Audio Based on Real Audio Using Generative Adversarial Network Model.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Anomaly Detection of Deepfake Audio Based on Real Audio Using Generative Adversarial Network Model

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.871938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.879023Z digest=sha256:acf1cdc8542323dfbce55189845ca7c7a4ea099ade38a1c98ff206147d8cf489

Observation 404e47bb-47cd-4bac-ba74-3198957bd337 · outbound

This paper cites The LJ Speech Dataset.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts The LJ Speech Dataset

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.857138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.883286Z digest=sha256:c6f447af495a6d53a70a03cb62e6e1c3b99f639ab96a70ca456cbb89e4c0ea4d

Observation 14d13191-c7e5-4e08-a7a7-e5bb031f2b9b · outbound

This paper cites JSUT corpus: free large-scale Japanese speech corpus for end-to-end speech synthesis.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts JSUT corpus: free large-scale Japanese speech corpus for end-to-end speech synthesis

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T18:27:06.887488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:27:06.887488Z digest=sha256:e9e0904ad95081f2a5e78fc11241d1e3714d6dc7e0820efb5cb7d54b6e40421e

Observation 1827f23e-bbb1-4799-b305-cfcc50b12e26 · outbound

This paper cites Weiss, Ye Jia, Zhifeng Chen, and Yonghui Wu.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Weiss, Ye Jia, Zhifeng Chen, and Yonghui Wu

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.843232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.892323Z digest=sha256:554ff2a418ddf8d49fd736ee967d4132e2ac77f566a3978df016fc79fe9b5483

Observation 774d222b-cd7e-47b7-93be-36b2462865a9 · outbound

This paper cites an unresolved cited work.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-08T18:27:07.828937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.896567Z digest=sha256:4891d7469defcd8fd36586a8dbac28b439708bd98e13d465e7f4deddda00a257

Observation 12244a3c-9def-4d3b-8101-cfdc337fe9fe · outbound

This paper cites FoR: A Dataset for Synthetic Speech Detection.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts FoR: A Dataset for Synthetic Speech Detection

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.815071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:06.996607Z digest=sha256:bdf3de6d70d5ca1f4480538941da76f7ca958bf0433d12c23c08e01855df529d

Observation 62f66595-a552-45b2-9ea1-69d2fb9238af · outbound

This paper cites MLAAD: The Multi-Language Audio Anti-Spoofing Dataset.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts MLAAD: The Multi-Language Audio Anti-Spoofing Dataset

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.801873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.001482Z digest=sha256:56b8a8cce3ee0ef7c0986b82d9e64a1ef99a69bafbf24f34a126a7bc475dacab

Observation 6f5aa142-2bc3-48a2-aacb-c38098d7bb69 · outbound

This paper cites V oxceleb: A large-scale speaker identification dataset.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts V oxceleb: A large-scale speaker identification dataset

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.788571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.005972Z digest=sha256:2af577425d76343b30c2620cecb0c41f5c08c994494a69c57904a6b9b64b7edd

Observation e07806b5-6be1-4536-aa64-d8a07c639065 · outbound

This paper cites V oxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts V oxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.774895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.010100Z digest=sha256:c04b839459556b50c906548cccd7eba7371f2e4f4b690f09de052588d80f290d

Observation 481466b0-5703-4a1f-843d-596148ccc12e · outbound

This paper cites GigaSpeech: An Evolving, Multi-Domain ASR Corpus with 10,000 Hours of Transcribed Audio.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts GigaSpeech: An Evolving, Multi-Domain ASR Corpus with 10,000 Hours of Transcribed Audio

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.761007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.014132Z digest=sha256:c13175350b71d801961fe8f52360462e352f8dd9e2449ac2187d77a2d8da1683

Observation f1810a2d-e77b-4bb9-912a-b16444e46067 · outbound

This paper cites AISHELL-1: An Open-Source Mandarin Speech Corpus and A Speech Recognition Baseline.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts AISHELL-1: An Open-Source Mandarin Speech Corpus and A Speech Recognition Baseline

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.747127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.018599Z digest=sha256:2b1b8b3438545ad946c82793fc4fae32645b3f84b479a4b12a786a1e65c92fe9

Observation d6558985-4281-44b3-9c45-64f940cfb0f8 · outbound

This paper cites Common V oice: A Massively-Multilingual Speech Corpus.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Common V oice: A Massively-Multilingual Speech Corpus

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.733550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.022899Z digest=sha256:e7f2aca3c35cca156e3306117190f660c33fb3abfafcc06b92acdd5af1609db5

Observation f76942e7-d44f-41b7-b86f-de0c2ace8218 · outbound

This paper cites Lotfian and C.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Lotfian and C

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.720140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.027002Z digest=sha256:a0c2127684339358035c379930759b47af518c91904044b35489617d172dcc90

Observation d2f4ff3d-ee56-4e2d-8762-776cf8916ce9 · outbound

This paper cites V oxCeleb2: Deep Speaker Recognition.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts V oxCeleb2: Deep Speaker Recognition

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.706136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.031435Z digest=sha256:b909e5c6084ded4f21f37a3c29dd2188ba129de387c712fca2cbf725ad1a8a05

Observation 28e56e21-3598-40a7-a0e9-db224d6625db · outbound

This paper cites Librispeech: An ASR Corpus Based on Public Domain Audio Books.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Librispeech: An ASR Corpus Based on Public Domain Audio Books

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.692210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.035800Z digest=sha256:4bdfa668627ed0e9af1f52163d988254f2e4f7f499d55fd5c98fca87680a1843

Observation 2192d576-be68-4307-bf01-b2dd2bced321 · outbound

This paper cites Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.678163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.040202Z digest=sha256:f5ca58d13a3fbec11fcc096d66ea982acd3e17859f6f1bbd5495dbfe4aceeb48

Observation 6b25fa22-7772-4c02-a645-e8d165d658fb · outbound

This paper cites Fastpitch: Parallel Text-to-Speech with Pitch Prediction.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Fastpitch: Parallel Text-to-Speech with Pitch Prediction

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.664034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.044519Z digest=sha256:29655dca5d4ca4bd5ca4c92ff96c513de9ea1288c7bf1dfb53204e5860de2f61

Observation 0fcedfcc-ea9d-4d40-94b0-7432cc1592fa · outbound

This paper cites Glow-TTS: A Generative Flow for Text-to-Speech via Monotonic Alignment Search.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Glow-TTS: A Generative Flow for Text-to-Speech via Monotonic Alignment Search

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.650420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.048635Z digest=sha256:6f3465f94ac6d81e32f806a9d37d3fd49c0283d3645158b522ec9802f6e2767f

Observation 4a28f3f6-e058-449a-bb3f-fe10516883b5 · outbound

This paper cites Grad-TTS: A Diffusion Probabilistic Model for Text-to-Speech.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Grad-TTS: A Diffusion Probabilistic Model for Text-to-Speech

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.636388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.053117Z digest=sha256:3785732a15a9496410fe9952b35792cca8de184fe2159f9ea79c0d91d7a36507

Observation 90fc0f9c-b398-4c03-9fd7-39c5f559cfe7 · outbound

This paper cites XTTS: a Massively Multilingual Zero-Shot Text-to-Speech Model.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts XTTS: a Massively Multilingual Zero-Shot Text-to-Speech Model

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-08T18:27:07.057268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:27:07.057268Z digest=sha256:51b7124cb20e87fa288e01147c2abb8bc8b868edbe5dc68ab2ccf7b8fe824851

Observation 311db61e-1d25-4690-b6e8-c12c226214a3 · outbound

This paper cites Superseded-CSTR VCTK Corpus: English Multi-Speaker Corpus for CSTR V oice Cloning Toolkit.University of Edinburgh, The Centre for Speech Technology Research, 2016.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Superseded-CSTR VCTK Corpus: English Multi-Speaker Corpus for CSTR V oice Cloning Toolkit.University of Edinburgh, The Centre for Speech Technology Research, 2016

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.622687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.061861Z digest=sha256:4dc2a005d0f85a0684be04af9b654392c9eb539492b2c77f75e84dc03aa1acbc

Observation 7458b232-9457-4ed8-a9c2-ee99b2a64701 · outbound

This paper cites Automatic Speaker Verification Spoofing and Deepfake Detection Using Wav2vec 2.0 and Data Augmenta- tion.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Automatic Speaker Verification Spoofing and Deepfake Detection Using Wav2vec 2.0 and Data Augmenta- tion

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.608973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.066380Z digest=sha256:f0d361651327431a3fcf2a9d6d182ec961a308fb24e9ab9182a4e189a7d873b4

Observation 1e463425-0cf7-4068-af30-253b8c4e421f · outbound

This paper cites XLS-R: Self-Supervised Cross-Lingual Speech Representation Learning at Scale.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts XLS-R: Self-Supervised Cross-Lingual Speech Representation Learning at Scale

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.595028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.070531Z digest=sha256:7208a49805ba52bcf72ed0c8eda5952ba8e19ae5233ebc1a1dcf4c10ef5f4299

Observation 3a15ffb2-b41f-4129-b7b4-fddc6517ceb7 · outbound

This paper cites AASIST: Audio Anti-Spoofing using Integrated Spectro-Temporal Graph Attention Networks.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts AASIST: Audio Anti-Spoofing using Integrated Spectro-Temporal Graph Attention Networks

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.581081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.074660Z digest=sha256:978ed9d8a3dcd5319c5b9bd7a4843557648da2bc1b20e05a6099fa2793b67930

Observation 2482f69b-e3d5-4a08-89e0-4619e118d5c1 · outbound

This paper cites The M-AILABS Speech Dataset.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts The M-AILABS Speech Dataset

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.566321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.078609Z digest=sha256:5c495918ea364ba15fb3ff22530d92a8e179611926063ce8ab42334452ae92d7

Observation 18f92873-d05e-44c8-a73e-52be411e6552 · outbound

This paper cites Chen, Marcus Bishop, and Nicholas Andrews.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Chen, Marcus Bishop, and Nicholas Andrews

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.553487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.082693Z digest=sha256:194ee38d98f75a98a2405aac3fa82619e6e21ebeb9617a6a86897be0ef22809e

Observation 6c55984f-ea6a-416d-982a-a68fa334b0cc · outbound

This paper cites A Simple Fix to Mahalanobis Distance for Improving Near-OOD Detection.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts A Simple Fix to Mahalanobis Distance for Improving Near-OOD Detection

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-08T18:27:07.086839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:27:07.086839Z digest=sha256:64048cdbc6baa016d21aca8f7f23940a73c942fc02f93184645b7bbf33bc0f8b

Observation 9cbb3065-cd4e-4809-bbe4-348ea853bd1c · outbound

This paper cites WaveNet: A Generative Model for Raw Audio.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts WaveNet: A Generative Model for Raw Audio

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.540383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.091165Z digest=sha256:a53bd1083981158830b75cea541cb640d14363d13639c99b19fc485271445f97

Observation b3f84057-5e38-47a3-a8bf-df0f5ebe2fd2 · outbound

This paper cites SampleRNN: An Unconditional End-to-End Neural Audio Generation Model.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts SampleRNN: An Unconditional End-to-End Neural Audio Generation Model

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.526810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.095135Z digest=sha256:c63a85d04ff4da342f3233a89b7218c3b9aed260aac9aecc3aec4e13e05914ab

Observation 22b8b01d-16e4-46c2-866c-6efcbc01564a · outbound

This paper cites Efficient Neural Audio Synthesis.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Efficient Neural Audio Synthesis

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.513691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.099531Z digest=sha256:c6e055e51d6b33711456d6039d902677ee2e5fb802832cb37e0f2351d7d6ee66

Observation 07b46796-ee72-4e50-b039-ffa45ed0e6b1 · outbound

This paper cites Generative Adversarial Networks.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Generative Adversarial Networks

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.499919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.103778Z digest=sha256:26fd289a6caddc9f30748c21a7d72f1b941f42b9caf411720de59e0fe7b32665

Observation 84df5bda-4dd6-4f0e-9c34-3f28e86a179b · outbound

This paper cites Parallel Wavegan: A Fast Waveform Genera- tion Model Based on Generative Adversarial Networks with Multi-Resolution Spectrogram.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Parallel Wavegan: A Fast Waveform Genera- tion Model Based on Generative Adversarial Networks with Multi-Resolution Spectrogram

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.471705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.111974Z digest=sha256:e15ff6ad5dd986b5e99c15bfc2479dc35c7ec8277dfb4cdcda45774dfb928616

Observation a5206af6-3bcf-4a2a-bcf4-670cb5326242 · outbound

This paper cites HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.457924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.116033Z digest=sha256:ac9494b9d49fe6ec3f03b115567e5e36a8548236815bd501cad34d1e628dd523

Observation b15a5e08-df5b-47df-83a4-7fd26e601f53 · outbound

This paper cites Multi-Band Melgan: Faster Waveform Generation For High-Quality Text-To-Speech.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Multi-Band Melgan: Faster Waveform Generation For High-Quality Text-To-Speech

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.486199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.119951Z digest=sha256:2edfc34621b859a275559047a790a227281edb063db59c489f31e905350bcd34

Observation ae242a82-1cf1-481d-bca0-fef5cdd0ea57 · outbound

This paper cites StyleMelGAN: An Efficient High-Fidelity Adversarial V ocoder with Temporal Adaptive Normalization.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts StyleMelGAN: An Efficient High-Fidelity Adversarial V ocoder with Temporal Adaptive Normalization

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.444100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.124001Z digest=sha256:740dbcfdfb20d045de3592532aa3d9f502e32a7ac6eae86d56bb746839d16919

Observation 154d65a8-2652-482a-9764-ac4279870a56 · outbound

This paper cites UnivNet: A Neural V ocoder with Multi-Resolution Spectrogram Discriminators for High-Fidelity Waveform Generation.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts UnivNet: A Neural V ocoder with Multi-Resolution Spectrogram Discriminators for High-Fidelity Waveform Generation

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.430393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.128125Z digest=sha256:1315d139ad7096b1c00a5466d1a868e8d57a9167e8239ade7ef78bedfd4d4089

Observation 9e2aa6ab-9519-4eca-afeb-82cebd79fdc6 · outbound

This paper cites BigVGAN: A Universal Neural V ocoder with Large-Scale Training.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts BigVGAN: A Universal Neural V ocoder with Large-Scale Training

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.416573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.132382Z digest=sha256:707e2f5204e9c27e14686dd8358fe33b81e1e914fac97ba256e5dfe4c64de52b

Observation 592013d2-3f24-4b7e-b53a-2c12d72f9694 · outbound

This paper cites BigVSAN: Enhancing GAN-based Neural V ocoders with Slicing Adversarial Network.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts BigVSAN: Enhancing GAN-based Neural V ocoders with Slicing Adversarial Network

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.402827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.136590Z digest=sha256:03168c0c369ec542f4ce51458281f389d15c47351599891f74029830fbfc33b3

Observation 2eec0362-3005-460d-877b-9079301163e5 · outbound

This paper cites SAN: Inducing Metrizability of GAN with Discriminative Normalized Linear Layer.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts SAN: Inducing Metrizability of GAN with Discriminative Normalized Linear Layer

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.388503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.140714Z digest=sha256:091e216eb7a9a4f3688b11821a56e4d3f6cf70efdf030130c84231ebaa794f74

Observation 0aad33af-809b-4c17-b119-63af765b82f5 · outbound

This paper cites ISTFTNET: Fast and Lightweight Mel-Spectrogram V ocoder Incorporating Inverse Short-Time Fourier Transform.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts ISTFTNET: Fast and Lightweight Mel-Spectrogram V ocoder Incorporating Inverse Short-Time Fourier Transform

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.373047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.144890Z digest=sha256:6d04669ef62f510631051d512487c4c09acb5d243f0efbde6d4ab689d2353c29

Observation 12288ee4-e635-40df-98c1-56af26aa67ea · outbound

This paper cites V ocos: Closing the Gap between Time-Domain and Fourier-Based Neural V ocoders for High-Quality Audio Synthesis.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts V ocos: Closing the Gap between Time-Domain and Fourier-Based Neural V ocoders for High-Quality Audio Synthesis

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.357188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.149067Z digest=sha256:6c1b6f753ce66be0d00fdee225dbf70ce19ca731e271fa7fb92719cc1206e3e9

Observation 7fda7123-047c-41dc-b85f-4e211872a320 · outbound

This paper cites APNet: An All-Frame-Level Neural V ocoder Incorporating Direct Prediction of Amplitude and Phase Spectra.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts APNet: An All-Frame-Level Neural V ocoder Incorporating Direct Prediction of Amplitude and Phase Spectra

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.341486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.153454Z digest=sha256:e553b4985292d94d45839086f7b156ec5406effbd8e897aea80b36df0d1f7ff1

Observation 5c07dec6-0589-40c0-90ec-6e0b360f4c38 · outbound

This paper cites APNet2: High-quality and High-efficiency Neural V ocoder with Direct Prediction of Amplitude and Phase Spectra.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts APNet2: High-quality and High-efficiency Neural V ocoder with Direct Prediction of Amplitude and Phase Spectra

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.327080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.157772Z digest=sha256:ba253a601a463e93b891d5bb5d1bb5e1def8a96a9e74cfccf0c5728cf2fa73fd

Observation b5b7905c-8c12-45f5-948b-9aeaf312f899 · outbound

This paper cites Waveglow: A Flow-Based Generative Network for Speech Synthesis.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Waveglow: A Flow-Based Generative Network for Speech Synthesis

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.313052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.162368Z digest=sha256:2f3011aee68177a44a7d8ecbbd6b109f98515f77d1affe13ef76f5cd57fa2427

Observation 9dd1eea5-3544-4e0b-8033-eef5b6de93ca · outbound

This paper cites Weiss, Mohammad Norouzi, and William Chan.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts Weiss, Mohammad Norouzi, and William Chan

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T18:27:07.298178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.166637Z digest=sha256:e03327a5dee3a4cce512f0b3e28d1ef48aef32542ddbc62e6e3d18865aa286d1

Observation b9789642-ca82-4cb6-a22a-b6737ea10d63 · outbound

This paper cites UTMOS: UTokyo-SaruLab System for V oiceMOS Challenge 2022.

ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts UTMOS: UTokyo-SaruLab System for V oiceMOS Challenge 2022

Reference 85

Resolution
malformed identifier
raw_fallback, observed 2026-08-08T18:27:07.282124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T18:27:07.170869Z digest=sha256:df00054ed16b862e2e449266a991792e6d53724326851b211639adfe736eda8d

Pith citing papers

Observation faae1c4c-b29a-46e3-806b-30c707938897 · inbound

Rapidly Adapting to New Voice Spoofing: Few-Shot Detection of Synthesized Speech Under Distribution Shifts cites this paper.

Rapidly Adapting to New Voice Spoofing: Few-Shot Detection of Synthesized Speech Under Distribution Shifts ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T19:08:14.134784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T19:08:14.134784Z digest=sha256:ec6a86f1754659feda8438e62198a591fd6be58b80e6d0bcbf55794e8066eb3b

Observation 265e0227-7795-4c2a-a43c-233ec782d76d · inbound

Alethia: A Foundational Encoder for Voice Deepfakes cites this paper.

Alethia: A Foundational Encoder for Voice Deepfakes ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:31.105348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:36130f32ca5f9027d7894309e3d6a5468efb3d4131660fa81e89365a07738d5c

Observation 3d28e6fa-c339-41b4-b38c-36395d548953 · inbound

Evaluating AI Models' Capability to Automate Voice Phishing Attacks cites this paper.

Evaluating AI Models' Capability to Automate Voice Phishing Attacks ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-14T14:15:11.793858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T14:15:11.793858Z digest=sha256:2a16a27934a15ae471111caed7a3c692a48334dd542c38ef05a45821a4edc921