Pith. sign in

Paper Citation Record · LEDGER

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection

As of 21 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 1 inbound Pith citation observation for arXiv:2509.04161.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.04161 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:24:21.774639Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-09T19:27:59.124425Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T15:41:32.859362Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy39
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c70deeb6-e849-41a6-a9ad-f82a07e070c5 · outbound

This paper cites Asvspoof 2019: Future horizons in spoofed and fake audio detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Asvspoof 2019: Future horizons in spoofed and fake audio detection,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.301745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.408376Z digest=sha256:0f7cb7476c886bbb45b067ada9fcdf2a2ec9c9962a5be232c6e50c376cb87531

Observation 24e11cd6-bc45-4e20-af29-f71bea87989e · outbound

This paper cites Asvspoof 2021: accelerating progress in spoofed and deepfake speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Asvspoof 2021: accelerating progress in spoofed and deepfake speech detection,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.276822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.414020Z digest=sha256:e8297e4c143d5bc3479380f406ba803cf747dd91aa74ac9f2d81d27c782b8d00

Observation 09a51a00-6777-465e-a99c-eced5c22e306 · outbound

This paper cites ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T10:24:21.420472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:24:21.420472Z digest=sha256:4d5cc2fcf2e75287a4254e948dc234f66d073ae0505161d2089c28742304346d

Observation 7610222b-8f79-4dc3-98e2-65bb3738040f · outbound

This paper cites Robust audio anti-spoofing with fusion-reconstruction learning on multi-order spectrograms,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Robust audio anti-spoofing with fusion-reconstruction learning on multi-order spectrograms,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.253666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.428804Z digest=sha256:bcaa03bdd4d74e7df9e8818a79b8713bf4641a67f4486a31990d86e8f862f7b3

Observation 313ac6e4-e15b-4fd3-9d91-491e9935e3c7 · outbound

This paper cites A comparative study on recent neural spoofing countermeasures for synthetic speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection A comparative study on recent neural spoofing countermeasures for synthetic speech detection,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.234638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.441067Z digest=sha256:c89ab200ef031a3a37b001cbed542911cd71a866e52c448a7a6d7098c8bd9cb1

Observation b9c69ebe-abab-47b5-8bf1-9461210140fe · outbound

This paper cites Channel-wise gated res2net: Towards robust detection of synthetic speech attacks,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Channel-wise gated res2net: Towards robust detection of synthetic speech attacks,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.212866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.448595Z digest=sha256:4ab0312804e96881e31955ce30ec9bb179c508aca2ceeabdc221ed548a109eef

Observation f779088e-a02c-4458-8928-d3d9cdd052c3 · outbound

This paper cites Fastaudio: A learnable audio front-end for spoof speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Fastaudio: A learnable audio front-end for spoof speech detection,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.185669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.460691Z digest=sha256:75723441a1febc05ea17b282c0dbac4f7ea42f9db77654822a43b2244b3f04cb

Observation 53c12967-7231-4a49-92bb-0a3ef463c365 · outbound

This paper cites The effect of silence and dual-band fusion in anti-spoofing system,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection The effect of silence and dual-band fusion in anti-spoofing system,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.152892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.468943Z digest=sha256:4e57a966298cfb433ad111f655162a5c1ce9346b86f23f52a41602f83d87a85b

Observation 09fc9929-97bd-4a22-857c-27ee94efc48d · outbound

This paper cites Towards end-to-end synthetic speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Towards end-to-end synthetic speech detection,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.131004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.474911Z digest=sha256:6438434570aa2cb15fa857b0ed44a7da8191d6c6351787cda6746ad81bce19b4

Observation 7e98b17d-5cbc-4802-8b61-c0c936ae7b81 · outbound

This paper cites Aasist: Audio anti-spoofing using integrated spectro-temporal graph attention networks,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Aasist: Audio anti-spoofing using integrated spectro-temporal graph attention networks,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.104623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.485405Z digest=sha256:1a7970e4e9cc3058ff4a8b3087836f75b9e5d95e87ac923aa4939ad14cce1996

Observation 1d8b3745-16f2-4547-a2c7-439829104472 · outbound

This paper cites Discriminative frequency information learning for end-to-end speech anti-spoofing,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Discriminative frequency information learning for end-to-end speech anti-spoofing,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.074341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.490772Z digest=sha256:2d9d466c5e3b393c002be4d66e59bac1974cd731f63b512fb1ed24f90fe85b7a

Observation 8a3306aa-c5fb-4575-a448-da837e11deee · outbound

This paper cites Robust data2vec: Noise-robust speech representation learning for asr by com- bining regression and improved contrastive learning,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Robust data2vec: Noise-robust speech representation learning for asr by com- bining regression and improved contrastive learning,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.050984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.497043Z digest=sha256:3b9552cd5d658a2a1a54923472f8dded4b6bbb2a44e11e9e625c07fd34ecf416

Observation 233bcfef-9b05-45e5-a7eb-3d707debecbd · outbound

This paper cites Self-supervised learning with cluster-aware-dino for high-performance robust speaker verification,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Self-supervised learning with cluster-aware-dino for high-performance robust speaker verification,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.023588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.503153Z digest=sha256:d267396b55de33b8c43cd4bb18470cb428cc9fc6a5c3e41cd8ffbe9d755b7aba

Observation 5d8db485-0f8b-461d-8f8b-6e9f14e405af · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection wav2vec 2.0: A framework for self-supervised learning of speech representations,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:23.000993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.511661Z digest=sha256:88b07d9108d68270d9e7b9350ed0800d43c9bd737f130cb243483e9f66a8b9ba

Observation b329de89-acad-41eb-a80b-bf587c782dd6 · outbound

This paper cites Hubert: Self-supervised speech representation learning by masked prediction of hidden units,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Hubert: Self-supervised speech representation learning by masked prediction of hidden units,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:24:21.519675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:24:21.519675Z digest=sha256:9c86136cae9c95d5f72431549babe6f23b6d6f8464162a81de344d3b98bd859e

Observation c00d1e3e-d147-4ed4-b3b5-6f6c52e20dfe · outbound

This paper cites Wavlm: Large-scale self-supervised pre- training for full stack speech processing,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Wavlm: Large-scale self-supervised pre- training for full stack speech processing,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.961333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.527515Z digest=sha256:9abf6b08d55c2b527c9128366d540be56404aff8ec3c47218263f16295848a39

Observation 7bcb0426-331c-4d2f-b582-39ffd30ca0fe · outbound

This paper cites Robust spoof speech detection based on multi-scale feature aggregation and dynamic convolution,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Robust spoof speech detection based on multi-scale feature aggregation and dynamic convolution,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.929032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.536857Z digest=sha256:550bb2472bd176fb4a7004e320ad0d8d3bec12a88ed438c267080c52c1018177

Observation 65e212ef-e50f-488d-a048-d3225b0b0feb · outbound

This paper cites Automatic speaker verification spoofing and deepfake detection using wav2vec 2.0 and data augmenta- tion,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Automatic speaker verification spoofing and deepfake detection using wav2vec 2.0 and data augmenta- tion,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.906918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.552238Z digest=sha256:db65067bdb222ffbb8ba5deb0fe50a49c52a07d1dd7ff73ffb8dc3fd113dc218

Observation 1d9f4958-6701-45e7-a0f6-f861346fc972 · outbound

This paper cites Investigating self-supervised front ends for speech spoofing countermeasures,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Investigating self-supervised front ends for speech spoofing countermeasures,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.866635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.561268Z digest=sha256:dd2ec8efd992a94a44a569ee735232ff317617a3e0dcaf3b9c4d36b930828fdc

Observation cd1d6e70-f61b-4430-8906-bc50e621ed06 · outbound

This paper cites Lora: Low-rank adaptation of large language models,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Lora: Low-rank adaptation of large language models,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.842890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.567817Z digest=sha256:daf9ee0dc830894adaa7c9d366fcdfcfb108630a550f0fd7fd10ac1613b3f483

Observation b25906c1-fa8d-4b93-8685-cad55e7f03e8 · outbound

This paper cites Parameter-efficient transfer learning for nlp,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Parameter-efficient transfer learning for nlp,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.817025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.575475Z digest=sha256:ed0c7c9f29abb00c14685bd6d3e8c387059724eb1b9d2d183655e5dcf9f506ce

Observation 6a8a19f9-1d50-4fc6-99b6-7edd2978f635 · outbound

This paper cites Audio deepfake detection with self- supervised xls-r and sls classifier,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Audio deepfake detection with self- supervised xls-r and sls classifier,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.784447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.582486Z digest=sha256:b73e4b491d8d4d23ea7466450a9a8b5f3027a9efbefeabe941a3a7a4b142ce53

Observation 0c23d29d-6f25-480b-b1b9-da2f266669d3 · outbound

This paper cites Attentive merging of hidden embeddings from pre-trained speech model for anti-spoofing detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Attentive merging of hidden embeddings from pre-trained speech model for anti-spoofing detection,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.754575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.589077Z digest=sha256:a84c0b165caeca1d7821ecad6c8eb7d5c2cec0989f7a93a43493002e62fbdfef

Observation a3f4acad-1ad8-4a97-8bb6-34d1718867cf · outbound

This paper cites Outrageously large neural networks: The sparsely-gated mixture-of-experts layer,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Outrageously large neural networks: The sparsely-gated mixture-of-experts layer,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.727153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.597639Z digest=sha256:59fd22f1216660067721235ef11bb8a9c12ce8a382a560b243d9f12c13fdff64

Observation 6d860ee5-5eb3-43d8-96c4-d83eddc48c9f · outbound

This paper cites Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.697394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.610876Z digest=sha256:e9e8c0df1a5ebe1d09887dc8365f3c6e7e7540b4ed71a349790ebf8d49c72f3c

Observation d961a671-a693-4f34-82dd-a6941668f64c · outbound

This paper cites Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.661369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.618851Z digest=sha256:d435b9d984712b77fb3451dddc966bbedbafd0fbf64d4c85df1a93437c4460dc

Observation 7fab0d9b-c156-4d25-989b-a3fea427a85b · outbound

This paper cites Fastdiff: A fast conditional diffusion model for high-quality speech synthesis,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Fastdiff: A fast conditional diffusion model for high-quality speech synthesis,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.633203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.624994Z digest=sha256:cb06eeee0464a4d1c2fcf9c1a62b68dbe272aa5c4f9751f1365ec3b93f35b1e8

Observation fa174add-e3d3-4ae7-bf5f-af9339b3057d · outbound

This paper cites Freevc: Towards high-quality text-free voice conversion,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Freevc: Towards high-quality text-free voice conversion,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.600627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.632325Z digest=sha256:27582ae05f520f0bb67c4b15e45e17edcb9601c585be5c9fe3c86f17c863a10b

Observation aedd6042-2cc0-4a3f-a886-dc06bcd8e49d · outbound

This paper cites Asvspoof 2019: A large- scale public database of synthesized, converted and replayed speech,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Asvspoof 2019: A large- scale public database of synthesized, converted and replayed speech,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.546545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.638598Z digest=sha256:025ce2c0f353468aec9f7b67518a29362585628844e92ff99171f5c614cd29df

Observation 9cdcd84e-a8fa-4c44-ba8d-ff4528ec75d9 · outbound

This paper cites Asvspoof 2021: Towards spoofed and deepfake speech detection in the wild,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Asvspoof 2021: Towards spoofed and deepfake speech detection in the wild,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.524777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.645949Z digest=sha256:ca9102dd6e6daf7b8fcda787ced356d50cc7f49a32e452f567b4e960e577a159

Observation f0870546-ba7c-40c5-979a-c21c66db1b5a · outbound

This paper cites Does audio deepfake detection generalize?,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Does audio deepfake detection generalize?,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T10:24:21.653089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:24:21.653089Z digest=sha256:7ef37e91cf031c5dae13cd43fe84ffd5ce314de27c52bc73cd734fd74bb59a23

Observation 33e3a0e4-df49-433f-bf4a-6bb631854fcf · outbound

This paper cites t-dcf: a detection cost function for the tandem assessment of spoofing countermeasures and automatic speaker verification,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection t-dcf: a detection cost function for the tandem assessment of spoofing countermeasures and automatic speaker verification,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.486726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.670151Z digest=sha256:7b925970c96115c97569bee688771768d13ee85f40cd6071862b0214e06aa3ce

Observation 01b0cda1-daa3-4201-b06f-1a95455a02e3 · outbound

This paper cites Rawboost: A raw data boosting and augmentation method applied to automatic speaker verification anti- spoofing,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Rawboost: A raw data boosting and augmentation method applied to automatic speaker verification anti- spoofing,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.436447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.677332Z digest=sha256:997614825c9378f98c5e4c7509f05ecadb1ed87c2f178657f428b3b1e84fa560

Observation ab76689c-a2af-4eaf-a937-958698e4afd7 · outbound

This paper cites Improving short utterance anti-spoofing with aasist2,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Improving short utterance anti-spoofing with aasist2,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.401284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.692373Z digest=sha256:d43755b2e571f39adf2d90c451e683ce42d3a07343d74f0e20df152c76eabfcf

Observation f77714c3-81f4-4924-9b58-4c158c99b7fc · outbound

This paper cites A conformer-based classifier for variable-length utterance processing in anti-spoofing,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection A conformer-based classifier for variable-length utterance processing in anti-spoofing,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.378433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.711964Z digest=sha256:f60ab224b77f8e85b022f6f1da0b724b7d27edacb66b007375201c69b9087133

Observation ca664df3-d1ec-4a9f-9947-994fd6e85791 · outbound

This paper cites Audio deepfake detection with self-supervised wavlm and multi-fusion attentive classi- fier,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Audio deepfake detection with self-supervised wavlm and multi-fusion attentive classi- fier,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.332612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.723573Z digest=sha256:7774c87f2c4311809633647d198a2681185113cef2315ff7c6833f078a3b5e12

Observation 8bb61193-c48a-465c-94a6-15c24fa27804 · outbound

This paper cites One class learning with adaptive centroid shift for audio deepfake detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection One class learning with adaptive centroid shift for audio deepfake detection,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.300309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.735335Z digest=sha256:9d3e2f5a45dbf60dcc856caba7a7c8453f8f1d8ddc5e7704072602497b5c8433

Observation 5be16f9c-d688-4e58-af75-6d045fd582b8 · outbound

This paper cites Temporal-channel modeling in multi-head self-attention for synthetic speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Temporal-channel modeling in multi-head self-attention for synthetic speech detection,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.273266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.740808Z digest=sha256:197437393b1495ef25517261b513a81e5e24bbc96e3c6d707c97f79f8142c9b2

Observation 16b44864-4e4b-4a03-bc9d-291c03435c4d · outbound

This paper cites A robust audio deepfake detection system via multi-view feature,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection A robust audio deepfake detection system via multi-view feature,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.221833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.748635Z digest=sha256:e058220ba4885be8b4008ca37d961ef432659f84b780a3a044bb6b23145a65c6

Observation 982ae0e7-4a3d-4158-a9a4-e1788409b501 · outbound

This paper cites Spoofed training data for speech spoofing countermeasure can be efficiently created using neural vocoders,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Spoofed training data for speech spoofing countermeasure can be efficiently created using neural vocoders,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.172178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.757104Z digest=sha256:8d2deeb6c48dd11f342bcec7ad1d501aef103ba8e809d55af48e804a09ad8f72

Observation e52fcafd-748c-402c-851c-12c99b215171 · outbound

This paper cites Can large-scale vocoded spoofed data improve speech spoofing countermeasure with a self-supervised front end?,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection Can large-scale vocoded spoofed data improve speech spoofing countermeasure with a self-supervised front end?,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.131597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.767444Z digest=sha256:f33f1a1dc83e6a2642c1796f072873152557348e80c4b0e518f854787a6e4536

Observation 4a7180f7-0ce7-4e35-84fd-32e0a20b35d3 · outbound

This paper cites One-class knowl- edge distillation for spoofing speech detection,.

Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection One-class knowl- edge distillation for spoofing speech detection,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:24:22.105531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:24:21.774639Z digest=sha256:990f6eba135ce03f85266b5939c34bb90c10bc604cf96b72ba51b4f3530c7ad6

Pith citing papers

Observation e60da123-abcf-460e-a92e-e8dbb05497e8 · inbound

Alethia: A Foundational Encoder for Voice Deepfakes cites this paper.

Alethia: A Foundational Encoder for Voice Deepfakes Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:32.948686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:6b64c6e6fcda3b49a098ca8c911a791f812779e482bd4f7e1c1628de1d7b1fae