Pith. sign in

Paper Citation Record · LEDGER

Alethia: A Foundational Encoder for Voice Deepfakes

As of 4 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 1 inbound Pith citation observation for arXiv:2605.00251.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.00251 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-09T19:27:59.124425Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T01:12:30.668116Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact27
  • verified fuzzy19
  • unresolved6
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch6

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 641850ff-482e-422d-952e-a48fbe8847fa · outbound

This paper cites V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning.

Alethia: A Foundational Encoder for Voice Deepfakes V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-11T15:41:33.204364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:0b898d2cea204993fc1da492dc5bc6c3831816f19721d3520b3265e5d2d550e1

Observation 9d1c0c32-53fb-4c50-9523-59b0afd95d6c · outbound

This paper cites Transferring audio deepfake detection capabil- ity across languages.

Alethia: A Foundational Encoder for Voice Deepfakes Transferring audio deepfake detection capabil- ity across languages

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.440406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:2c675ed3df565bd60784218fed034e88cf48b73414a31abea38fdc532ba62308

Observation 7bbd27c9-76fe-4fe9-b43c-b64d3365c9ba · outbound

This paper cites Learning by Reconstruction Produces Uninformative Features For Perception.

Alethia: A Foundational Encoder for Voice Deepfakes Learning by Reconstruction Produces Uninformative Features For Perception

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:33.295146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:e1e09005a78086377b85ab775b2de9801ccfea93318cae10e0da3eed90ab92d0

Observation 48e383b4-6414-4ccd-9d0e-b2937a6c30c6 · outbound

This paper cites an unresolved cited work.

Alethia: A Foundational Encoder for Voice Deepfakes Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-05-25T21:26:15.437118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:0d228ede6734cd40765adefb92d192fa71512ebb3175291b751586721dcc249b

Observation 620aedc2-c342-4ea4-9c12-2a2fd41977c0 · outbound

This paper cites Deepfake-Eval-2024: A Multi-Modal In-the-Wild Benchmark of Deepfakes Circulated in 2024.

Alethia: A Foundational Encoder for Voice Deepfakes Deepfake-Eval-2024: A Multi-Modal In-the-Wild Benchmark of Deepfakes Circulated in 2024

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-28T02:04:08.212974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:a34af3382adc86b8b569c7eeba236b8a4f227f10c90141e0b6ff85c5f8b1f2ec

Observation a0a9749f-0886-48cb-9feb-c36f9e79e3d5 · outbound

This paper cites CoLLD: Contrastive layer-to-layer distil- lation for compressing multilingual pre-trained speech encoders.

Alethia: A Foundational Encoder for Voice Deepfakes CoLLD: Contrastive layer-to-layer distil- lation for compressing multilingual pre-trained speech encoders

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.443193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:4f90a62ade209bee206c48d8a3b5fbd7a624c19840fcc0147b355da4f262c88f

Observation b3c39e7d-ebbd-417a-bfa5-7ab57dcca0a4 · outbound

This paper cites USAD: Universal Speech and Audio Representation via Distillation.

Alethia: A Foundational Encoder for Voice Deepfakes USAD: Universal Speech and Audio Representation via Distillation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:33.721376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:4591676f82f30be5f63feaa275b11cfd4b733e916be21bb35a0a18b4f7623f19

Observation 959fdc7d-61b5-4113-81c0-59a1e5dfdb88 · outbound

This paper cites Demir¨ors, M., Ozbayoglu, A.

Alethia: A Foundational Encoder for Voice Deepfakes Demir¨ors, M., Ozbayoglu, A

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.427302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:de763ff27172f36978f9539bdd79ef8d994f9379f8154438a6ca8dcda6d48fe4

Observation ec839b8c-84cd-4a80-8dd8-28bb7f05535e · outbound

This paper cites Trident of poseidon: A generalized approach for detecting deepfake voices.

Alethia: A Foundational Encoder for Voice Deepfakes Trident of poseidon: A generalized approach for detecting deepfake voices

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.422498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:f5b9d0fa498bc78f81cf0ba1c7b8da65caaae0eb7eb0c67d135a104e6bb50053

Observation cd5b9cba-767b-4ba0-b146-8cd0edf47473 · outbound

This paper cites T., Manrique, R., and Nunes, B.

Alethia: A Foundational Encoder for Voice Deepfakes T., Manrique, R., and Nunes, B

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.424855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:0e61021657b7fee32ff57881fe50aa2b373f3df1672e21cabf085f6c560981f8

Observation 265e0227-7795-4c2a-a43c-233ec782d76d · outbound

This paper cites ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts.

Alethia: A Foundational Encoder for Voice Deepfakes ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:31.105348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:143cc017dadf57dde1d1c7a03a3b940a05f30feaa3d582a4a089bbcef0b994f6

Observation 7640aea8-f2a2-4703-9bdc-21922a0b421c · outbound

This paper cites Post-training for deepfake speech de- tection.arXiv preprint arXiv:2506.21090.

Alethia: A Foundational Encoder for Voice Deepfakes Post-training for deepfake speech de- tection.arXiv preprint arXiv:2506.21090

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:35.515762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:ae4fa55a6efc3aaeb81924a9dcae39b8a696ef4f3946cd94d93375fa3065d03a

Observation 7933d3d9-79f2-4a52-ab0a-670c1d26897c · outbound

This paper cites ReMASC: Realistic Replay Attack Corpus for Voice Controlled Systems.

Alethia: A Foundational Encoder for Voice Deepfakes ReMASC: Realistic Replay Attack Corpus for Voice Controlled Systems

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-04T23:45:46.948480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:e2e1b97633a315d8a28d255847c93a356cfc6ecb57e9cfb672660fee40850b3b

Observation bfa253f2-f954-4410-b27e-8284d9e9d1f1 · outbound

This paper cites R., Pimentel, A., Avila, A.

Alethia: A Foundational Encoder for Voice Deepfakes R., Pimentel, A., Avila, A

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.434718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:49ed1326a489291726dce6f48d78f1da2905ad41af3b93d0d867ad41eebd556e

Observation ffd426c4-a3c5-4a1e-b09c-22283fe12c2b · outbound

This paper cites An Efficient End-to-End Approach to Noise Invariant Speech Features via Multi-Task Learning.

Alethia: A Foundational Encoder for Voice Deepfakes An Efficient End-to-End Approach to Noise Invariant Speech Features via Multi-Task Learning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:30.569243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:53147f744010b8c3ee2ca2e01e9c6eef91408982a1a1c7fe2325dda46d3c0bee

Observation e60da123-abcf-460e-a92e-e8dbb05497e8 · outbound

This paper cites Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection.

Alethia: A Foundational Encoder for Voice Deepfakes Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:32.948686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:d8b3debf8172209b5aff265d181b5bb415bf97879978674f0fb8f963a27cd028

Observation 22f2ad38-e67a-41a7-a0b6-3621a7587aca · outbound

This paper cites Manipulated Regions Localization For Partially Deepfake Audio: A Survey.

Alethia: A Foundational Encoder for Voice Deepfakes Manipulated Regions Localization For Partially Deepfake Audio: A Survey

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:34.085537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:a32ca98685b4b58066b26c18bf6f1f0628452e55cca57e1d083afe67a1a19a86

Observation bf7a20ec-4abf-4975-8772-f9a17653c7c5 · outbound

This paper cites SpeechFake: A Large-Scale Multilingual Speech Deepfake Dataset Incorporating Cutting-Edge Generation Methods.

Alethia: A Foundational Encoder for Voice Deepfakes SpeechFake: A Large-Scale Multilingual Speech Deepfake Dataset Incorporating Cutting-Edge Generation Methods

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:35.077427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:5726563e34fc6dc58c2ab05f310af3382f9adb94df92c8ab78cba7d6b961de60

Observation 7aec9a64-e12e-4b43-b87a-772f5749caa8 · outbound

This paper cites UniCodec: Unified Audio Codec with Single Domain-Adaptive Codebook.

Alethia: A Foundational Encoder for Voice Deepfakes UniCodec: Unified Audio Codec with Single Domain-Adaptive Codebook

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:32.535227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:1f72a44a5b93cc76df4271e69a505783a29c8a410d4bb299a9ee653e1abf1f34

Observation 004b232c-6971-43c3-8c4b-3c74bd0f50e9 · outbound

This paper cites S., et al.

Alethia: A Foundational Encoder for Voice Deepfakes S., et al

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.414891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:aaf635d5961b70b0d5d89974a9ad5f5d2721e3ea435c0a68cd62803c21d57c68

Observation 95430da1-9ed2-4b28-8b7d-cc3f1eab6971 · outbound

This paper cites Source Tracing of Audio Deepfake Systems.

Alethia: A Foundational Encoder for Voice Deepfakes Source Tracing of Audio Deepfake Systems

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:33.959367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:e405a2490e2a417e373290c99739a4adb26387b13b786d992774abc2fb83dcf4

Observation db490d85-3db8-45f7-a3c5-e367e0dc8cf0 · outbound

This paper cites IndieFake Dataset: A Benchmark Dataset for Audio Deepfake Detection.

Alethia: A Foundational Encoder for Voice Deepfakes IndieFake Dataset: A Benchmark Dataset for Audio Deepfake Detection

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:32.825794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:70c977232defe31a7591701a74a923037f4a0efdd2c56638c1ad3d7ff96ae798

Observation c21f4f4c-7721-4560-a33d-f6a9edee7702 · outbound

This paper cites A survey on speech deepfake detection.ACM Computing Surveys, 57 (7):1–38, 2025a.

Alethia: A Foundational Encoder for Voice Deepfakes A survey on speech deepfake detection.ACM Computing Surveys, 57 (7):1–38, 2025a

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.420026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:03109e7f20456749f0ec2dcdd2e47d951602e898c89966b9c9e283b321878371

Observation 45bb5d01-2935-453b-8288-39b70abcc2d3 · outbound

This paper cites Measuring the Robustness of Audio Deepfake Detection under Real-World Corruption.

Alethia: A Foundational Encoder for Voice Deepfakes Measuring the Robustness of Audio Deepfake Detection under Real-World Corruption

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-07T02:18:24.167264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:bf75d276fd36229e014106fdfd821c9cb948e2228445b558cf35115cbaf4fe40

Observation 303860d8-5039-4cbd-a4bf-4b284ebb126e · outbound

This paper cites Flow Matching for Generative Modeling.

Alethia: A Foundational Encoder for Voice Deepfakes Flow Matching for Generative Modeling

Reference 25

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T15:41:34.564428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:3f3c9e2629a194f8f57952440471ebfe95794961ac2cbaa53dbc197d04b9dc92

Observation 54d60b93-9d0e-4c9f-bbf3-47f739e9f73c · outbound

This paper cites Generative Pre-training for Speech with Flow Matching.

Alethia: A Foundational Encoder for Voice Deepfakes Generative Pre-training for Speech with Flow Matching

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:33.447165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:935fb9956512a294e96a8cfacd7fda2208ba81da20b988ef06968e3dceb75125

Observation 73407916-7033-4ba0-a8ac-9eefc9445da5 · outbound

This paper cites A., and Chng, E.

Alethia: A Foundational Encoder for Voice Deepfakes A., and Chng, E

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.412803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:f8cb71031a02f33f452f0f64d944e83e1594fe8063e8a1bb20aa82b5ad52462e

Observation 3ab933b8-abae-4b07-bd9f-47b793255611 · outbound

This paper cites Can Emotion Fool Anti-spoofing?.

Alethia: A Foundational Encoder for Voice Deepfakes Can Emotion Fool Anti-spoofing?

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:30.964905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:0d4a7cea547b383773545c7d337e4685687c66d51531c43d35cc2ec601418573

Observation 14e2efef-894b-4138-b102-c69a60f8db54 · outbound

This paper cites Discrete audio tokens: More than a survey!.

Alethia: A Foundational Encoder for Voice Deepfakes Discrete audio tokens: More than a survey!

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:31.016439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:2ad51bcfe36e545e9b407aa70f468a0edd835c261e7a429c3bf26055ae24ced2

Observation a032c811-d589-4306-a92b-c6326b7389ac · outbound

This paper cites Replay Attacks Against Audio Deepfake Detection.

Alethia: A Foundational Encoder for Voice Deepfakes Replay Attacks Against Audio Deepfake Detection

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:34.689109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:0ca5cb0dc05666d75162cfc3969f75d15218436684cf8acfd23ceb9f509308aa

Observation d350a536-2e4e-4c06-976a-453ad440e1ac · outbound

This paper cites Speech is Silver, Silence is Golden: What do ASVspoof-trained Models Really Learn?.

Alethia: A Foundational Encoder for Voice Deepfakes Speech is Silver, Silence is Golden: What do ASVspoof-trained Models Really Learn?

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:41:30.779331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:417c022cf139167caa9564b315006dba41bfeb47495bb1389c809c1edfe49f22

Observation da469195-752a-499e-ae69-431c1a6f9e30 · outbound

This paper cites H., Vest- man, V ., Todisco, M., Delgado, H., Sahidullah, M., Yam- agishi, J., and Lee, K.

Alethia: A Foundational Encoder for Voice Deepfakes H., Vest- man, V ., Todisco, M., Delgado, H., Sahidullah, M., Yam- agishi, J., and Lee, K

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.417453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:37ef5239bb9acd92e29e2988f25c46c85a619af7c98965ce5d88757db0192a60

Observation e04cab96-3f1c-47ae-86d1-43213eebbbf4 · outbound

This paper cites an unresolved cited work.

Alethia: A Foundational Encoder for Voice Deepfakes Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-05-25T21:26:15.432278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:b3074f5c9a47a6bea1fc6016fc893c362e614530b03f302d3c228bd2cc3939b7

Observation abfa597c-d941-4c84-855a-86aa33d20697 · outbound

This paper cites Singing voice graph modeling for singfake detection.

Alethia: A Foundational Encoder for Voice Deepfakes Singing voice graph modeling for singfake detection

Reference 34

Resolution
malformed identifier
doi, observed 2026-05-09T19:30:37.083915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:744329836788b1eb38a9f557273489ff76939286160a2688bec34869953be112

Observation 73cbbff2-7736-4303-bc6e-4260e8bc823b · outbound

This paper cites SRC4VC: Smartphone-Recorded Corpus for Voice Conversion Benchmark.

Alethia: A Foundational Encoder for Voice Deepfakes SRC4VC: Smartphone-Recorded Corpus for Voice Conversion Benchmark

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:31.478174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:1f6025973efa0fa9c3b3612cdf06042947586add67594e49ebf6be1eb3d7477a

Observation 03a1d533-cd23-418f-bf37-2266a05e30a1 · outbound

This paper cites JSUT corpus: free large-scale Japanese speech corpus for end-to-end speech synthesis.

Alethia: A Foundational Encoder for Voice Deepfakes JSUT corpus: free large-scale Japanese speech corpus for end-to-end speech synthesis

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:34.878535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:2fd4eb864eee83459918fc9f1ea13785fc89aebe8726fe5b5e707b5fbd4be986

Observation 30c8161e-e91a-4a12-a01a-09b12f8d5b39 · outbound

This paper cites Rawboost: A raw data boosting and augmentation method applied to automatic speaker verification anti-spoofing.

Alethia: A Foundational Encoder for Voice Deepfakes Rawboost: A raw data boosting and augmentation method applied to automatic speaker verification anti-spoofing

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.429930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:88c56b7f489be9025db74bf8a90856bf687ae60bf1817d2779bf0c5cbb123fca

Observation 6e8d119b-6a94-4ae4-b973-2d391c135383 · outbound

This paper cites JMD: Japanese multi-dialect corpus for speech synthesis.

Alethia: A Foundational Encoder for Voice Deepfakes JMD: Japanese multi-dialect corpus for speech synthesis

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.380956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:cc39daf2f8ad5b5a47004475e771d96b73053b2c0cdfad26c8e07a5ca1bb654a

Observation c15dff90-1d12-423e-baa2-7835ea0f993e · outbound

This paper cites JSSS: free Japanese speech corpus for summarization and simplification.

Alethia: A Foundational Encoder for Voice Deepfakes JSSS: free Japanese speech corpus for summarization and simplification

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:31.707924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:e780935673db7f9b9656758ab13c641ce1e32360a6f40b9ad1d5320bb8b7cbe9

Observation aa430382-204a-426d-b3c3-a247a26503d3 · outbound

This paper cites ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale.

Alethia: A Foundational Encoder for Voice Deepfakes ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:35.345209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:374f715a6a6e53f7e263c261f9a77e0bb3d8fe4a77c7b5630b0ad624c114b4ef

Observation 756857e7-d796-44cd-9e2a-c0ca260e2418 · outbound

This paper cites CodecFake: Enhancing Anti-Spoofing Models Against Deepfake Audios from Codec-Based Speech Synthesis Systems.

Alethia: A Foundational Encoder for Voice Deepfakes CodecFake: Enhancing Anti-Spoofing Models Against Deepfake Audios from Codec-Based Speech Synthesis Systems

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:34.406645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:dc89d38603be53f31db7d09d433baf38b10c9b42fc90c5f8d6692a005edf3587

Observation 0c4fa906-d077-4b9b-9385-a08bd0821cd7 · outbound

This paper cites Neural Codec Source Tracing: Toward Comprehensive Attribution in Open-Set Condition.

Alethia: A Foundational Encoder for Voice Deepfakes Neural Codec Source Tracing: Toward Comprehensive Attribution in Open-Set Condition

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:41:32.564889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:8258f7f244f33fa98dceaedff26f1c4c2486f5dd97b798257f9ae03af819e265

Observation d1459f72-da4a-4128-91d9-16cff074ccbf · outbound

This paper cites ASVspoof 2021: accelerating progress in spoofed and deepfake speech detection.

Alethia: A Foundational Encoder for Voice Deepfakes ASVspoof 2021: accelerating progress in spoofed and deepfake speech detection

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:41:31.868407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:47bb417076833c79b5fc8c6283138b92ba289fbe40d37cf38a29a14ad2b78fec

Observation de63e1e9-9ecb-4c86-80fa-fdce143c610e · outbound

This paper cites MSceneSpeech: A Multi-Scene Speech Dataset For Expressive Speech Synthesis.

Alethia: A Foundational Encoder for Voice Deepfakes MSceneSpeech: A Multi-Scene Speech Dataset For Expressive Speech Synthesis

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:35.279155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:0d2da88fb0cc9e670a2989c350b3319352dd760bbdce060be446ee2a58851b6f

Observation 1d174128-92b1-4b5d-90be-656fa86a94bc · outbound

This paper cites SUPERB: Speech processing Universal PERformance Benchmark.

Alethia: A Foundational Encoder for Voice Deepfakes SUPERB: Speech processing Universal PERformance Benchmark

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:41:33.505238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:2349907355fa00444e2e03efccfb7f83b793e79e313085d7f5de10f716983d7a

Observation 9df8832f-7b6d-434f-8990-93f145edc78a · outbound

This paper cites SPEAR: A Unified SSL Framework for Learning Speech and Audio Representations.

Alethia: A Foundational Encoder for Voice Deepfakes SPEAR: A Unified SSL Framework for Learning Speech and Audio Representations

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-06-23T04:13:40.777736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:c7ac4fd3a1534e8d73b57c041e7695aef7ae5a27b51f0040384df7706facc130

Observation 57bc7879-8a94-460a-82e6-a68a2874a80b · outbound

This paper cites Half-Truth: A Partially Fake Audio Detection Dataset.

Alethia: A Foundational Encoder for Voice Deepfakes Half-Truth: A Partially Fake Audio Detection Dataset

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:32.138168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:09cc060dd31400cf4cb45801e1d516f3afee47045bcf3cd80a86e317b9bd9619

Observation 9b8e9469-d344-4d12-b3e9-869e0c2e5a80 · outbound

This paper cites Singfake: Singing voice deepfake detection.

Alethia: A Foundational Encoder for Voice Deepfakes Singfake: Singing voice deepfake detection

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.378444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:97f4fb7d39087783984c1af8ac36b73836b5159d60420f0d820efd08eb206ea3

Observation 18842dae-1a5e-48ea-8bba-30bff6ddf03d · outbound

This paper cites Audio deepfake detection: What has been achieved and what lies ahead.Sensors (Basel, Switzerland), 25(7):1989.

Alethia: A Foundational Encoder for Voice Deepfakes Audio deepfake detection: What has been achieved and what lies ahead.Sensors (Basel, Switzerland), 25(7):1989

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.384022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:783ebd82111073a0eaf91d85ee577ffdf05986d52f84146cde70b36e7f2b322b

Observation dd743721-520f-4299-b43f-758fd58a2cf5 · outbound

This paper cites SVDD 2024: The inaugural singing voice deepfake detection challenge.

Alethia: A Foundational Encoder for Voice Deepfakes SVDD 2024: The inaugural singing voice deepfake detection challenge

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.387024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:45b51bf112f40921f58617ee3efbb6436a320eef63865a440a4cf101e71fcb28

Observation 443d723c-b3cb-413d-8d98-66e5d4368898 · outbound

This paper cites AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors.

Alethia: A Foundational Encoder for Voice Deepfakes AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-06-04T02:07:39.638614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:9ffa50443a78e5d9faad19d0bdb30df6186a6e1f7e7d450b9851c117b1e09199

Observation 5ccc53a1-bc74-456c-80b1-75dfd457111a · outbound

This paper cites an unresolved cited work.

Alethia: A Foundational Encoder for Voice Deepfakes Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-05-25T21:26:15.392964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:c8b656c3d114248d609cb770accff2988d05af04d76d21a874453dd22170c715

Observation ed41190a-b2a8-414b-859c-6980c7572207 · outbound

This paper cites all-step.

Alethia: A Foundational Encoder for Voice Deepfakes all-step

Reference 53

Resolution
malformed identifier
raw_fallback, observed 2026-05-25T21:26:15.375713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:fa61ae65182638f17e2a6cd3ac3899c6a4bd1d67dd4b1b87f74d6f8688e5cca4

Observation 961a8366-95c9-4244-9868-3a701b2f7757 · outbound

This paper cites Other Tasks PFSL.For dataset configuration and model inference, we adapt the framework provided by Luong et al.

Alethia: A Foundational Encoder for Voice Deepfakes Other Tasks PFSL.For dataset configuration and model inference, we adapt the framework provided by Luong et al

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.390327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:55fba5d8c61f8447949c1077c10bcce3a3ea7628007d6fd041e2afd1f9903e7a

Observation 9c7f63d8-7453-4836-b9d6-95cc554222ad · outbound

This paper cites A VDD.For this task, we adapt code from the FakeA VCeleb repository2.

Alethia: A Foundational Encoder for Voice Deepfakes A VDD.For this task, we adapt code from the FakeA VCeleb repository2

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.401907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:b88ecb9ce4d1fc169f56a8454d0ead09b6f98879511ecae3b36e0afa1f404d86

Observation c2034fe2-0aa4-4feb-b8a1-db9d751822a6 · outbound

This paper cites an unresolved cited work.

Alethia: A Foundational Encoder for Voice Deepfakes Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-05-25T21:26:15.404444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:11606501e09228283f3c3347199a369ab0be3d0c3b1ee4673fb0e1151734c646

Observation cf07b65d-a843-4998-aab0-4a0d012b14e0 · outbound

This paper cites preprocessed.

Alethia: A Foundational Encoder for Voice Deepfakes preprocessed

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.407306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:a946ec271cf5bb47f7d61a535fd8fef1e649a9bc3a4404e3dfffa074fe49f8b2

Observation 6704ddd1-53eb-40f8-86b6-04feb069d2e6 · outbound

This paper cites an unresolved cited work.

Alethia: A Foundational Encoder for Voice Deepfakes Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-05-25T21:26:15.396178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:5346cafea2817c459d9dd6e2366756e9665976415052e42dc6d6c635366dc3a2

Observation 3e3e94e6-e8af-4352-84c8-5cb0397b6f33 · outbound

This paper cites an unresolved cited work.

Alethia: A Foundational Encoder for Voice Deepfakes Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-05-25T21:26:15.398943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:7ac2e20411378275e205fc3ce20c0c722114a825c7b581e70b9ebc8878b3e0ae

Observation 88836099-047e-4aee-b12a-d881db205f19 · outbound

This paper cites Other Experimental Results G.1.

Alethia: A Foundational Encoder for Voice Deepfakes Other Experimental Results G.1

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T21:26:15.410512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:c4d6e2c2f54ab6af7813a82240acb95bede8c9ee70b5403922a93808effbcb5a

Pith citing papers

Observation 31613142-6035-4ac0-af9c-3f340d2897df · inbound

Large Audio Language Models for Spoofing-Aware Speaker Verification cites this paper.

Large Audio Language Models for Spoofing-Aware Speaker Verification Alethia: A Foundational Encoder for Voice Deepfakes

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T01:12:30.668116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:12:30.668116Z digest=sha256:576d8cb1814e2e56e641740e3cc93e00c7ea9ffefb8aef8aeca1f23062b121c8