Pith. sign in

Paper Citation Record · LEDGER

Audio Deepfake Detection at the First Greeting: "Hi!"

As of 6 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2601.19573.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2601.19573 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-16T11:05:33.744633Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-16T11:05:33.744633Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-05-16T11:07:48.301303Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact4
  • verified fuzzy24
  • unresolved0
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40bb36b5-548b-4f0e-9067-a622009676f1 · outbound

This paper cites Audio Deepfake Detection at the First Greeting: "Hi!".

Audio Deepfake Detection at the First Greeting: "Hi!" Audio Deepfake Detection at the First Greeting: "Hi!"

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:07:48.302826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:f1377fccaad60853f0efc742c1d34ac124c6c4a2eac4773e54fec4347a4a07ef

Observation 0272d145-5a0f-49c2-8be6-bd61a2f8fdd1 · outbound

This paper cites Framework Architecture Our work extends the Multi-Granularity Adaptive Time- Frequency Attention (MGAA) framework for ADD [7] to ultra-short utterances (0.5s–2s).

Audio Deepfake Detection at the First Greeting: "Hi!" Framework Architecture Our work extends the Multi-Granularity Adaptive Time- Frequency Attention (MGAA) framework for ADD [7] to ultra-short utterances (0.5s–2s)

Reference 2

Resolution
malformed identifier
raw_fallback, observed 2026-05-16T11:12:47.548000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:25f2bc9f9a189e0f804ce2adfecd26bb76c7ec7239c0c6b80714cda104a75f1c

Observation 45005e85-cdf9-4559-93b7-67450980ad07 · outbound

This paper cites Deep” and “Shallow.

Audio Deepfake Detection at the First Greeting: "Hi!" Deep” and “Shallow

Reference 3

Resolution
malformed identifier
raw_fallback, observed 2026-05-16T11:12:47.534986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:c084d6f85015dbef1b2af2903f728f282f1b332ddff582cde0fc9c4ac0bb05e1

Observation a161ed1d-3018-4725-9f8f-5c10303f52b8 · outbound

This paper cites The framework enhances discriminative repre- sentation learning from limited audio inputs, enabling reliable detection within durations as short as a greeting phrase.

Audio Deepfake Detection at the First Greeting: "Hi!" The framework enhances discriminative repre- sentation learning from limited audio inputs, enabling reliable detection within durations as short as a greeting phrase

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.552761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:b993565acfae4db64e3b3d46bbf54240e4bfa381dbc7648c0b6264bfde970eae

Observation 1fefbf82-8b2f-479e-b822-1b0b32706a6c · outbound

This paper cites A survey on speech deepfake detection.

Audio Deepfake Detection at the First Greeting: "Hi!" A survey on speech deepfake detection

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.554778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:80e4ad350cf8df18fdc3801949e5293a94acce8e8b0c7efa5f994018cf439633

Observation f680f200-6087-486d-95ef-47fdb0a18367 · outbound

This paper cites ASVspoof 2019: Future Horizons in Spoofed and Fake Audio Detection.

Audio Deepfake Detection at the First Greeting: "Hi!" ASVspoof 2019: Future Horizons in Spoofed and Fake Audio Detection

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:07:48.296408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:4577f273bae2f934afaa930eab98decb905a471b25fc1e45f51ac782d0bc9798

Observation ff05468d-5f95-4c31-ac72-3974f86ca448 · outbound

This paper cites Asvspoof 2021: Towards spoofed and deep- fake speech detection in the wild.

Audio Deepfake Detection at the First Greeting: "Hi!" Asvspoof 2021: Towards spoofed and deep- fake speech detection in the wild

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.545173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:71ce6c974f858b0f41c885fb160131653faf7eb3210081e74f09dea68e6707e3

Observation 78078ef2-2817-4516-a26f-3a4d218dc82b · outbound

This paper cites Asvspoof 5: Crowdsourced speech data, deepfakes, and adversarial attacks at scale.

Audio Deepfake Detection at the First Greeting: "Hi!" Asvspoof 5: Crowdsourced speech data, deepfakes, and adversarial attacks at scale

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.532418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:d955e2f38bba39e9e1d0f838e32fb5a188a139b8bb4c155e54de8854204cacfc

Observation c3b4b70a-42c4-4667-8cbb-6ec84d08f241 · outbound

This paper cites Add 2022: the first audio deep synthesis detection challenge.

Audio Deepfake Detection at the First Greeting: "Hi!" Add 2022: the first audio deep synthesis detection challenge

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.537694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:65697eab1ee5bfbc0656cf657651761e422c6467838f50b9aac33e70d7090ccc

Observation 838be841-5a04-4221-b9f2-27641d050b11 · outbound

This paper cites Benchmarking audio deepfake detection ro- bustness in real-world communication scenarios.

Audio Deepfake Detection at the First Greeting: "Hi!" Benchmarking audio deepfake detection ro- bustness in real-world communication scenarios

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.540060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:de378ea5c5dc0a639d3f21353f2b32d8ff29f1626051d9f7b2d00260ca7a0333

Observation 929dec8c-b7be-43d5-a4bc-a70806d3029e · outbound

This paper cites Multi-Granularity Adaptive Time-Frequency Attention Framework for Audio Deepfake Detection under Real-World Communication Degradations.

Audio Deepfake Detection at the First Greeting: "Hi!" Multi-Granularity Adaptive Time-Frequency Attention Framework for Audio Deepfake Detection under Real-World Communication Degradations

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:07:48.299756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:53278ec26a409a48de5fb65262ed6c1ac6be5eefb9578192808b363f0d841541

Observation dcc8a845-f90a-455b-8556-7d856ac0e626 · outbound

This paper cites Aasist: Audio anti- spoofing using integrated spectro-temporal graph atten- tion networks.

Audio Deepfake Detection at the First Greeting: "Hi!" Aasist: Audio anti- spoofing using integrated spectro-temporal graph atten- tion networks

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.550361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:5e327c1401c12fd71d06142088df7f8c027dcb05e5cd25211a55c8b3be663c8a

Observation aa6793e1-fd2f-49c8-8513-965ca704864f · outbound

This paper cites Domain general- ization via aggregation and separation for audio deepfake detection.

Audio Deepfake Detection at the First Greeting: "Hi!" Domain general- ization via aggregation and separation for audio deepfake detection

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.542739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:c81f33666f16c97060a6769118ec65c470c34c778562651673d973ce61eaa1cb

Observation c02beabc-3684-43c0-b769-1a3f35b3a77f · outbound

This paper cites End-to-end spectro-temporal graph at- tention networks for speaker verification anti-spoofing and speech deepfake detection.

Audio Deepfake Detection at the First Greeting: "Hi!" End-to-end spectro-temporal graph at- tention networks for speaker verification anti-spoofing and speech deepfake detection

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.586188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:ee8a10387a666327434506d705fb6395344bdcbbae8ca88e0141cc52c6b48958

Observation 78acbf48-ee40-494f-9a5e-f249dd4677c7 · outbound

This paper cites A comparative study on physical and perceptual features for deepfake audio detection.

Audio Deepfake Detection at the First Greeting: "Hi!" A comparative study on physical and perceptual features for deepfake audio detection

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.588637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:d40a452a8cfdb6d9f8f312248d85069e40e53c7869102949d1a9a4200569214f

Observation 938f9066-47ce-410e-9f98-04b1d06b8316 · outbound

This paper cites A conformer-based classifier for variable-length utterance processing in anti-spoofing.

Audio Deepfake Detection at the First Greeting: "Hi!" A conformer-based classifier for variable-length utterance processing in anti-spoofing

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.577007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:80f6b8ec9d6e70a1bd4830cd94fba028dc6b35d4453771c9af1c0e08c2ac8d40

Observation 13e433ee-fb48-4ae7-97ac-ea7584c9497b · outbound

This paper cites Low- rank adaptation method for wav2vec2-based fake audio detection.

Audio Deepfake Detection at the First Greeting: "Hi!" Low- rank adaptation method for wav2vec2-based fake audio detection

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.579206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:e2ea5bbffa916c39cffd46f1132e6fef3249cb296b483e350d5a14ba0a9bf505

Observation 05359564-6c10-4865-b6ff-91e0c8f3a1ea · outbound

This paper cites Im- proving short utterance anti-spoofing with aasist2.

Audio Deepfake Detection at the First Greeting: "Hi!" Im- proving short utterance anti-spoofing with aasist2

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.590606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:2b700e63373af76a65c5d0a39810416c628cff082fb58ee282075b74c406ac05

Observation 97547fcc-6978-4765-b30d-005e5ea3ef75 · outbound

This paper cites Xlsr-mamba: A dual-column bidirectional state space model for spoofing attack detec- tion.

Audio Deepfake Detection at the First Greeting: "Hi!" Xlsr-mamba: A dual-column bidirectional state space model for spoofing attack detec- tion

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.556904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:93bc9403e4c4a488581b73714a65886d13a8a8e2bd9af00aecec496da1222c76

Observation fe6439f6-00af-493e-bfe4-2e77cf3701e7 · outbound

This paper cites Dynamic ensemble teacher-student distillation framework for light- weight fake audio detection.

Audio Deepfake Detection at the First Greeting: "Hi!" Dynamic ensemble teacher-student distillation framework for light- weight fake audio detection

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.574835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:8d3528c7ded8744be1e50f9d790b1a586b03882faff6414038267729177192cc

Observation c0799b43-19a6-4809-8c63-002762923e7b · outbound

This paper cites The partialspoof database and countermeasures for the detection of short fake speech segments embedded in an utterance.

Audio Deepfake Detection at the First Greeting: "Hi!" The partialspoof database and countermeasures for the detection of short fake speech segments embedded in an utterance

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.569835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:f288c3dbf91a918e1bc026d0946f6aeded2bfbb8138fa03da8975b26626cad53

Observation d03ba036-1944-4188-a613-977ff69fd265 · outbound

This paper cites End-to-end anti-spoofing with rawnet2.

Audio Deepfake Detection at the First Greeting: "Hi!" End-to-end anti-spoofing with rawnet2

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.571941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:0ca7e4a154892f6ceee31b36651a859883f84909dc2f3117bc959398a4eebe00

Observation ef197cc4-dfc6-4899-9bd4-3c98a5f77c83 · outbound

This paper cites For: A dataset for synthetic speech detection.

Audio Deepfake Detection at the First Greeting: "Hi!" For: A dataset for synthetic speech detection

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.581301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:e3bfa748573b43f0270778fc1a0a528d80edbaf888528a0aab55afd3b0287171

Observation 9c2dceaa-600a-41d5-9d19-b8cd4c11baca · outbound

This paper cites WaveFake: A Data Set to Facilitate Audio Deepfake Detection.

Audio Deepfake Detection at the First Greeting: "Hi!" WaveFake: A Data Set to Facilitate Audio Deepfake Detection

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:07:48.305813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:3f62904c68d46e1306df8f9c96c90ddaf421ff3635f4d14c2a219b50d7ce438e

Observation 1c2a4745-7506-483a-8f6c-67c9f12ab7f0 · outbound

This paper cites The lj speech dataset.

Audio Deepfake Detection at the First Greeting: "Hi!" The lj speech dataset

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.567866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:be10ee957e9b7c4fa5ca3dc525f186f7b06f48e56f24d7ad4567b784839bd55f

Observation 83d685a2-ab54-47e6-a736-e7cd680dde3e · outbound

This paper cites Mlaad: The multi-language audio anti-spoofing dataset.

Audio Deepfake Detection at the First Greeting: "Hi!" Mlaad: The multi-language audio anti-spoofing dataset

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.583813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:6ccebf39f0790f2c5829b42897d6cd113b92722d8936846df0adf23fa366c012

Observation aec6bc1b-6eaf-448e-9736-da4e676fd6b2 · outbound

This paper cites The m-ailabs speech dataset.

Audio Deepfake Detection at the First Greeting: "Hi!" The m-ailabs speech dataset

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.563839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:6415dc85e2f65caeb44eace99baf52d837ed3ca5030f2464d225b8c7c022e1ba

Observation 5f731df1-8f89-4178-9eb5-b1cd301f192e · outbound

This paper cites Optimization methods for large-scale machine learning.

Audio Deepfake Detection at the First Greeting: "Hi!" Optimization methods for large-scale machine learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.566002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:c498a25908c968d7e88b38cad763cc52f4035727a4dda82bbcd391a7e0cb686f

Observation 54d73f30-2de9-4df6-8dbb-592a9060c6e7 · outbound

This paper cites Decoupled weight decay regularization.

Audio Deepfake Detection at the First Greeting: "Hi!" Decoupled weight decay regularization

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.559674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:cee1ddf4f3705108bafc4a468f944581591fbcfca500adb4665cd66b96a09bd8

Observation f4cf640d-661b-43c7-b2df-3bd30a015e2e · outbound

This paper cites SGDR: Stochastic gradient descent with warm restarts.

Audio Deepfake Detection at the First Greeting: "Hi!" SGDR: Stochastic gradient descent with warm restarts

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:12:47.561987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:31abf8b82cb7de740dd820c9c3801a1e170d5e546afbf6c5a1d6d6a5df49fbbd

Pith citing papers

Observation 40bb36b5-548b-4f0e-9067-a622009676f1 · inbound

Audio Deepfake Detection at the First Greeting: "Hi!" cites this paper.

Audio Deepfake Detection at the First Greeting: "Hi!" Audio Deepfake Detection at the First Greeting: "Hi!"

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:07:48.302826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:05:33.744633Z digest=sha256:f1377fccaad60853f0efc742c1d34ac124c6c4a2eac4773e54fec4347a4a07ef