Pith. sign in

Paper Citation Record · LEDGER

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge

As of 11 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 2 inbound Pith citation observations for arXiv:2606.10791.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.10791 v2

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-29T05:19:26.613301Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T01:12:30.652884Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact4
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 232f632b-1db2-4788-92c3-4c5dec6ffe09 · outbound

This paper cites Compspoof: A dataset and joint learning framework for component- level audio anti-spoofing countermeasures,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Compspoof: A dataset and joint learning framework for component- level audio anti-spoofing countermeasures,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:323781178d1c0884ae2a8cd60e0e21f57b09f305a06e84dfa96207e9fd1a4ade

Observation 19d3c7f0-05cf-44db-bb8c-db9954dd9bc4 · outbound

This paper cites Environmental sound deepfake detection challenge: An overview,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Environmental sound deepfake detection challenge: An overview,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:8e47bd6ba196cf2211686e56f415420139daa2687668078bc8f6253c2ca8d02d

Observation 2bf0a484-d791-4479-8812-b50d30a90bc8 · outbound

This paper cites Esdd 2026: Environmental sound deepfake detection challenge evalu- ation plan.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Esdd 2026: Environmental sound deepfake detection challenge evalu- ation plan

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:13:45.457367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:3f430d4f77db36e442b318850a968e9b98600c03f5fb2a1a491797ca84467c1a

Observation b57569d5-309e-4f9d-8d95-3de276efdd9f · outbound

This paper cites The First Environmental Sound Deepfake Detection Challenge: Benchmarking Robustness, Evaluation, and Insights.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge The First Environmental Sound Deepfake Detection Challenge: Benchmarking Robustness, Evaluation, and Insights

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-06-29T17:13:45.460141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:827ee44649d56f1a181a8210ab53b7e9269659bf1081afd71590aa82d0c0b3fc

Observation 9ab58271-24bf-4981-9bf9-8086b5837f7c · outbound

This paper cites Audiocaps: Generating captions for audios in the wild,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Audiocaps: Generating captions for audios in the wild,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:c2e19fdf21c6252916646e418c4974b79892f5f08446025dd6f8a5e382140943

Observation 042c6abe-6b0e-4fd5-8bf6-590efff56f36 · outbound

This paper cites Vggsound: A large-scale audio-visual dataset,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Vggsound: A large-scale audio-visual dataset,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:d7d0f0abbea02898fc1056a0acf221d9c4ab5a87d51e554565859f0048096906

Observation 07539c61-dfc3-428f-9337-3b6ef6df7c10 · outbound

This paper cites Common voice: A massively-multilingual speech corpus,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Common voice: A massively-multilingual speech corpus,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:639af653cb931fdc4f19fa50a7660d00492eddbc4dcbc75e5170b3b42cd2fd9f

Observation 56ae0c1a-0697-4f19-a148-40ae387b432a · outbound

This paper cites Libritts: A corpus derived from librispeech for text-to-speech,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Libritts: A corpus derived from librispeech for text-to-speech,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:1316605a6a21df364af9f168ae1b68288dc7ebdbf2a40ca26654830cbbccb60a

Observation f5336ba8-0f17-41cb-990e-8f9ceda0d07d · outbound

This paper cites Enhancing Speaking Styles in Conversational Text-to- Speech Synthesis with Graph-Based Multi-Modal Context Modeling,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Enhancing Speaking Styles in Conversational Text-to- Speech Synthesis with Graph-Based Multi-Modal Context Modeling,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:80b730788c55063c75519946c75b65e1681c8ded1a60f2b8d03e99391dac1d21

Observation 3ff9eef4-abe3-4784-a5eb-70cfeb577d1c · outbound

This paper cites TAU urban acoustic scenes 2019 open set, development dataset,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge TAU urban acoustic scenes 2019 open set, development dataset,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:0f0f0aca210c0d3c11b2346d522e2d370a190a21bfa9be4f04e4084ef2e527f7

Observation f996d014-f2a6-4539-9b04-341c18729f27 · outbound

This paper cites Tut database for acoustic scene classification and sound event detection,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Tut database for acoustic scene classification and sound event detection,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:b1323d4ec06f04ee42bba6d61779acaa3b5b7ab170780294a3abbdc004f82923

Observation 48cdeae3-1491-4c41-84fa-7d32dcf00ba7 · outbound

This paper cites Tut sound events 2017, development dataset,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Tut sound events 2017, development dataset,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:34095b0ec72b17c9974a8b19a82349f42ac10682542ef181eb6c0bdd170da5f7

Observation 253bd668-0da4-4fef-a7b1-ac8188142c26 · outbound

This paper cites Tut sound events 2017, evaluation dataset,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Tut sound events 2017, evaluation dataset,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:613f22530bafc6ad2c153714d4dac24b3c58875ad47b66e6857c75272e3d30e2

Observation c5f8cd24-7d0f-48fc-aa2e-a672c9eff077 · outbound

This paper cites A dataset and taxonomy for urban sound research,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge A dataset and taxonomy for urban sound research,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:de9ed9e86d70dbde91a2d0070ea22fdacaaf026585c4584ab91601b5a7e0f1bc

Observation 9fee1ee3-771f-4ddf-9651-48ab1da8805c · outbound

This paper cites Envsdd: Benchmarking environ- mental sound deepfake detection,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Envsdd: Benchmarking environ- mental sound deepfake detection,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:c285f4eadbcfc0c371f4a9f68d0430c7c92abbc3f324c196bec828e26e22fbb9

Observation 5c1e80e0-bc0a-40bb-9e6a-9f96660dfd06 · outbound

This paper cites VCapA V: A Video-Caption Based Audio-Visual Deepfake Detection Dataset,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge VCapA V: A Video-Caption Based Audio-Visual Deepfake Detection Dataset,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:1f58c8a7c015818e06c0e49ddae7bf62268d748a2e22d6bf88611bd8890994cc

Observation dcaa45db-68c8-48af-a0a2-6b1e3ebf5e71 · outbound

This paper cites Asvspoof 5: crowdsourced speech data, deepfakes, and adversarial attacks at scale,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Asvspoof 5: crowdsourced speech data, deepfakes, and adversarial attacks at scale,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:399c67db0b29b6b110e1ae33f75bd8a76f8e6884d16107936e9a31d23025d688

Observation c9b0f8ee-ca98-4520-9edc-d0aa8b844723 · outbound

This paper cites Mlaad: The multi-language audio anti-spoofing dataset,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Mlaad: The multi-language audio anti-spoofing dataset,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:48e5c61e7f0b228b528e02d1219a0238e203697dc1bea5ee3150233310337602

Observation 45700f7c-301d-4e92-8be6-0916ff1e4497 · outbound

This paper cites Deepfake Audio Detection Using Self-supervised Fusion Representations.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Deepfake Audio Detection Using Self-supervised Fusion Representations

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-06-29T17:13:45.451160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:0449feb1c3fc9f9e76a2bbd26a58fc09a7d7bbb6dc94a55c50476ce302111d55

Observation 34989123-f903-4c8c-a4a0-320f63b578d3 · outbound

This paper cites Xls-r: Self-supervised cross-lingual speech representation learning at scale,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Xls-r: Self-supervised cross-lingual speech representation learning at scale,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:5a54410b39af96bd6b7fa7a60f2b3b3a15a8851db1d4594674d123950ec49ee7

Observation 4bc75510-8d98-4fe5-b805-e10e7612c8d3 · outbound

This paper cites Eat: self-supervised pre-training with efficient audio transformer,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Eat: self-supervised pre-training with efficient audio transformer,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:de556dda93ed8fa574a1afa83f6fd8de3e79d2e793a82ae14d93fa65867af9fd

Observation 215b379f-5aa8-4fe5-8d6f-64555c8b5c2a · outbound

This paper cites Sslam: Enhancing self-supervised models with audio mixtures for polyphonic soundscapes,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Sslam: Enhancing self-supervised models with audio mixtures for polyphonic soundscapes,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:9a41bb4c85043778b6bb5c2b5d819f4f76e4dd417aac6797195d015256f191ef

Observation 7c2df84a-9a4a-41aa-97e5-802c8f3a8665 · outbound

This paper cites Scaling up masked audio encoder learning for general audio classification,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Scaling up masked audio encoder learning for general audio classification,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:7fec8a31cbbcb14cca944c6891cdbdeeaa4845c010492ac40114c11f0d841dd3

Observation 403b2a7d-c133-40e2-b9c2-cde0f0be6a0d · outbound

This paper cites Do compact ssl backbones matter for audio deepfake detection? a controlled study with raptor.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Do compact ssl backbones matter for audio deepfake detection? a controlled study with raptor

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:13:45.454386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:3213342c8992ccac09f01f0335e553f81443665fefec1cb76fb28bb215f41dac

Observation ad31a4fe-884f-4eaf-9ea1-359698f3885f · outbound

This paper cites Xlsr-mamba: A dual-column bidirectional state space model for spoofing attack detection,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Xlsr-mamba: A dual-column bidirectional state space model for spoofing attack detection,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:82344b5d8cdb27de56429d2efd5ddbe54bc22e51b4cf8d177ac33bc57022c752

Observation e5c67944-6bc0-4057-b0a0-b6c4dd7392ce · outbound

This paper cites Audio deepfake detection with self-supervised xls-r and sls classifier,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Audio deepfake detection with self-supervised xls-r and sls classifier,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:cc23f3def90290aae10fe551a0812face6d3363e41ed09ab42123d3e8c097ab8

Observation 7a9f7525-0672-4550-9bcf-38d0ea49b8b0 · outbound

This paper cites Temporal-channel modeling in multi-head self-attention for synthetic speech detection,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Temporal-channel modeling in multi-head self-attention for synthetic speech detection,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:1800465abe896a2ee93d8414191172d186c9b6cbd67e48d59f75da6c6187e914

Observation d4b91ff5-e510-412a-b532-4bdf4fd1ddfd · outbound

This paper cites Rawboost: A raw data boosting and augmentation method applied to automatic speaker verification anti-spoofing,.

Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Rawboost: A raw data boosting and augmentation method applied to automatic speaker verification anti-spoofing,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-29T05:19:26.613301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:19:26.613301Z digest=sha256:df440bf269547b66b8fd589a12f0da0167c9fc5278b9e5ce36f81fee1151e450

Pith citing papers

Observation 5bf1ba35-306a-4e43-8ba6-63af0b17ca99 · inbound

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation cites this paper.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:768b2f98d79a7ef296cf42433eb799aa30b45064ef663ece163f7022624c0277

Observation d1286339-70b2-46c1-99ef-d8ef0b441fa1 · inbound

Large Audio Language Models for Spoofing-Aware Speaker Verification cites this paper.

Large Audio Language Models for Spoofing-Aware Speaker Verification Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T01:12:30.652884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:12:30.652884Z digest=sha256:1713614c58b8296de23d307beff100a16214905f322cc0e3ef5caf0cf10ae572