Pith. sign in

Paper Citation Record · LEDGER

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation

As of 20 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 0 inbound Pith citation observations for arXiv:2607.04848.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.04848 v1

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-11T12:34:20.057072Z

measured 33 of 33 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

33 of 33 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved29
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b31a340e-8cf2-4574-b2cd-126cebe3975b · outbound

This paper cites Audio deepfakes: A survey,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Audio deepfakes: A survey,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:9dd17270bd142c2979c37f241c84610ef7affc7086a264ff6c51ca106c757326

Observation b784e968-7d39-4882-b3fb-5c68b8493bbd · outbound

This paper cites Available: https://www.frontiersin.org/journals/big- data/articles/10.3389/fdata.2022.1001063.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Available: https://www.frontiersin.org/journals/big- data/articles/10.3389/fdata.2022.1001063

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:c8783565da2ce1086a359367ca4c7776ed8d982468d9f8ae619122ec899eeb23

Observation 548ce5f1-1e12-4bee-897a-0985d5b8c57e · outbound

This paper cites Audiobox: Unified audio generation with natural language prompts,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Audiobox: Unified audio generation with natural language prompts,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:f85a60d6ce04793b8ce2cd90487bddb53ac0056826da25e763daedb52e61f4ea

Observation a3d23cec-77d5-403f-9493-0bf03a9ec394 · outbound

This paper cites Audiobox: Unified Audio Generation with Natural Language Prompts.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Audiobox: Unified Audio Generation with Natural Language Prompts

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:df3c670fe2b5e418a197c29495b30b8bb5e1b3e6978c9efa580a8cb11cb7bffb

Observation 257c9411-0d83-40b1-9986-577e822c9546 · outbound

This paper cites Asvspoof 2021: Towards spoofed and deepfake speech detection in the wild,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Asvspoof 2021: Towards spoofed and deepfake speech detection in the wild,

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-11T12:38:00.628922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:0f9d7a05474d6cfdafb6303072934e0d2cdb03173e0a95edc7d4c532737e3a10

Observation 060d2041-d4bf-4a37-a1ee-64839087e3c3 · outbound

This paper cites Environmental sound recognition: A survey,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Environmental sound recognition: A survey,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:a468c44f4525e79eb5c7bc0870d2a8403b9cc611990464606c15ad4784324745

Observation 6723f634-3684-4055-9022-570d53feff0a · outbound

This paper cites Audio Surveillance: a Systematic Review.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Audio Surveillance: a Systematic Review

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:87381c946b417af4c8aa173ca5f68632ff1b1e28e143d430b9b08de3c5856ceb

Observation 1a336e2f-a948-40e8-8210-740617a19f44 · outbound

This paper cites ASVspoof 2019: Future horizons in spoofed and fake audio detection,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation ASVspoof 2019: Future horizons in spoofed and fake audio detection,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:9b475c646e201ec135a68ea2b7201f98546a951b96c3ec7789322123bd782f4e

Observation 55b29a97-e982-492a-903e-26b198859c87 · outbound

This paper cites ASVspoof 2021: Automatic speaker verification spoofing and countermeasures challenge evaluation plan,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation ASVspoof 2021: Automatic speaker verification spoofing and countermeasures challenge evaluation plan,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:3e4ea2e089edc8720a57d3baee6dc59b0c2afc15a92f991153653790966ba6cb

Observation e9aa5ae3-14f1-40e6-973f-331f647b1b22 · outbound

This paper cites FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:32a68b7737b986245cd23eaa0fed5ac87814833204f176db516c69aba5a3c079

Observation 3e6779e1-87aa-4b93-aad5-977786733425 · outbound

This paper cites TO-Rawnet: Improving RawNet with TCN and Orthogonal Regularization for Fake Audio Detection.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation TO-Rawnet: Improving RawNet with TCN and Orthogonal Regularization for Fake Audio Detection

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:3575acddd62153c71bbc5b2cb27ae29ceda24faa0925b308c5268677045ec904

Observation b904a9a1-d87c-4620-b28a-6b45b8a2ea24 · outbound

This paper cites Twice attention networks for synthetic speech detection,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Twice attention networks for synthetic speech detection,

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-11T12:38:00.611125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:f9578a91e14c162cc2e70feddc55222b494a48b9f73fc1c8915d0b0e61db820d

Observation 0501abae-9a1c-49fb-92c4-141c5446984b · outbound

This paper cites Wavefake: A data set to facilitate audio deepfake detection,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Wavefake: A data set to facilitate audio deepfake detection,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:df31dd977b6805c80874694401ad034aca69113cd1247f9de777dde8e6419fb7

Observation 7e86ee54-e84c-4270-a00c-aef23850b9b1 · outbound

This paper cites Multi- lingual deepfake speech dataset for robust and generalizable detection,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Multi- lingual deepfake speech dataset for robust and generalizable detection,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:afaf7de3644aa3954887a8fc3ee343d698f975c798556f20f7efaa4caa653ad8

Observation e929f1ce-217c-4328-9d7b-9d9299f121cb · outbound

This paper cites Simple and controllable music generation,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Simple and controllable music generation,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:1c7218fe0ecb96e357549ec41574f78ed217e7f5ad28fe6393ce561b786b0589

Observation 6b618a80-5245-450a-b915-ba7b6712dcc8 · outbound

This paper cites AudioLDM: Text-to-Audio Generation with Latent Diffusion Models.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation AudioLDM: Text-to-Audio Generation with Latent Diffusion Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:42f7fe36d3b28a580c80f2ab1336d4d6012e058097ec48fa4d6c05f4d5f6bd82

Observation 6655ad29-f814-4a34-aaf4-c09ed6593027 · outbound

This paper cites AudioLDM 2: Learning Holistic Audio Generation with Self-supervised Pretraining.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation AudioLDM 2: Learning Holistic Audio Generation with Self-supervised Pretraining

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:cfed11857e967c2ab0c1d13eacc17201d08fe7f80ee4a03f66cdf0cd726cb594

Observation a6e4163b-9bb6-49a9-a4ca-347fa63eeb76 · outbound

This paper cites Stable audio: Fast timing-conditioned latent audio diffusion,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Stable audio: Fast timing-conditioned latent audio diffusion,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:5ba1aa642cca3a19f5dc905a386641bf4481bbf2804b121eea377a597738571c

Observation ebd3d5fd-e6bf-4fe5-a15b-a259e7a3eb97 · outbound

This paper cites Available: https://stability.ai/research/stable-audio-fast- timing-conditioned-latent-audio-diffusion.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Available: https://stability.ai/research/stable-audio-fast- timing-conditioned-latent-audio-diffusion

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:5654e97f01f0542215b01ee2419a58cf5688abb014cca3c273db93fa716ff439

Observation 59f4529d-a39f-4692-a3f8-c01a8d30b9ad · outbound

This paper cites Diffsound: Discrete diffusion model for text-to-sound generation,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Diffsound: Discrete diffusion model for text-to-sound generation,

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-11T12:38:00.631632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:df229bbfe8eda20295b68c8db421576144bf3bd1d9ce552a2fc5e148db620e3b

Observation bc96dca8-13c6-4bd5-963f-c5bb8a245132 · outbound

This paper cites Make-an-audio: Text-to-audio generation with prompt-enhanced diffusion models,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Make-an-audio: Text-to-audio generation with prompt-enhanced diffusion models,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:f419b3e50b208009639cf9e33890872c053b22e00e6225ddadff80bee6b0e26e

Observation 85d98edd-c600-454c-a9dc-983ee4baa906 · outbound

This paper cites MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:1a1cd685733b01eb2858deb3c651a455914c7f13b8abaa3bea9b88f4eece31ad

Observation 8b3a5318-e541-410b-ab9d-f7fa63760899 · outbound

This paper cites TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:b0871c1a723cb625a58b85ee54158e175453484a53294071c48634adae3caa40

Observation 7f73c209-a9d3-45c4-984e-6d42919806ec · outbound

This paper cites Envsdd: Benchmarking environmental sound deepfake detection,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Envsdd: Benchmarking environmental sound deepfake detection,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:100025c02c4dac3cde422254445dc259750ddaeb03f53dac74badc4dcf65c823

Observation 2cf5ffeb-b342-463b-a9f5-2db4f313ef98 · outbound

This paper cites Compspoof: A dataset and joint learning framework for component-level audio anti-spoofing countermeasures,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Compspoof: A dataset and joint learning framework for component-level audio anti-spoofing countermeasures,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:c54134313aced38c5bb915f9c7c583e6e172e79429824061e06bfd4ddefb990e

Observation 5bf1ba35-306a-4e43-8ba6-63af0b17ca99 · outbound

This paper cites Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:de946fc56c3d787e4079341568bc7cf9df2db6b34167044b1168dbc86503c1d4

Observation 869d5770-2430-43ba-9208-5f21e1d6f003 · outbound

This paper cites Efficient audio transformer and aasist for environment sound deepfake detection in the esdd 2026 challenge,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Efficient audio transformer and aasist for environment sound deepfake detection in the esdd 2026 challenge,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:59f23023feb353e3c1d320ebd99cbabf5220136abe00504f3ddb6943758b564d

Observation 9c9316cb-aba9-4b1b-969f-d713cf7a2147 · outbound

This paper cites Clotho: an audio captioning dataset,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Clotho: an audio captioning dataset,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:d31d01cf0506cb0b632e2dd8c76c4cc78151c066d495b0af72c656d132e7ac19

Observation 9c414d3e-7c7b-461e-b38e-b9a14b3d2470 · outbound

This paper cites an unresolved cited work.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:724cd7d517d6dee2dd6e2627cddbe3756563cc4e5ceab3735631dc5a293a11fa

Observation bf598def-c5fb-4b6e-8134-b6cb613b0a98 · outbound

This paper cites TACOS: Temporally-aligned audio captions for language-audio pretraining,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation TACOS: Temporally-aligned audio captions for language-audio pretraining,

Reference 30

Resolution
malformed identifier
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:43fcfbe84852ab08125fa3a30167d5ef015757515da230aea989aca276875213

Observation e946d516-1ce3-42b5-b897-81db2a2cbc5d · outbound

This paper cites WavCaps: A ChatGPT-assisted weakly-labelled audio captioning dataset for audio-language multimodal research,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation WavCaps: A ChatGPT-assisted weakly-labelled audio captioning dataset for audio-language multimodal research,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:c8a8add23c22afe1d72d7323bbfb4bbbf308d875c1e37a44039b6f4035e7200e

Observation 002ccebc-1351-4527-83be-c7d0c55709e4 · outbound

This paper cites AASIST: Audio anti-spoofing using integrated spectro-temporal graph attention networks,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation AASIST: Audio anti-spoofing using integrated spectro-temporal graph attention networks,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:4d8cf150f9f5e4e351a29e868c282657508df9c171340785396f54440ade7a6f

Observation e0f64efa-bbe6-4238-b750-25c4dd1a3d78 · outbound

This paper cites A dataset and taxonomy for urban sound research,.

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation A dataset and taxonomy for urban sound research,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-11T12:34:20.057072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:34:20.057072Z digest=sha256:6864d391567a7f987fd24590c8f6284ed4492382ede466c056a33f697b7b3ed2

Pith citing papers

No inbound Pith citation observations are available.