Pith. sign in

Paper Citation Record · LEDGER

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification

As of 8 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 1 inbound Pith citation observation for arXiv:2507.21642.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21642 v2

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:36:19.265120Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:36:14.939053Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T12:36:19.393272Z

Reference resolution

46 of 46 outbound references displayed

  • verified exact0
  • verified fuzzy40
  • unresolved5
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0d7fdc80-86a8-45d4-8f83-7ff408980450 · outbound

This paper cites In-the-wild (ITW) refers to data that is not recorded or collected in a highly controlled environ- ment [1].

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification In-the-wild (ITW) refers to data that is not recorded or collected in a highly controlled environ- ment [1]

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.347827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:14.694066Z digest=sha256:7af5bd227eef59eba8643b98f02632cdc808e066f0449a9958e12311b0c4526e

Observation d39270fe-4d4d-49d2-aa0c-7197b77f5b29 · outbound

This paper cites Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification

Reference 2

Resolution
malformed identifier
local_arxiv, observed 2026-08-06T12:36:19.486175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:14.939053Z digest=sha256:2859543338d8995375a184a26a473905707d3f26e6810d154475a87c6a305f22

Observation 3d8a8247-85f7-42fe-9e62-1715c538b78f · outbound

This paper cites For the first training stage, simple filename and label pairs are derived from non-ITW datasets.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification For the first training stage, simple filename and label pairs are derived from non-ITW datasets

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.320435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:15.203573Z digest=sha256:2ef5eb9f48c9e980893e5b7d0a6ab0f1dd7e2abb6d05acc2a57e900fe59fe79f

Observation deb1ed1e-34e7-4258-87eb-c7cb1c86c98d · outbound

This paper cites an unresolved cited work.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:36:20.311411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:15.367337Z digest=sha256:b5758e86b0ba34fc8d0cd97b3ffed7b6a48a9a22cce797878384c578a84c01e2

Observation 18f3e325-9386-4531-9dd6-f7eb3020ba32 · outbound

This paper cites an unresolved cited work.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:36:20.302330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:15.474633Z digest=sha256:6fbb482af8e268508a5bf519ac2ae864cacc97c484c499ccb6cf94e68af34328

Observation b202d257-f806-44c3-a2b5-8e725b0e3ef8 · outbound

This paper cites noisiness.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification noisiness

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.293300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:15.633256Z digest=sha256:505cfd11b9a740c6461fad3e30360c854984502ad02492a39b34f66dfeeff58b

Observation 0eab6f77-b80c-489d-b6d9-13c2aa1b0544 · outbound

This paper cites It was shown that the AITW dataset can be used to train classifiers on multispeaker, foreign language, background music, noise and synthetic speech labels.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification It was shown that the AITW dataset can be used to train classifiers on multispeaker, foreign language, background music, noise and synthetic speech labels

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.284045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:15.759015Z digest=sha256:d022deb5d956c42ddf0fb1740c600319a297a469ef7bf7c0448f568e367495fe

Observation 662c5dbe-1be9-445e-801f-2669df316465 · outbound

This paper cites An open-source speaker gender detection framework for moni- toring gender equality,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification An open-source speaker gender detection framework for moni- toring gender equality,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.213158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:17.012882Z digest=sha256:1ffb6a69754413577d8a923cfbf2c59e8d45a3794a1d1915d8a17f36778eb6bd

Observation f66a5262-6e0e-4108-a8df-557f08781bcd · outbound

This paper cites Label Studio: Data labeling software,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Label Studio: Data labeling software,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.203845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:17.188737Z digest=sha256:d04eac92683eabbe9130ca3e5f209a8d216aec19b72222dd138388e4b4465a36

Observation 80080ab1-ade4-4845-89ee-65c9459317ce · outbound

This paper cites V oxceleb2: Deep speaker recognition,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification V oxceleb2: Deep speaker recognition,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.274372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:15.952302Z digest=sha256:af422085867d57ad68f59afe21f8ff75c1a8febdd01392ea44ae6947f2a9e980

Observation 7f0c58ca-f948-4e55-ad21-a87f4cf104a9 · outbound

This paper cites YODAS: Youtube-Oriented Dataset for Audio and Speech,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification YODAS: Youtube-Oriented Dataset for Audio and Speech,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.263840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:16.062098Z digest=sha256:5776456bd75777ca714a43faf1af7e38ea143628da3066ed69984f6c91bb8116

Observation d4a39cdd-a493-4d96-90f3-f2082407f8dd · outbound

This paper cites Emilia: An Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Speech Generation.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Emilia: An Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Speech Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T12:36:16.250981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:36:16.250981Z digest=sha256:fbb1f5d4986b7c9898a7c22f5b3d091380f99007493cdaacf61bf4297d66a7b1

Observation d42dd31b-d7a2-4b0c-ac1c-20430aa0225f · outbound

This paper cites Speak, read and prompt: High-fidelity text-to-speech with min- imal supervision,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Speak, read and prompt: High-fidelity text-to-speech with min- imal supervision,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.253100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:16.377864Z digest=sha256:b1f8a2154de44c290dafc250db36d190ba12296673aece0f87d538dffa5ae3b0

Observation 4e3daa7b-011d-4649-8449-6688a6aeeeb6 · outbound

This paper cites synthetic speech and non-target languages in many cases).

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification synthetic speech and non-target languages in many cases)

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.339065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:14.761813Z digest=sha256:a47e1f7642d5bec20a0c676b8680c27eb4d4e665919e5518a4816839a1d66c0c

Observation 788c4853-3cc5-42c1-b3fe-f9e889214510 · outbound

This paper cites mhubert-147: A compact multilingual hubert model,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification mhubert-147: A compact multilingual hubert model,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.242353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:16.582222Z digest=sha256:7c1a0f69ce097383e369d575e784caa3cc184cf93c9d1db8104c7b852c22fae1

Observation 3a99ca04-b170-4b91-9a1c-1a49d79b0ac9 · outbound

This paper cites Sentence level intelligibility eval- uation for mandarin text-to-speech systems using semantically unpredictable sentences,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Sentence level intelligibility eval- uation for mandarin text-to-speech systems using semantically unpredictable sentences,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.232020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:16.739548Z digest=sha256:ea9c7b17c057d5fc6874d10548e7439320dc79ed6c7e78190cc2ce7795d501b6

Observation 8095009c-91eb-4157-b76f-2d882ec09be2 · outbound

This paper cites Data-filtering methods for self-training of auto- matic speech recognition systems,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Data-filtering methods for self-training of auto- matic speech recognition systems,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.222623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:16.856553Z digest=sha256:79fe9a21bcfb29374eb57c1ab391cc547b5100470b337edb37f3022cd17dcce0

Observation dfd21f4b-5d2b-4ef0-ac35-e1014274d4fb · outbound

This paper cites Attention is All you Need,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Attention is All you Need,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.026366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:17.993141Z digest=sha256:ad5dcf47650e06eefacb229004d4b47156d3cb216e93560d19947436462df949

Observation b9c5a3b8-08c8-40f9-9502-6c92053cb87b · outbound

This paper cites pyannote.audio 2.1 speaker diarization pipeline: prin- ciple, benchmark, and recipe,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification pyannote.audio 2.1 speaker diarization pipeline: prin- ciple, benchmark, and recipe,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.194492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:17.257923Z digest=sha256:8228fba49c0b6aaaab712acd8ec0c4525469fe4d18fffb1c90bef77ab2a96beb

Observation 6c8d09a5-d94e-415b-ae78-4a814d3773de · outbound

This paper cites FoR: A dataset for synthetic speech detection,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification FoR: A dataset for synthetic speech detection,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.092229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:17.347739Z digest=sha256:83687aad50b3e3e3f8655868f3c8ad3a07243917d892374d271e271dc75a6d4e

Observation 9a699321-9cf0-4061-93d0-a07c96283c51 · outbound

This paper cites AASIST: Audio Anti-Spoofing Us- ing Integrated Spectro-Temporal Graph Attention Networks,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification AASIST: Audio Anti-Spoofing Us- ing Integrated Spectro-Temporal Graph Attention Networks,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.083301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:17.436804Z digest=sha256:081608b2c3743ba1fd9f94d27711ab053a8d4e7246e2c5dcf278a38d47269020

Observation d2e535ba-d3a9-4939-8abf-99a1931349bd · outbound

This paper cites This as- sumption is validated later in our results, cf.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification This as- sumption is validated later in our results, cf

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.329846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:15.048489Z digest=sha256:e767da5be4614f6e370795147b42e06ebeda50c86cafccad4f756c42c4c90fa6

Observation 431ed380-c6a3-4748-9ba2-f678217b81d3 · outbound

This paper cites Robust speech recognition via large-scale weak su- pervision,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Robust speech recognition via large-scale weak su- pervision,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.073766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:17.518641Z digest=sha256:13c895d736ff1204d873c5c7da8fb94e2ead519d134f47342eccff67a7619166

Observation 7a73256a-16f2-4004-b601-e7a39733765d · outbound

This paper cites Perceive and predict: Self-supervised speech representation based loss func- tions for speech enhancement,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Perceive and predict: Self-supervised speech representation based loss func- tions for speech enhancement,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.064610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:17.608036Z digest=sha256:13132ecef10c27de6dbe3ac8e65c824712e26e5118dd2214dc8b986b8980b96f

Observation 1d4fe041-0f67-4bbe-9392-c30dc86605ce · outbound

This paper cites Transcription-free fine-tuning of speech separation models for noisy and reverberant multi-speaker automatic speech recognition,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Transcription-free fine-tuning of speech separation models for noisy and reverberant multi-speaker automatic speech recognition,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.055009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:17.687623Z digest=sha256:3cf4a0dd6d77ca377376828c37f66e8269b5fa90b7a90030090223a7cb63d620

Observation 4bd74790-bffe-434a-8f27-88d484b90167 · outbound

This paper cites Non-intrusive speech intelligibility pre- diction for hearing-impaired users using intermediate asr features and human memory models,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Non-intrusive speech intelligibility pre- diction for hearing-impaired users using intermediate asr features and human memory models,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.045407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:17.818034Z digest=sha256:4a9198969ede89a9a4b0fbee4feeb4d75ba6d70d1c594a364f2217b99ec80200

Observation 9a794ed0-4674-4f02-afdc-075e8d938a42 · outbound

This paper cites Heterogeneity over homogeneity: Investigating multilingual speech pre-trained models for detecting audio deepfake,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Heterogeneity over homogeneity: Investigating multilingual speech pre-trained models for detecting audio deepfake,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.036295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:17.904511Z digest=sha256:c47a67dc01a8b03cef39eadb620481dc03b1c1172591d70f517d5b0cf15a9663

Observation f78d6fe7-9e8c-4e81-9b18-d2446077e1d7 · outbound

This paper cites Nisqa: A deep cnn-self-attention model for multidimensional speech quality pre- diction with crowdsourced datasets,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Nisqa: A deep cnn-self-attention model for multidimensional speech quality pre- diction with crowdsourced datasets,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.016269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:18.073975Z digest=sha256:a3c5ddda58a38b5e0867d299e81591018b828985db0c83f5c53d08dbc803a6e6

Observation 19dd9b2a-49d7-4688-a944-9e8599a75257 · outbound

This paper cites Hallucination in perceptual metric-driven speech enhancement networks,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Hallucination in perceptual metric-driven speech enhancement networks,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:20.005475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:18.121425Z digest=sha256:671c1a99d4dffbb3333c2276e354f51059a4833ed66a895562b52cd896755b8d

Observation e2518e7a-a00f-478f-abe5-525e202aa34b · outbound

This paper cites An overview of multi-task learning in deep neural networks,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification An overview of multi-task learning in deep neural networks,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.994398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:18.183181Z digest=sha256:1f14cebb5e2b38c9fbed193c9ba11ef985b5c5e50c09a43e38a774fd96661b9f

Observation 05818646-6738-4a57-9ddd-199d562323d4 · outbound

This paper cites BEATs: Audio pre-training with acoustic to- kenizers,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification BEATs: Audio pre-training with acoustic to- kenizers,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.985310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:18.246729Z digest=sha256:ce142e36e029e05a2d6df355164f54a2e20a9f4ee9500a64d9b3a34801df8543

Observation aebc4646-88f0-4a16-b3e8-57bcbf75fec7 · outbound

This paper cites Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.976010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:18.331090Z digest=sha256:3902ded3e073d555f5280229da28f85143e1d13206ecab93039a91a643ce0bc6

Observation 014a8bb2-4168-48c3-a442-168f1bee4121 · outbound

This paper cites Wavesplit: End-to-end speech separation by speaker clustering,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Wavesplit: End-to-end speech separation by speaker clustering,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.966245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:18.453605Z digest=sha256:605ed167b61beb0beac373f6582b3fc91d31f31b9bfce29290779a99aa378b51

Observation fb1510eb-4c64-47b5-aada-3fdee4d7de85 · outbound

This paper cites The AMI Meeting Corpus: A Pre- announcement,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification The AMI Meeting Corpus: A Pre- announcement,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T12:36:18.476532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:36:18.476532Z digest=sha256:45f86b4faa9f56c7a536e64377c0e36cf8a4968fa039da274aafd08e2b45f85d

Observation e2ccdfa0-b169-4d5b-b4b3-0a34d3d08f00 · outbound

This paper cites M2Met: The ICASSP 2022 Multi-Channel Multi-Party Meeting Transcription Challenge,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification M2Met: The ICASSP 2022 Multi-Channel Multi-Party Meeting Transcription Challenge,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.950665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:18.591304Z digest=sha256:18c89013a71fc9b2b90abe06d55bc31e8b7930c72ea0f514aa1802f6bbd50d83

Observation 5a0b28c5-50c1-44f0-bc0d-ddb1dd0f03e8 · outbound

This paper cites MLS: A Large-Scale Multilingual Dataset for Speech Research,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification MLS: A Large-Scale Multilingual Dataset for Speech Research,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.940318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:18.612570Z digest=sha256:b015d45674878a6c96a2ae515c3e43a43910458a83593f30bf0a416d20fb6223

Observation 79768340-eee1-4774-b730-8ef1af09d142 · outbound

This paper cites MUSAN: A Music, Speech, and Noise Corpus.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification MUSAN: A Music, Speech, and Noise Corpus

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T12:36:18.698419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:36:18.698419Z digest=sha256:e77e67826b59b0607924a8fb1a8e918cc460083c9b49ee1f955c42f99bc691bf

Observation 153e9b18-4ea5-417b-b93c-d74ff886df1b · outbound

This paper cites An open dataset of synthetic speech,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification An open dataset of synthetic speech,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.930422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:18.763442Z digest=sha256:bfa3ea011b4137a1c07659c8d91d2d67312ee4cfebd92e2d9c29c413c5d21bba

Observation 3d43015b-4aec-426c-a567-8d67381e9d58 · outbound

This paper cites OpenMIC-2018: An Open Dataset for Multiple Instrument Recognition,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification OpenMIC-2018: An Open Dataset for Multiple Instrument Recognition,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.920880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:18.817911Z digest=sha256:57846b16ea74cded6b663fa670642c25676f25b3bdf36b0de4fcd1f187fa5938

Observation 94bc4d53-bd70-44fb-8698-49640fe024cb · outbound

This paper cites Learning sound event classifiers from web audio with noisy labels,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Learning sound event classifiers from web audio with noisy labels,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.910904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:18.927422Z digest=sha256:b416c2a7a9fff5ca49c1f2b4617871c58cd5c59fbff68b29c076776169a93b77

Observation 071634a6-5ab4-47b2-82e3-4d333f345586 · outbound

This paper cites The Diverse Environments Multi-channel Acoustic Noise Database (DEMAND): A database of multichannel environmental noise recordings,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification The Diverse Environments Multi-channel Acoustic Noise Database (DEMAND): A database of multichannel environmental noise recordings,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.900996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:18.964872Z digest=sha256:6ef6b7940a8a0f985878225d9572e4c95ef81029f57aae00f6df19ff31f5f067

Observation 1bd364b6-7016-4a1a-b27a-88cae80a4116 · outbound

This paper cites Adam: A method for stochastic opti- mization,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Adam: A method for stochastic opti- mization,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.890136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:19.040456Z digest=sha256:68b9a9adf553d22e185444164298454bd9b8d7c06543330182e8617c91fa718e

Observation a5aa999c-6a45-46ea-9f94-54606eccca78 · outbound

This paper cites Data augmen- tation for speech separation,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Data augmen- tation for speech separation,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.879696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:19.120411Z digest=sha256:a1f2fb23c72b3f300d8b101ef650e1ca39eef74e761addf501461b421531c910

Observation e52c97ed-3e81-4024-aec6-8138acbdf9b0 · outbound

This paper cites Dnsmos: A non- intrusive perceptual objective speech quality metric to evaluate noise suppressors,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Dnsmos: A non- intrusive perceptual objective speech quality metric to evaluate noise suppressors,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.870112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:19.139425Z digest=sha256:655dded8d6605c9fc71f111bf27f54057baea83f2875d593e75f2ee55aee6c26

Observation 9b35680e-7f3d-4ca2-95a1-ec19dc7d79f5 · outbound

This paper cites Convolutional recurrent neural networks for poly- phonic sound event detection,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Convolutional recurrent neural networks for poly- phonic sound event detection,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.834641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:19.156870Z digest=sha256:69e99a59ec1ed1cba6f35b34e53e5b7f0a6daac7b932e3736fc8e70fe813e05f

Observation 2a14e91f-c44d-491f-ab8d-9d88d3921049 · outbound

This paper cites Mlaad: The multi- language audio anti-spoofing dataset,.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Mlaad: The multi- language audio anti-spoofing dataset,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:36:19.652398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:19.265120Z digest=sha256:bbd9e53f5181056a507ab53f650353deaf280d1bfca0857efb1ee5b3286ca46f

Pith citing papers

Observation d39270fe-4d4d-49d2-aa0c-7197b77f5b29 · inbound

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification cites this paper.

Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification Whilter: A Whisper-based Data Filter for "In-the-Wild" Speech Corpora Using Utterance-level Multi-Task Classification

Reference 2

Resolution
malformed identifier
local_arxiv, observed 2026-08-06T12:36:19.486175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T12:36:14.939053Z digest=sha256:2859543338d8995375a184a26a473905707d3f26e6810d154475a87c6a305f22