AudioSet-R reannotates AudioSet with a three-stage Qwen-Audio, Mistral, and DeepSeek R1 pipeline followed by CLAP filtering, and shows consistent mAP gains on AST, PANNs, SSAST, and AudioMAE.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.SD 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
AudioSet-R: A Refined AudioSet with Multi-Stage LLM Label Reannotation
AudioSet-R reannotates AudioSet with a three-stage Qwen-Audio, Mistral, and DeepSeek R1 pipeline followed by CLAP filtering, and shows consistent mAP gains on AST, PANNs, SSAST, and AudioMAE.