A combination of extra audio features, curated training data, and agent-based label correction raises CA-SDRi from 11.088 dB to 12.726 dB on DCASE 2025 Task 4.
Model training The audio-tagging and source-separation models were trained in- dependently
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
eess.AS 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Performance improvement of spatial semantic segmentation with enriched audio features and agent-based error correction for DCASE 2025 Challenge Task 4
A combination of extra audio features, curated training data, and agent-based label correction raises CA-SDRi from 11.088 dB to 12.726 dB on DCASE 2025 Task 4.