A ResNet-Conformer network with shared weights and multi-scale attention, trained on synthetic and augmented real audio, improves detection and direction estimation over the DCASE 2024 baseline while distance error stays flat.
A four-stage data augmentation approach to resnet-conformer based acoustic modeling for sound event localization and de- tection,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
method 1
citation-polarity summary
fields
cs.SD 1years
2025 1verdicts
CONDITIONAL 1roles
method 1polarities
use method 1representative citing papers
citing papers explorer
-
Resnet-conformer network with shared weights and attention mechanism for sound event localization, detection, and distance estimation
A ResNet-Conformer network with shared weights and multi-scale attention, trained on synthetic and augmented real audio, improves detection and direction estimation over the DCASE 2024 baseline while distance error stays flat.