A tfc-SE attention block inserted after CNN layers of a CRNN reduces sound event detection error rate from 0.2538 to 0.2026 on the synthetic CRESIM overlap-3 dataset.
The tfc-SE block was inserted after each convolution layer of the CRNN model
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
eess.AS 1years
2019 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Sound Event Detection in Multichannel Audio using Convolutional Time-Frequency-Channel Squeeze and Excitation
A tfc-SE attention block inserted after CNN layers of a CRNN reduces sound event detection error rate from 0.2538 to 0.2026 on the synthetic CRESIM overlap-3 dataset.