X-CrossNet applies the CrossNet separation backbone to target speaker extraction with cross-attention speaker embedding fusion, reporting small improvements on WSJ0-2mix and WHAMR!.
X-tf-gridnet: A time–frequency domain target speaker extraction network with adaptive speaker embedding fusion,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.SD 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
X-CrossNet: A complex spectral mapping approach to target speaker extraction with cross attention speaker embedding fusion
X-CrossNet applies the CrossNet separation backbone to target speaker extraction with cross-attention speaker embedding fusion, reporting small improvements on WSJ0-2mix and WHAMR!.