A rate-distortion trained speech codec using a channel-wise entropy model and CNN-RWKV blocks reports 53.51% average BD-rate savings over four baselines.
High- fidelity audio compression with improved rvqgan,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
other 1
citation-polarity summary
fields
eess.AS 1years
2025 1verdicts
CONDITIONAL 1roles
other 1polarities
unclear 1representative citing papers
citing papers explorer
-
Rate-Aware Learned Speech Compression
A rate-distortion trained speech codec using a channel-wise entropy model and CNN-RWKV blocks reports 53.51% average BD-rate savings over four baselines.