Crowdsourced MUSHRA tests on Prolific and MTurk reproduce expert codec rankings for generative speech codecs, and SCOREQ tracks subjective quality more consistently than PESQ, POLQA, or ViSQOL.
Conducting these studies requires significant time and cost, often relying on specialized labs or resorting to small- sample internal listening tests with expert listeners
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
eess.AS 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
Crowdsourcing MUSHRA Tests in the Age of Generative Speech Technologies: A Comparative Analysis of Subjective and Objective Testing Methods
Crowdsourced MUSHRA tests on Prolific and MTurk reproduce expert codec rankings for generative speech codecs, and SCOREQ tracks subjective quality more consistently than PESQ, POLQA, or ViSQOL.