A late-fusion model that combines ASR-transcribed lyrics and speech embeddings detects AI-written lyrics from audio alone, achieving 94.9% recall in-domain and staying robust to attacks.
Title resolution pending
1 Pith paper cite this work, alongside 19 external citations. Polarity classification is still indexing.
1
Pith paper citing it
19
external citations · OpenAlex
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Double Entendre: Robust Audio-Based AI-Generated Lyrics Detection via Multi-View Fusion
A late-fusion model that combines ASR-transcribed lyrics and speech embeddings detects AI-written lyrics from audio alone, achieving 94.9% recall in-domain and staying robust to attacks.