A two-stage contrastive training method aligns song audio with text semantics and then with user-favored song pairs, improving music classification and recommendation over prior models.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.SD 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Bridging the Gap Between Semantic and User Preference Spaces for Multi-modal Music Representation Learning
A two-stage contrastive training method aligns song audio with text semantics and then with user-favored song pairs, improving music classification and recommendation over prior models.