RecLM uses two-turn collaborative instruction tuning plus a reinforcement-learning reward model to generate user and item profiles that improve cold-start recommendation performance when plugged into existing recommenders.
Recommendation as language processing (rlp): A unified pretrain, personalized prompt & predict paradigm (p5)
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.IR 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
RecLM: Recommendation Instruction Tuning
RecLM uses two-turn collaborative instruction tuning plus a reinforcement-learning reward model to generate user and item profiles that improve cold-start recommendation performance when plugged into existing recommenders.