REVIEW 4 cited by
End-to-end Training for Recommendation with Language-based User Profiles
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
There is a growing interest in natural language-based user profiles for recommender systems, which aims to enhance transparency and scrutability compared with embedding-based methods. Existing studies primarily generate these profiles using zero-shot inference from large language models (LLMs), but their quality remains insufficient, leading to suboptimal recommendation performance. In this paper, we introduce LangPTune, the first end-to-end training framework to optimize LLM-generated user profiles. Our method significantly outperforms zero-shot approaches by explicitly training the LLM for the recommendation objective. Through extensive evaluations across diverse training configurations and benchmarks, we demonstrate that LangPTune not only surpasses zero-shot baselines but can also matches the performance of state-of-the-art embedding-based methods. Finally, we investigate whether the training procedure preserves the interpretability of these profiles compared to zero-shot inference through both GPT-4 simulations and crowdworker user studies. Implementation of LangPTune can be found at https://github.com/ZhaolinGao/LangPTune.
Forward citations
Cited by 4 Pith papers
-
Biases in LLM-Generated Musical Taste Profiles for Recommendation
Users identify more with LLM-generated music taste profiles for some genres and user groups than others, and these biases differ across models.
-
RECAP: Feedback-Driven Streaming Semantic User Profiles for Short-Video Recommendation
RECAP trains a streaming LLM profile updater with GRPO rewards from a dual-tower evaluator, gaining +0.0084 uAUC (cleaned eval) and +0.139% online usage time.
-
Prediction Is Not Memory: Dual-Timescale Gated Profile Writing for Persistent User Modeling
A lightweight write-risk gate reduces harmful persistent-profile updates from 22.45% to about 14.5% on MicroLens-100K, and next-item ranking confidence is a poor substitute for write-risk scoring.
-
Music Recommendation with Large Language Models: Challenges, Opportunities, and Evaluation
A review and position paper proposing a six-dimension success framework and risk diagnostics for evaluating LLM-based music recommendation systems.
Discussion (0). Sign in to comment.