A per-instance gated alignment between biased and unbiased ranking towers improved unbiased recommendation accuracy and, in live deployment at Kuaishou, slightly raised retention and interest diversity with negligible latency cost.
TDR-CL: Targeted Doubly Robust Collaborative Learning for Debiased Recommendations
1 Pith paper cite this work, alongside 14 external citations. Polarity classification is still indexing.
abstract
Bias is a common problem inherent in recommender systems, which is entangled with users' preferences and poses a great challenge to unbiased learning. For debiasing tasks, the doubly robust (DR) method and its variants show superior performance due to the double robustness property, that is, DR is unbiased when either imputed errors or learned propensities are accurate. However, our theoretical analysis reveals that DR usually has a large variance. Meanwhile, DR would suffer unexpectedly large bias and poor generalization caused by inaccurate imputed errors and learned propensities, which usually occur in practice. In this paper, we propose a principled approach that can effectively reduce bias and variance simultaneously for existing DR approaches when the error imputation model is misspecified. In addition, we further propose a novel semi-parametric collaborative learning approach that decomposes imputed errors into parametric and nonparametric parts and updates them collaboratively, resulting in more accurate predictions. Both theoretical analysis and experiments demonstrate the superiority of the proposed methods compared with existing debiasing methods.
fields
cs.IR 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
ConAlign: Conditional Alignment Framework for Balancing Biased and Unbiased Recommendation
A per-instance gated alignment between biased and unbiased ranking towers improved unbiased recommendation accuracy and, in live deployment at Kuaishou, slightly raised retention and interest diversity with negligible latency cost.