REVIEW 3 cited by
StableDR: Stabilized Doubly Robust Learning for Recommendation on Data Missing Not at Random
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
In recommender systems, users always choose the favorite items to rate, which leads to data missing not at random and poses a great challenge for unbiased evaluation and learning of prediction models. Currently, the doubly robust (DR) methods have been widely studied and demonstrate superior performance. However, in this paper, we show that DR methods are unstable and have unbounded bias, variance, and generalization bounds to extremely small propensities. Moreover, the fact that DR relies more on extrapolation will lead to suboptimal performance. To address the above limitations while retaining double robustness, we propose a stabilized doubly robust (StableDR) learning approach with a weaker reliance on extrapolation. Theoretical analysis shows that StableDR has bounded bias, variance, and generalization error bound simultaneously under inaccurate imputed errors and arbitrarily small propensities. In addition, we propose a novel learning approach for StableDR that updates the imputation, propensity, and prediction models cyclically, achieving more stable and accurate predictions. Extensive experiments show that our approaches significantly outperform the existing methods.
Forward citations
Cited by 3 Pith papers
-
Invariant debiasing learning for recommendation via biased imputation
KD-Debias distills a fusion of invariant and variant user preferences into a lightweight matrix-factorization student, improving unbiased recommendation metrics on Yahoo!R3, Coat, and MIND.
-
Doubly Robust Estimation of Causal Effect on CVR with Targeted Regularization
A targeted-regularized, doubly robust estimator for causal effects on post-click conversion rates, with theoretical convergence rates and experiments showing gains over existing CVR causal estimators.
-
EGEAN: An Exposure-Guided Embedding Alignment Network for Post-Click Conversion Estimation
EGEAN combines exposure-guided embedding alignment with a parameter-varying doubly robust loss and reports higher CVR and GMV in offline and online advertising experiments.
Discussion (0). Continue with ORCID to comment.