Pith. sign in

REVIEW 1 cited by

Contextual Bandit with Herding Effects: Algorithms and Recommendation Applications

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2408.14432 v2 pith:CYQMW7DW submitted 2024-08-26 cs.LG cs.AIcs.IR

classification cs.LGcs.AIcs.IR
keywords effectsfeedbackherdingcontextualrecommendationbanditsbiasts-conf
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Contextual bandits serve as a fundamental algorithmic framework for optimizing recommendation decisions online. Though extensive attention has been paid to tailoring contextual bandits for recommendation applications, the "herding effects" in user feedback have been ignored. These herding effects bias user feedback toward historical ratings, breaking down the assumption of unbiased feedback inherent in contextual bandits. This paper develops a novel variant of the contextual bandit that is tailored to address the feedback bias caused by the herding effects. A user feedback model is formulated to capture this feedback bias. We design the TS-Conf (Thompson Sampling under Conformity) algorithm, which employs posterior sampling to balance the exploration and exploitation tradeoff. We prove an upper bound for the regret of the algorithm, revealing the impact of herding effects on learning speed. Extensive experiments on datasets demonstrate that TS-Conf outperforms four benchmark algorithms. Analysis reveals that TS-Conf effectively mitigates the negative impact of herding effects, resulting in faster learning and improved recommendation accuracy.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Scalable and Interpretable Contextual Bandits: A Literature Review and Retail Offer Prototype

    cs.LG 2025-05 reject novelty 3.0 of 10

    The paper reviews contextual bandit methods and sketches a category-level logistic-regression prototype for retail offers with LLM-generated member profiles, but provides no empirical validation.

Pith tools