Pith. sign in

Title resolution pending

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

stat.ML 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

Catoni Contextual Bandits are Robust to Heavy-tailed Rewards

stat.ML · 2025-02-04 · conditional · novelty 7.0

Contextual bandits with general function approximation can achieve regret scaling with cumulative reward variance and only logarithmically with the reward range, using Catoni robust mean estimators, with a matching lower bound for the leading term.

citing papers explorer

Showing 1 of 1 citing paper.

  • Catoni Contextual Bandits are Robust to Heavy-tailed Rewards stat.ML · 2025-02-04 · conditional · none · ref 6

    Contextual bandits with general function approximation can achieve regret scaling with cumulative reward variance and only logarithmically with the reward range, using Catoni robust mean estimators, with a matching lower bound for the leading term.