Pith. sign in

Action elimination and stopping conditions for the multi-armed bandit and reinforcement le arning problems

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

citation-role summary

background 1

citation-polarity summary

fields

math.ST 1

years

2025 1

verdicts

REJECT 1

roles

background 1

polarities

support 1

representative citing papers

Early Stopping in Contextual Bandits and Inferences

math.ST · 2025-02-05 · reject · novelty 4.0

Early stopping rules for linear contextual bandits are derived from regret upper bounds and from estimated estimator variances, with a proposed post-stopping conditional inference procedure.

citing papers explorer

Showing 1 of 1 citing paper.

  • Early Stopping in Contextual Bandits and Inferences math.ST · 2025-02-05 · reject · none · ref 7

    Early stopping rules for linear contextual bandits are derived from regret upper bounds and from estimated estimator variances, with a proposed post-stopping conditional inference procedure.