Coarse Q-learning aggregates feedback within exogenous similarity classes in stochastic bandit problems, yielding mean-field dynamics whose high payoff-sensitivity limits include multiple stable strict equilibria, a unique globally stable mixed indifference equilibrium, or convergence to a stable限周期
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
econ.TH 1years
2024 1verdicts
UNVERDICTED 1representative citing papers
citing papers explorer
-
Coarse Q-learning: Indifference, Indeterminacy, and Instability
Coarse Q-learning aggregates feedback within exogenous similarity classes in stochastic bandit problems, yielding mean-field dynamics whose high payoff-sensitivity limits include multiple stable strict equilibria, a unique globally stable mixed indifference equilibrium, or convergence to a stable限周期