Q-learning firms converge to a supracompetitive price forever if the Q-function at the end of experimentation favors that price in the relevant states.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
econ.GN 1years
2025 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
Learning to Charge More: A Theoretical Study of Collusion by Q-Learning Agents
Q-learning firms converge to a supracompetitive price forever if the Q-function at the end of experimentation favors that price in the relevant states.