Pith. sign in

Is q-learning minimax optimal? a tight sample complexity analysis.Operations Research, 72(1): 222–236, 2024

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

math.OC 1

years

2026 1

verdicts

ACCEPT 1

representative citing papers

Mathematical methods of reinforcement learning

math.OC · 2026-07-08 · accept · novelty 0.0

A survey unifying the operator-theoretic, probabilistic, and optimization-based mathematical structures underlying modern reinforcement learning algorithms.

citing papers explorer

Showing 1 of 1 citing paper.

  • Mathematical methods of reinforcement learning math.OC · 2026-07-08 · accept · none · ref 33

    A survey unifying the operator-theoretic, probabilistic, and optimization-based mathematical structures underlying modern reinforcement learning algorithms.