← back to paper
arxiv: 2607.17823 · 2 revisions
Theoretical Foundations of $\max$@$k$ Reinforcement Learning