← back to paper
arxiv: 2606.24633 · 2 revisions
Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation