Pith. sign in

A Minimalist Approach to Offline Reinforcement Learning, December 2021

5 Pith papers cite this work. Polarity classification is still indexing.

5 Pith papers citing it

verdicts

UNVERDICTED 5

representative citing papers

Offline Reinforcement Learning with Implicit Q-Learning

cs.LG · 2021-10-12 · unverdicted · novelty 8.0

IQL achieves policy improvement in offline RL by implicitly estimating optimal action values through state-conditional upper expectiles of value functions, without querying Q-functions on out-of-distribution actions.

Improving Robotic Generalist Policies via Flow Reversal Steering

cs.RO · 2026-06-11 · unverdicted · novelty 7.0

Flow Reversal Steering steers flow matching generalist policies by reversing suboptimal actions to nearby better modes, enabling improved zero-shot control, quick distillation, and RL bootstrapping in robotic manipulation.

citing papers explorer

Showing 5 of 5 citing papers.