← back to paper
arxiv: 2604.18530 · 2 revisions
OGER: A Robust Offline-Guided Exploration Reward for Hybrid Reinforcement Learning