Pith. sign in

arXiv preprint arXiv:2305.04412 , year =

2 Pith papers cite this work. Polarity classification is still indexing.

2 Pith papers citing it

fields

cs.LG 1 cs.RO 1

years

2026 2

verdicts

UNVERDICTED 2

representative citing papers

Intrinsic Vicarious Conditioning for Deep Reinforcement Learning

cs.LG · 2026-05-12 · unverdicted · novelty 7.0

Vicarious conditioning is proposed as a new intrinsic reward in RL that implements attention, retention, reproduction, and reinforcement via memory methods to enable low-shot learning from others without their policies or rewards, yielding longer episodes in tested environments.

citing papers explorer

Showing 2 of 2 citing papers.

  • Intrinsic Vicarious Conditioning for Deep Reinforcement Learning cs.LG · 2026-05-12 · unverdicted · none · ref 59

    Vicarious conditioning is proposed as a new intrinsic reward in RL that implements attention, retention, reproduction, and reinforcement via memory methods to enable low-shot learning from others without their policies or rewards, yielding longer episodes in tested environments.

  • MAGNIFIED: RL Fine-tuning of Multimodal Large Language Models for Motion Planning cs.RO · 2026-06-02 · unverdicted · none · ref 22

    MAGNIFIED applies RL fine-tuning to MLLMs for autonomous driving motion planning, yielding over 10.5% lower overlap rate and 38.9% lower off-road rate than SFT baseline on Waymo Open Motion Dataset.