Pith. sign in

Deep reinforcement learning from human preferences.Ad- vances in Neural Information Processing Systems , 30,

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.LG 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

Imitation Learning via Focused Satisficing

cs.LG · 2025-05-20 · conditional · novelty 5.0

MinSubFI directly minimizes subdominance, a margin-based measure of failing to be acceptable, and empirically reports higher demonstrator acceptability than prior imitation methods.

citing papers explorer

Showing 1 of 1 citing paper.

  • Imitation Learning via Focused Satisficing cs.LG · 2025-05-20 · conditional · none · ref 10

    MinSubFI directly minimizes subdominance, a margin-based measure of failing to be acceptable, and empirically reports higher demonstrator acceptability than prior imitation methods.