Pith. sign in

REVIEW 1 cited by

When to Ask for Help: Proactive Interventions in Autonomous Reinforcement Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2210.10765 v1 pith:U7MKIHTA submitted 2022-10-19 cs.LG

classification cs.LG
keywords whenagentsirreversiblestatesalgorithmdesignhelplearning
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

A long-term goal of reinforcement learning is to design agents that can autonomously interact and learn in the world. A critical challenge to such autonomy is the presence of irreversible states which require external assistance to recover from, such as when a robot arm has pushed an object off of a table. While standard agents require constant monitoring to decide when to intervene, we aim to design proactive agents that can request human intervention only when needed. To this end, we propose an algorithm that efficiently learns to detect and avoid states that are irreversible, and proactively asks for help in case the agent does enter them. On a suite of continuous control environments with unknown irreversible states, we find that our algorithm exhibits better sample- and intervention-efficiency compared to existing methods. Our code is publicly available at https://sites.google.com/view/proactive-interventions

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Approximating Safety Feedback Without a Safety Oracle via Model Predictive Control

    cs.LG 2025-10 conditional novelty 5.0 of 10

    RL-SA VMPC shields an RL policy by planning, via MPPI in a black-box simulator, a path from the next state back to the previous state, aborting when no such path exists.

Pith tools