Pith. sign in

REVIEW

Consciousness-Inspired Spatio-Temporal Abstractions for Better Generalization in Reinforcement Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2310.00229 v4 pith:CS24ZATW submitted 2023-09-30 cs.AI cs.LG

classification cs.AIcs.LG
keywords abstractionsbettergeneralizationlearningplanningreinforcementskipperspatio-temporal
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Inspired by human conscious planning, we propose Skipper, a model-based reinforcement learning framework utilizing spatio-temporal abstractions to generalize better in novel situations. It automatically decomposes the given task into smaller, more manageable subtasks, and thus enables sparse decision-making and focused computation on the relevant parts of the environment. The decomposition relies on the extraction of an abstracted proxy problem represented as a directed graph, in which vertices and edges are learned end-to-end from hindsight. Our theoretical analyses provide performance guarantees under appropriate assumptions and establish where our approach is expected to be helpful. Generalization-focused experiments validate Skipper's significant advantage in zero-shot generalization, compared to some existing state-of-the-art hierarchical planning methods.

Discussion (0). Sign in to comment.

Pith tools