Pith. sign in

REVIEW 2 cited by

Curiosity Driven Exploration of Learned Disentangled Goal Spaces

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1807.01521 v3 pith:VQQLYWTY submitted 2018-07-04 cs.LG cs.AIcs.NEcs.ROstat.ML

classification cs.LGcs.AIcs.NEcs.ROstat.ML
keywords goalexplorationspacedisentangledenvironmentslearningalgorithmscomplex
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Intrinsically motivated goal exploration processes enable agents to autonomously sample goals to explore efficiently complex environments with high-dimensional continuous actions. They have been applied successfully to real world robots to discover repertoires of policies producing a wide diversity of effects. Often these algorithms relied on engineered goal spaces but it was recently shown that one can use deep representation learning algorithms to learn an adequate goal space in simple environments. However, in the case of more complex environments containing multiple objects or distractors, an efficient exploration requires that the structure of the goal space reflects the one of the environment. In this paper we show that using a disentangled goal space leads to better exploration performances than an entangled goal space. We further show that when the representation is disentangled, one can leverage it by sampling goals that maximize learning progress in a modular manner. Finally, we show that the measure of learning progress, used to drive curiosity-driven exploration, can be used simultaneously to discover abstract independently controllable features of the environment.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Learning to Explore in Motion and Interaction Tasks

    cs.RO 2019-08 conditional novelty 6.0 of 10

    A learned generative model of past task motions, used as exploration noise in DDPG, speeds up learning of new robot manipulation and contact tasks by more than two times in simulation.

  2. A survey on intrinsic motivation in reinforcement learning

    cs.LG 2019-08 accept novelty 4.0 of 10

    A survey that classifies intrinsic motivation methods in deep RL as knowledge acquisition or skill learning and proposes their unification through information compression.

Pith tools