Pith. sign in

REVIEW 1 cited by

Assessing Quality-Diversity Neuro-Evolution Algorithms Performance in Hard Exploration Problems

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2211.13742 v2 pith:NKAQZTYF submitted 2022-11-24 cs.NE cs.AI

classification cs.NEcs.AI
keywords problemsexplorationalgorithmscontrolmethodsaspectbenchmarkscommunity
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

A fascinating aspect of nature lies in its ability to produce a collection of organisms that are all high-performing in their niche. Quality-Diversity (QD) methods are evolutionary algorithms inspired by this observation, that obtained great results in many applications, from wing design to robot adaptation. Recently, several works demonstrated that these methods could be applied to perform neuro-evolution to solve control problems in large search spaces. In such problems, diversity can be a target in itself. Diversity can also be a way to enhance exploration in tasks exhibiting deceptive reward signals. While the first aspect has been studied in depth in the QD community, the latter remains scarcer in the literature. Exploration is at the heart of several domains trying to solve control problems such as Reinforcement Learning and QD methods are promising candidates to overcome the challenges associated. Therefore, we believe that standardized benchmarks exhibiting control problems in high dimension with exploration difficulties are of interest to the QD community. In this paper, we highlight three candidate benchmarks and explain why they appear relevant for systematic evaluation of QD algorithms. We also provide open-source implementations in Jax allowing practitioners to run fast and numerous experiments on few compute resources.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Scaling Policy Gradient Quality-Diversity with Massive Parallelization via Behavioral Variations

    cs.NE 2025-01 conditional novelty 6.0 of 10

    ASCII-ME replaces actor-critic updates in policy-gradient MAP-Elites with reward-weighted interpolation between action sequences, mapped to policy parameters through a Jacobian, enabling fast GPU-parallel quality-diversity.

Pith tools