Pith. sign in

REVIEW 1 cited by

AnyMorph: Learning Transferable Polices By Inferring Agent Morphology

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2206.12279 v1 pith:74HK42V2 submitted 2022-06-17 cs.LG cs.AIcs.RO

AnyMorph: Learning Transferable Polices By Inferring Agent Morphology

classification cs.LG cs.AIcs.RO
keywords morphologyagentlearningagentsdescriptionreinforcementwithoutapproach
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

The prototypical approach to reinforcement learning involves training policies tailored to a particular agent from scratch for every new morphology. Recent work aims to eliminate the re-training of policies by investigating whether a morphology-agnostic policy, trained on a diverse set of agents with similar task objectives, can be transferred to new agents with unseen morphologies without re-training. This is a challenging problem that required previous approaches to use hand-designed descriptions of the new agent's morphology. Instead of hand-designing this description, we propose a data-driven method that learns a representation of morphology directly from the reinforcement learning objective. Ours is the first reinforcement learning algorithm that can train a policy to generalize to new agent morphologies without requiring a description of the agent's morphology in advance. We evaluate our approach on the standard benchmark for agent-agnostic control, and improve over the current state of the art in zero-shot generalization to new agents. Importantly, our method attains good performance without an explicit description of morphology.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Sampling Strategies for Robust Universal Quadrupedal Locomotion Policies

    cs.RO 2025-10 conditional novelty 6.0

    A particle-filter-based adaptive sampling of morphologies and wide PD-gain randomization yields a single quadruped locomotion policy that transfers zero-shot to ANYmal hardware.