Pith. sign in

REVIEW 3 cited by

Transform2Act: Learning a Transform-and-Control Policy for Efficient Agent Design

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2110.03659 v3 pith:OX5C6TJD submitted 2021-10-07 cs.LG cs.AIcs.CVcs.NEcs.RO

classification cs.LGcs.AIcs.CVcs.NEcs.RO
keywords designagentjointpolicytransform2actactionsdesignsacross
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

An agent's functionality is largely determined by its design, i.e., skeletal structure and joint attributes (e.g., length, size, strength). However, finding the optimal agent design for a given function is extremely challenging since the problem is inherently combinatorial and the design space is prohibitively large. Additionally, it can be costly to evaluate each candidate design which requires solving for its optimal controller. To tackle these problems, our key idea is to incorporate the design procedure of an agent into its decision-making process. Specifically, we learn a conditional policy that, in an episode, first applies a sequence of transform actions to modify an agent's skeletal structure and joint attributes, and then applies control actions under the new design. To handle a variable number of joints across designs, we use a graph-based policy where each graph node represents a joint and uses message passing with its neighbors to output joint-specific actions. Using policy gradient methods, our approach enables joint optimization of agent design and control as well as experience sharing across different designs, which improves sample efficiency substantially. Experiments show that our approach, Transform2Act, outperforms prior methods significantly in terms of convergence speed and final performance. Notably, Transform2Act can automatically discover plausible designs similar to giraffes, squids, and spiders. Code and videos are available at https://sites.google.com/view/transform2act.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Transformer Transformer: A Unified Model for Motion-Conditioned Robot Co-design

    cs.RO 2026-07 conditional novelty 7.0 of 10

    A single diffusion transformer trains on tokenized robot bodies and motions to generate and optimize robot designs for unseen rewards and trajectories, outpacing evolutionary search in speed and often in reward.

  2. House of Dextra: Cross-embodied Co-design for Dexterous Hands

    cs.RO 2025-12 unverdicted novelty 6.0 of 10

    A cross-embodied co-design framework learns task-specific hand morphologies and control policies, achieving sim-to-real rotation up to 3.3 rad/s and full fabrication in under 24 hours.

  3. VLMgineer: Vision Language Models as Robotic Toolsmiths

    cs.RO 2025-07 conditional novelty 6.0 of 10

    VLMgineer combines VLM-generated URDF tool designs with evolutionary search to co-design tools and action plans, outperforming human-specified and existing tools on a new simulated manipulation benchmark.

Pith tools