REVIEW 6 cited by
Deep Reinforcement Learning for Robotics: A Survey of Real-World Successes
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Reinforcement learning (RL), particularly its combination with deep neural networks referred to as deep RL (DRL), has shown tremendous promise across a wide range of applications, suggesting its potential for enabling the development of sophisticated robotic behaviors. Robotics problems, however, pose fundamental difficulties for the application of RL, stemming from the complexity and cost of interacting with the physical world. This article provides a modern survey of DRL for robotics, with a particular focus on evaluating the real-world successes achieved with DRL in realizing several key robotic competencies. Our analysis aims to identify the key factors underlying those exciting successes, reveal underexplored areas, and provide an overall characterization of the status of DRL in robotics. We highlight several important avenues for future work, emphasizing the need for stable and sample-efficient real-world RL paradigms, holistic approaches for discovering and integrating various competencies to tackle complex long-horizon, open-world tasks, and principled development and evaluation procedures. This survey is designed to offer insights for both RL practitioners and roboticists toward harnessing RL's power to create generally capable real-world robotic systems.
Forward citations
Cited by 6 Pith papers
-
Fast Estimation of Globally Optimal Independent Contact Regions for Robust Grasping and Manipulation
A Delaunay-triangulation-based divide-and-conquer algorithm computes epsilon-optimal independent contact regions for planar grasps with up to 7 contacts, at speeds suitable for real-time planning.
-
Latent Activation Editing: Inference-Time Refinement of Learned Policies for Safer Multirobot Navigation
Editing a frozen RL policy's latent activations at inference time, using a collision world model, cuts collisions by about 90% on a curated set of hard multirotor scenarios and on real Crazyflies.
-
Is Inter-Seed Cross-Play Enough? Evaluating the Robustness of Zero-Shot Coordination Algorithms to Implementation Details
For Other-Play in Yokai, agents trained with different implementation details coordinate across implementations about as well as across seeds, supporting inter-seed cross-play as a proxy for cross-implementation evaluation.
-
From Motion to Behavior: Hierarchical Modeling of Humanoid Generative Behavior Control
A new 124K-clip dataset with hierarchical text annotations, plus a pipeline that couples an LLM planner, a text-to-pose VAE, diffusion in-betweening, and physics control to generate long-horizon human behaviors.
-
RL as Regressor: A Reinforcement Learning Approach for Function Approximation
Framing regression as an actor-critic reinforcement learning problem with a Gaussian reward can fit a noisy sine wave, but the demonstration reduces to squared-error minimization and adds no new method.
-
Perspective on Utilizing Foundation Models for Laboratory Automation in Materials Research
A perspective article reviews the state of using foundation models for laboratory automation and proposes a roadmap for fully autonomous experiments.
Discussion (0). Sign in to comment.