Pith. sign in

REVIEW 1 cited by

Deep Dive into Model-free Reinforcement Learning for Biological and Robotic Systems: Theory and Practice

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2405.11457 v1 pith:UZ7XQHG4 submitted 2024-05-19 cs.RO cs.AIcs.LG

classification cs.ROcs.AIcs.LG
keywords learningreinforcementdeeproboticanimalbodiescontrolenvironments
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Animals and robots exist in a physical world and must coordinate their bodies to achieve behavioral objectives. With recent developments in deep reinforcement learning, it is now possible for scientists and engineers to obtain sensorimotor strategies (policies) for specific tasks using physically simulated bodies and environments. However, the utility of these methods goes beyond the constraints of a specific task; they offer an exciting framework for understanding the organization of an animal sensorimotor system in connection to its morphology and physical interaction with the environment, as well as for deriving general design rules for sensing and actuation in robotic systems. Algorithms and code implementing both learning agents and environments are increasingly available, but the basic assumptions and choices that go into the formulation of an embodied feedback control problem using deep reinforcement learning may not be immediately apparent. Here, we present a concise exposition of the mathematical and algorithmic aspects of model-free reinforcement learning, specifically through the use of \textit{actor-critic} methods, as a tool for investigating the feedback control underlying animal and robotic behavior.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Optimizing Metachronal Paddling with Reinforcement Learning at Low Reynolds Number

    physics.flu-dyn 2025-07 conditional novelty 6.0 of 10

    In a simulated low-Reynolds-number swimmer with 2 to 4 rigid paddle pairs, reinforcement learning recovers the biologically common back-to-front metachronal wave as the most efficient stroke, while front-to-back or pa...

Pith tools