Pith. sign in

REVIEW 2 cited by

Coprocessor Actor Critic: A Model-Based Reinforcement Learning Approach For Adaptive Brain Stimulation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2406.06714 v2 pith:ARR66C6B submitted 2024-06-10 cs.LG cs.AIcs.HC

classification cs.LGcs.AIcs.HC
keywords learningbrainstimulationcoprocessorapproachneuralreinforcementactor
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Adaptive brain stimulation can treat neurological conditions such as Parkinson's disease and post-stroke motor deficits by influencing abnormal neural activity. Because of patient heterogeneity, each patient requires a unique stimulation policy to achieve optimal neural responses. Model-free reinforcement learning (MFRL) holds promise in learning effective policies for a variety of similar control tasks, but is limited in domains like brain stimulation by a need for numerous costly environment interactions. In this work we introduce Coprocessor Actor Critic, a novel, model-based reinforcement learning (MBRL) approach for learning neural coprocessor policies for brain stimulation. Our key insight is that coprocessor policy learning is a combination of learning how to act optimally in the world and learning how to induce optimal actions in the world through stimulation of an injured brain. We show that our approach overcomes the limitations of traditional MFRL methods in terms of sample efficiency and task success and outperforms baseline MBRL approaches in a neurologically realistic model of an injured brain.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Temporal Basis Function Models for Closed-Loop Neural Stimulation

    cs.LG 2025-07 conditional novelty 6.0 of 10

    Temporal basis function models predict the spatiotemporal LFP response to optogenetic stimulation with test-set R2 around 0.46, beating linear state-space and LSTM baselines while training 30 to 100 times faster.

  2. Neurophysiologically Realistic Environment for Comparing Adaptive Deep Brain Stimulation Algorithms in Parkinson Disease

    q-bio.NC 2025-04 conditional novelty 6.0 of 10

    DBS-Gym is a configurable Kuramoto-based simulation environment that unifies 15 spatial, temporal, and bandwidth features for benchmark testing of adaptive DBS controllers, with RL and classical baselines evaluated at...

Pith tools