REVIEW 2 cited by
Solving Time-Continuous Stochastic Optimal Control Problems: Algorithm Design and Convergence Analysis of Actor-Critic Flow
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We propose an actor-critic framework to solve the time-continuous stochastic optimal control problem. A least square temporal difference method is applied to compute the value function for the critic. The policy gradient method is implemented as policy improvement for the actor. Our key contribution lies in establishing a linear rate of convergence for our proposed actor-critic flow. Theoretical findings are further validated through numerical examples, showing the efficacy of our approach in practical applications.
Forward citations
Cited by 2 Pith papers
-
Solving nonconvex Hamilton--Jacobi--Isaacs equations with PINN-based policy iteration
A PINN-based policy iteration method for nonconvex Hamilton-Jacobi-Isaacs equations with a convergence analysis and tests up to 10 dimensions.
-
Simulating Fokker-Planck equations via mean field control of score-based normalizing flows
A mean field control formulation using score-based normalizing flows simulates Fokker-Planck equations deterministically, with a convergence theorem for Ornstein-Uhlenbeck processes and experiments on Langevin and cha...
Discussion (0). Sign in to comment.