REVIEW 2 cited by
Gaussian process policy iteration with additive Schwarz acceleration for forward and inverse HJB and mean field game problems
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
In this paper, we propose a Gaussian Process (GP)-based policy iteration framework for addressing both forward and inverse problems in Hamilton--Jacobi--Bellman (HJB) equations and mean field games (MFGs). Policy iteration is formulated as an alternating procedure between evaluating the value function under a fixed control policy and improving the policy. In our approach, we model the unknown fields using GPs within a policy-iteration framework that converts the nonlinear system into a sequence of linear PDE subproblems. Then, leveraging the linear structure, the updates for the value function and, in the MFG setting, the population density admit explicit representer formulas under linear PDE collocation constraints. The policy is subsequently updated pointwise via a Legendre transform step, which involves a low-dimensional maximization over the control variable. This maximization is explicit for standard quadratic costs. For smooth, strictly convex costs, this pointwise maximization is solved through its first-order optimality condition, whereas in constrained or non-smooth cases, it becomes a low-dimensional constrained maximization problem. To improve convergence, we incorporate the additive Schwarz acceleration as a preconditioning step following each policy update. Numerical experiments demonstrate the effectiveness of the Schwarz acceleration in improving computational efficiency.
Forward citations
Cited by 2 Pith papers
-
Recovering the initial condition and physical coefficients in a nonlinear PDE model of cell invasion
Global uniqueness with Lipschitz stability for two reaction coefficients and logarithmic stability for initial condition in a nonlinear cell invasion PDE, plus a decoupled numerical reconstruction method.
-
A Globally Convergent Flow for Time-Dependent Mean Field Games and a Solver-Agnostic Framework for Inverse Problems
A discretize-then-flow Hessian-Riemannian method globally converges for time-dependent MFGs while preserving positivity and mass, paired with a solver-agnostic bilevel inverse framework using implicit adjoint differen...
Discussion (0). Continue with ORCID to comment.