Pith. sign in

REVIEW 2 cited by

Gaussian process policy iteration with additive Schwarz acceleration for forward and inverse HJB and mean field game problems

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2505.00909 v3 pith:KQUANU6Z submitted 2025-05-01 cs.LG math.OC

classification cs.LGmath.OC
keywords policymaximizationaccelerationiterationlinearschwarzadditiveconstrained
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In this paper, we propose a Gaussian Process (GP)-based policy iteration framework for addressing both forward and inverse problems in Hamilton--Jacobi--Bellman (HJB) equations and mean field games (MFGs). Policy iteration is formulated as an alternating procedure between evaluating the value function under a fixed control policy and improving the policy. In our approach, we model the unknown fields using GPs within a policy-iteration framework that converts the nonlinear system into a sequence of linear PDE subproblems. Then, leveraging the linear structure, the updates for the value function and, in the MFG setting, the population density admit explicit representer formulas under linear PDE collocation constraints. The policy is subsequently updated pointwise via a Legendre transform step, which involves a low-dimensional maximization over the control variable. This maximization is explicit for standard quadratic costs. For smooth, strictly convex costs, this pointwise maximization is solved through its first-order optimality condition, whereas in constrained or non-smooth cases, it becomes a low-dimensional constrained maximization problem. To improve convergence, we incorporate the additive Schwarz acceleration as a preconditioning step following each policy update. Numerical experiments demonstrate the effectiveness of the Schwarz acceleration in improving computational efficiency.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Recovering the initial condition and physical coefficients in a nonlinear PDE model of cell invasion

    math.AP 2026-06 unverdicted novelty 7.0 of 10

    Global uniqueness with Lipschitz stability for two reaction coefficients and logarithmic stability for initial condition in a nonlinear cell invasion PDE, plus a decoupled numerical reconstruction method.

  2. A Globally Convergent Flow for Time-Dependent Mean Field Games and a Solver-Agnostic Framework for Inverse Problems

    math.OC 2026-03 conditional novelty 6.0 of 10

    A discretize-then-flow Hessian-Riemannian method globally converges for time-dependent MFGs while preserving positivity and mass, paired with a solver-agnostic bilevel inverse framework using implicit adjoint differen...

Pith tools