REVIEW 3 major objections 4 minor 30 references
For risk-sensitive exit-time control problems with path-dependent coefficients, this paper proves that the log-transformed value converges to the value of a deterministic control problem as the noise intensity vanishes.
Reviewed by Pith at T0; open to challenge. T0 means a machine referee read the full paper against a public rubric. the ladder, T0–T4 →
T0 review · deepseek-v4-flash
2026-08-01 15:43 UTC pith:6K5WLSPV
load-bearing objection Real path-dependent extension of Boué–Dupuis with a genuine gap in the Lipschitz reduction; deserves refereeing, but the key lemma needs to be either fixed or deferred to a published [9]. the 3 major comments →
Risk-sensitive exit-time control for stochastic differential equations with path-dependent coefficients
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
Core claim
On the paper's own terms, the discovery is Theorem 2.4: for every initial condition (t,ω) belonging to a set T of regular initial conditions, every sequence of initial conditions converging to (t,ω), and every noise level ε_n→0, the logarithmically scaled risk-sensitive exit-time value V^{ε_n}_{t_n,ω_n} converges to V^0_{t,ω}, the value of a deterministic path-dependent control problem. The deterministic limiting problem has two ingredients not visible at finite noise: an added drift control drawn from the Cameron–Martin space and a quadratic penalty for its energy. The regularity set T excludes boundary behaviour where the δ-blow-up value V^{0,δ} does not converge to V^0, and the paper give
What carries the argument
The load-bearing object is the variational representation E^*_t(φ)=E'_t(φ) (Theorem 3.1) for the log-transform of a path-dependent stochastic control problem. It expresses the entropically transformed value as a supremal expectation over relaxed control rules carrying an extra drift control z and a quadratic penalty ∥z∥²/2. The proof identifies both sides as viscosity solutions of the same path-dependent Hamilton–Jacobi–Bellman equation—one with the maximization Hamiltonian G and one with its Legendre-transformed counterpart G̃—and uses a comparison principle for convex expectations on path space to reduce the class of admissible payoffs from upper semianalytic to Lipschitz functions.
Load-bearing premise
The proof chain assumes that the comparison theorem for convex expectations on path space, proved by the authors for a Markovian setting, remains valid in the present path-dependent relaxed-control framework; Lemma 3.2 invokes it without a self-contained proof, and the reduction from general payoffs to Lipschitz payoffs—a step Theorem 3.1 cannot do without—collapses if that comparison fails.
What would settle it
Exhibit a path-dependent SDE satisfying Assumption 2.1 and a bounded upper semianalytic payoff φ for which the claimed variational equality (3.1) fails, or find an initial condition (t,ω)∈T and sequences (t_n,ω_n), ε_n→0 for which V^{ε_n}_{t_n,ω_n} does not converge to V^0_{t,ω}. More narrowly, checking whether the comparison theorem from the authors' preprint holds for relaxed control rules with unbounded z would settle the load-bearing step.
If this is right
- Small-noise asymptotics for risk-sensitive exit problems are now available when drift and diffusion depend on the entire past, not just the current state.
- The limiting deterministic control problem is explicit, so optimal limiting strategies can be computed: a path-dependent ODE with an added Cameron–Martin drift and a quadratic control cost.
- The variational formula Theorem 3.1 applies to measurable functions of the path beyond the exit-time reward g(τ_D), making it a standalone tool for entropic transforms.
- The memory example yields a fully computed limit and explicit optimal control, demonstrating that the abstract condition (t,ω)∈T holds in a concrete non-Markovian model.
- Because convergence holds along sequences of initial conditions, the result has a stability property: small perturbations of initial data do not break the limit.
Where Pith is reading between the lines
- If the comparison principle invoked in Lemma 3.2 extends as the authors expect, the same variational formula should hold for bounded upper semianalytic payoffs far beyond exit times, including quantile-like and constraint-type criteria in path-dependent control.
- The condition (t,ω)∈T is the real substance of the theorem: the paper's boundary analysis suggests that, in Markovian models, T is essentially the set of starting points from which the boundary can be left with a finite-energy control; testing this characterization in fully path-dependent models is a natural next step.
- The Hausdorff convergence of relaxed control rules in Corollary 4.3 is proved under quadratic moment bounds; one could expect analogous convergence under p-th moment bounds, which would extend the method to rewards with polynomial growth.
- The memory example indicates that the limiting control problem sometimes reduces to a finite-dimensional Hilbert-space projection; similar reductions could be systematically derived for affine or linearized path-dependent dynamics.
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The paper studies the small-noise asymptotics of a risk-sensitive exit-time control problem for SDEs with path-dependent coefficients. The main result, Theorem 2.4, asserts that for initial conditions in a set T, the log-transformed value ε log sup E[exp(g(τ_D)/ε)] converges to a deterministic path-dependent control value V^0, defined through Cameron–Martin controls h and relaxed controls m. The proof proceeds by first establishing a variational representation (Theorem 3.1) using path-dependent PDE techniques and convex expectations on path space, then proving a relaxed-control convergence result (Corollary 4.3), and finally combining these with semicontinuity arguments in Section 5. The paper also contains a worked one-dimensional example with memory.
Significance. If correct, the paper would provide a path-dependent extension of the Boué–Dupuis variational approach and of classical small-noise limits for risk-sensitive escape control, with an explicit computable example. The combination of PPDE comparison principles, convex expectations, and relaxed-control convergence is original and the overall strategy is well motivated. The paper is clearly written and the semicontinuity manipulations in Section 5 are transparent. However, the main theorem rests on two load-bearing points that are not fully justified: the reduction to Lipschitz payoffs is deferred to an unpublished comparison result, and the deterministic limit V^0 defined in Section 2 does not obviously coincide with the relaxed-control limit actually obtained in Section 5. These issues affect the central claim and require attention.
major comments (3)
- [Section 3.1, Lemma 3.2] After applying Corollary 4.3, the limsup is bounded by sup_{P∈R^0_M(t,ω)} E_P[g(τ^t_{Dδ}) - cost]. The next inequality '≤ V^{0,δ}_{t,ω}' is asserted without argument. Here V^{0,δ} is defined in Section 2 as a supremum over (h,m) with product structure m(ds,dλ)⊗δ_{h'(s)}(dz), whereas R^0 consists of relaxed controls with arbitrary joint measures M*(ds,dλ,dz). For a general M*, the marginal m on Λ and the averaged h'(s)=∫ z M*_s(dλ,dz) do not reproduce the drift ∫σ(s,X,λ) z M*; the equality would require, e.g., σ independent of λ or z independent of λ conditionally. The product controls are a strict subset of R^0, and the supremum over R^0 can be strictly larger. No convexity or density argument is provided to justify the reduction. This makes the upper bound (5.1) unproved and Theorem 2.4, as stated, unsupported.
- [Section 3.2, Lemma 3.6(b)] The proof of Lemma 3.2 is one sentence: it asserts that E* and E' are convex expectations as in [9] and then invokes the comparison theorem [9, Theorem 2.12]. This is load-bearing because Theorem 3.1 is later applied in Section 5 to exit-time payoffs g(τ_D^{t_n}), which are not Lipschitz. The present paper does not verify that E', defined via relaxed control rules with unbounded control z and quadratic cost on the infinite-dimensional path space C(R_+;R^d), satisfies the axioms of [9, Definition 2.1] or the hypotheses of the comparison theorem. The dependence on an unpublished preprint makes this gap more serious. Please either include a self-contained verification of the convex-expectation axioms and the comparison, or provide a direct proof of the reduction for the specific structure of E'.
- [Section 4, Proposition 4.2(a)] Lemma 3.6(b) proves only the subsolution property and explicitly leaves the supersolution property and the d-Lipschitz continuity to the reader. These properties are necessary for Lemma 3.9, where v=e^{tilde v} is shown to solve the G-backward equation, and for the uniqueness argument. The omitted parts are not routine in the presence of an unbounded action variable z and the relaxed-control topology. Similarly, Proposition 4.2(a) omits the martingale-problem argument and the compactness proof that underpin Corollary 4.3. Since these results are essential for both the variational formula and the convergence argument, the full proofs or precise statements of how they follow from the cited references should be included.
minor comments (4)
- [Discussion 2.5] The reduction to λ≡1 is justified heuristically ('Intuitively, ...', 'Approximating controls in A by piecewise constant controls, we find ...'). Since this is a claimed explicit characterization of the limit, the argument should either be made rigorous or explicitly labeled as heuristic.
- [General] The proof that the stated geometric conditions imply (t,ω)∈T is informal; some estimates gloss over measurability and the treatment of the case τ_D=∞ is abbreviated. Please clarify these points.
- [Appendix A] The notation V^0 vs V^{0,δ} and the distinction between the deterministic control problem in Section 2 and the relaxed-control value in Sections 4–5 is confusing; a remark explaining the intended equivalence would help.
- [Section 3.1, proof of Lemma 3.2] The definition of C^{1,2}_{pol} is delegated to the appendix, but the appendix defines derivatives only on the space of càdlàg paths. The passage between continuous and càdlàg paths should be stated more carefully.
Circularity Check
No definitional circularity: the limit V^0 is an independent deterministic control problem and no parameter is fitted; the main concern is load-bearing reliance on the authors' own prior results [9]–[11], which is a verification/correctness issue, not circular reasoning.
full rationale
The paper's derivation chain is not circular. The stochastic value V^ε is defined via ε log of a supremal exponential reward, while the claimed limit V^0 is an independent deterministic control problem: sup over Cameron–Martin controls h and relaxed controls m of g(τ_D(Y^{h,m})) minus half the L^2 cost of h. The limiting object is not defined in terms of V^ε, and no constants are fitted to data. The proof of Theorem 2.4 combines a variational representation (Theorem 3.1) with an independent convergence argument for relaxed control rules (Section 4) that uses tightness, Aldous' criterion, martingale-problem convergence, and Hausdorff convergence of the admissible-control sets. The only circularity-adjacent feature is that the pivotal Lemma 3.2 defers to the authors' own unpublished preprint [9] for the comparison of convex expectations that reduces upper semianalytic payoffs to Lipschitz payoffs, and Lemma 3.6 similarly relies on the authors' prior papers [10] and [11] for viscosity-solution identifications. These are load-bearing self-citations, and the present paper does not verify the hypotheses of [9, Theorem 2.12] for the relaxed-control framework with unbounded z. That is a genuine correctness risk, but it is not a circular reduction: [9], [10], and [11] are fixed mathematical results whose assumptions do not include the target convergence theorem, and they are not the conclusions of the present paper restated as premises. No equation in the paper is shown to equal its own input by construction, so the circularity score stays low.
Axiom & Free-Parameter Ledger
axioms (7)
- domain assumption Assumption 2.1(A)-(D): coefficients µ, σ are Borel measurable, continuous in (t, λ), non-anticipative, uniformly equi-Lipschitz, and of linear growth.
- domain assumption Assumption 2.1(E): reward g is increasing, bounded, and continuous.
- standard math Comparison principle for viscosity solutions of second-order path-dependent HJB equations (Zhou 2023, Theorem 6.1).
- domain assumption Comparison theorem for convex expectations on path space (Criens & Kupper [9], Theorems 2.21, 5.2).
- standard math Existence and uniqueness of strong solutions to SDEs with path-dependent coefficients (Jacod [19, Theorem 14.30]).
- standard math Martingale convergence and tightness results (Aldous' criterion; Jacod-Shiryaev Proposition IX.1.12).
- standard math Functional Itô formula for non-anticipative functionals (Cont-Fournié [7, Proposition 7]).
read the original abstract
In this work, we study small-noise asymptotics of risk-sensitive exit-time control problems governed by stochastic differential equations with path-dependent coefficients. Our main result establishes the convergence of the $\log$-transformed exit-time problem to a deterministic control problem with path-dependent coefficients. For its proof, we first derive a novel variational representation for general $\log$-transformed stochastic control problems with path-dependent coefficients, combining tools from the theory of path-dependent partial differential equations and convex expectations on path spaces. In a second step, we use probabilistic methods to analyze the convergence of the resulting variational formulas. To illustrate the scope of our analysis, we consider a computable example for a stochastic differential equation with memory and characterize the limiting problem and associated control strategies.
Reference graph
Works this paper leans on
-
[1]
Nonexponential Sanov and Schilder the- orems on Wiener space: BSDEs, Schr¨ odinger problems and control,
J. Backhoff-Veraguas, D. Lacker, and L. Tangpi, “Nonexponential Sanov and Schilder the- orems on Wiener space: BSDEs, Schr¨ odinger problems and control,”Ann. Appl. Probab., vol. 30, no. 3, pp. 1321–1367, 2020
2020
-
[2]
On ws-convergence of product measures.,
E. J. Balder, “On ws-convergence of product measures.,”Math. Oper. Res., vol. 26, no. 3, pp. 494–518, 2001
2001
-
[3]
D. P. Bertsekas and S. E. Shreve,Stochastic Optimal Control. The Discrete Time Case(Math. Sci. Eng.). Elsevier, Amsterdam, 1978, vol. 139
1978
-
[4]
A variational representation for certain functionals of Brownian motion,
M. Bou´ e and P. Dupuis, “A variational representation for certain functionals of Brownian motion,”Ann. Probab., vol. 26, no. 4, pp. 1641–1659, 1998
1998
-
[5]
Risk-sensitive and robust escape control for degenerate diffusion processes,
M. Bou´ e and P. Dupuis, “Risk-sensitive and robust escape control for degenerate diffusion processes,”Math. Control Signals Syst., vol. 14, no. 1, pp. 62–85, 2001
2001
-
[6]
Carmona and F
R. Carmona and F. Delarue,Probabilistic theory of mean field games with applications I. Mean field FBSDEs, control, and games(Probab. Theory Stoch. Model.). Springer, 2018, vol. 83
2018
-
[7]
Change of variable formulas for non-anticipative functionals on path space,
R. Cont and D.-A. Fourni´ e, “Change of variable formulas for non-anticipative functionals on path space,”J. Funct. Anal., vol. 259, no. 4, pp. 1043–1072, 2010
2010
-
[8]
Crandall-Lions viscosity solutions for path-dependent PDEs: The case of heat equation,
A. Cosso and F. Russo, “Crandall-Lions viscosity solutions for path-dependent PDEs: The case of heat equation,”Bernoulli, vol. 28, no. 1, pp. 481–503, 2022
2022
-
[9]
D. Criens and M. Kupper,Representation Theorems for Convex Expectations and Semigroups on Path Space, (to appear inMath. Oper. Res.), 2025. arXiv:2503.10572 [math.OC]
arXiv 2025
-
[10]
Nonlinear continuous semimartingales,
D. Criens and L. Niemann, “Nonlinear continuous semimartingales,”Electron. J. Probab., vol. 28, p. 146, 2023
2023
-
[11]
Nonlinear semimartingales and Markov processes with jumps,
D. Criens and L. Niemann, “Nonlinear semimartingales and Markov processes with jumps,” J. Evol. Equ., vol. 25, no. 1, p. 39, 2025, Id/No 21
2025
-
[12]
Risk-sensitive and robust escape criteria,
P. Dupuis and W. M. McEneaney, “Risk-sensitive and robust escape criteria,”SIAM J. Con- trol Optim., vol. 35, no. 6, pp. 2021–2049, 1997
2021
-
[13]
Martingale measures and stochastic calculus,
N. El Karoui and S. M´ el´ eard, “Martingale measures and stochastic calculus,”Probab. Theory Relat. Fields, vol. 84, no. 1, pp. 83–101, 1990
1990
-
[14]
Existence of an optimal Markovian filter for the control under partial observations,
N. El Karoui, D. Nguyen, and M. Jeanblanc-Picqu´ e, “Existence of an optimal Markovian filter for the control under partial observations,”SIAM J. Control Optim., vol. 26, no. 5, pp. 1025–1061, 1988
1988
-
[15]
N. El Karoui and X. Tan,Capacities, Measurable Selection and Dynamic Programming Part II: Application in Stochastic Control Problems, 2024. arXiv:1310.3364 [math.OC]
Pith/arXiv arXiv 2024
-
[16]
W. H. Fleming and H. M. Soner,Controlled Markov processes and viscosity solutions(Stoch. Model. Appl. Probab.), 2nd ed. New York, NY: Springer, 2006, vol. 25
2006
-
[17]
Large deviations in safety-critical systems with probabilistic initial conditions,
A. R. Gomez, M. L. Bujorianu, and R. Wisniewski, “Large deviations in safety-critical systems with probabilistic initial conditions,”Automatica J. IF AC, vol. 189, Paper No. 113021, 7, 2026
2026
-
[18]
Hu and N
S. Hu and N. S. Papageorgiou,Handbook of Multivalued Analysis. Volume I: Theory(Math. Appl., Dordr.). Dordrecht: Kluwer Academic Publishers, 1997, vol. 419
1997
-
[19]
Jacod,Calcul Stochastique et Probl` emes de Martingales(Lect
J. Jacod,Calcul Stochastique et Probl` emes de Martingales(Lect. Notes Math.). Springer, Cham, 1979, vol. 714
1979
-
[20]
Jacod and A
J. Jacod and A. N. Shiryaev,Limit Theorems for Stochastic Processes.(Grundlehren Math. Wiss.), 2nd ed. Berlin: Springer, 2003, vol. 288
2003
-
[21]
Asymptotic analysis of nonlinear stochastic risk-sensitive control and differen- tial games,
M. R. James, “Asymptotic analysis of nonlinear stochastic risk-sensitive control and differen- tial games,”Math. Control Signals Syst., vol. 5, no. 4, pp. 401–417, 1992
1992
-
[22]
Optimal control of diffusion processes and Hamilton-Jacobi-Bellman equations. I: The dynamic programming principle and applications,
P.-L. Lions, “Optimal control of diffusion processes and Hamilton-Jacobi-Bellman equations. I: The dynamic programming principle and applications,”Commun. Partial Differ. Equations, vol. 8, pp. 1101–1174, 1983. 24 REFERENCES
1983
-
[23]
Large deviations for non-Markovian diffusions and a path-dependent eikonal equation,
J. Ma, Z. Ren, N. Touzi, and J. Zhang, “Large deviations for non-Markovian diffusions and a path-dependent eikonal equation,”Ann. Inst. Henri Poincar´ e, Probab. Stat., vol. 52, no. 3, pp. 1196–1216, 2016
2016
-
[24]
Survey on path-dependent PDEs,
S. Peng, Y. Song, and F. Wang, “Survey on path-dependent PDEs,”Chin. Ann. Math., Ser. B, vol. 44, no. 6, pp. 837–856, 2023
2023
-
[25]
Comparison of viscosity solutions of semilinear path- dependent PDEs,
Z. Ren, N. Touzi, and J. Zhang, “Comparison of viscosity solutions of semilinear path- dependent PDEs,”SIAM J. Control Optim., vol. 58, no. 1, pp. 277–302, 2020
2020
-
[26]
D. W. Stroock and S. R. S. Varadhan,Multidimensional diffusion processes.(Class. Math.), Reprint of the 2nd correted printing (1997). Berlin: Springer, 2006
1997
-
[27]
Some minimax theorems,
F. Terkelsen, “Some minimax theorems,”Math. Scand., vol. 31, pp. 405–413, 1973
1973
-
[28]
Viscosity solutions to second order path-dependent Hamilton-Jacobi-Bellman equa- tions and applications,
J. Zhou, “Viscosity solutions to second order path-dependent Hamilton-Jacobi-Bellman equa- tions and applications,”Ann. Appl. Probab., vol. 33, no. 6B, pp. 5564–5612, 2023
2023
-
[29]
J. Zhou, N. Touzi, and J. Zhang,Viscosity Solutions for HJB Equations on the Process Space,
-
[2025]
arXiv:2401.04920 [math.OC]. D. Criens - University of Freiburg, Ernst-Zermelo-Str. 1, 79104 Freiburg, Germany. Email address:david.criens@stochastik.uni-freiburg.de F. Fuchs - Department of AI, Data and Decision Sciences, Luiss Guido Carli, Viale Romania 32, 00197 Roma, Italy. Email address:ffuchs@luiss.it
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.