REVIEW 3 major objections 4 minor 50 references
Truncated signatures learn smooth path functionals at rate K^{-2γ}.
Reviewed by Pith at T0; open to challenge. T0 means a machine referee read the full paper against a public rubric. the ladder, T0–T4 →
T0 review · deepseek-v4-flash
2026-08-01 16:46 UTC pith:CVKAQ7KA
load-bearing objection The minimax rate theorem is a real contribution and the paper should go to review, but the OLS consistency proof has a specific gap: the covariance concentration bound ignores exponential K-dependence in the signature moments. the 3 major comments →
How Fast Do Signatures Learn? Statistical Theory and Applications for Path Regression
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
Core claim
The paper proves that for (X,h) in the smooth diffusion-functional class U_{γ,R}, the projection residual E|ξ^X_K(Y)|^2 is bounded by C H_γ(h)^2 K^{-2γ}, and the supremum over the class is of exact order K^{-2γ}. Smoother coefficient functions yield faster convergence of the truncated signature approximation, and the exponent 2γ cannot be improved uniformly. It then shows that, under a uniform spectral-gap condition on the signature covariance matrix and a polynomial decay bound on the truncation residual, Signature-OLS is consistent and asymptotically normal, Signature-LASSO recovers the active signature coordinates with high probability, and Signature-Logistic estimates the latent logit co
What carries the argument
The central object is the time-augmented path signature: the ordered hierarchy of iterated integrals of cX_t = (t, X_t). Theorem 1 is built on a localization argument that confines the diffusion path to a box of size K^β; Jackson polynomial approximation of the coefficient functions on that box; the shuffle identity, which converts polynomial coefficient functionals into level-K signature functionals; and sub-Gaussian tail control for the event that the path leaves the box. For Brownian motion, a signature-chaos lemma shows that the level-K signature space is exactly the polynomial part of the first K Wiener chaos kernels, which is why the minimax lower bound is sharp. The same rate enters t
Load-bearing premise
Every consistency result assumes the covariance matrix of the level-K signature features keeps its smallest eigenvalue bounded away from zero for all K, and that the truncation residual decays at a polynomial rate; the authors themselves note that the full signature dictionary grows exponentially, so these conditions are only plausible for analytic targets or a pre-selected feature subset whose spectral properties are not demonstrated.
What would settle it
Compute the smallest eigenvalue of the sample covariance matrix of the level-K signature features used in the battery or EEG application for K = 1,2,3,4; if λ_min is not bounded below by a positive constant as K grows, Assumption 1(i) fails and the Signature-OLS and Signature-Logistic consistency theorems do not govern those fitted models.
If this is right
- Truncation level K can now be treated as a smoothing parameter: for a target of regularity γ, squared bias decays like K^{-2γ}, so the optimal K in finite samples balances this bias against the d_K/n variance term.
- Signature-OLS is consistent and asymptotically normal when d_K^2/n → 0 and K^Q/d_K^2 → ∞; equivalently, the truncation bias must be negligible relative to estimation error.
- Signature-LASSO can achieve support recovery even when the full signature dictionary is much larger than n, provided the active set stays small and the irrepresentability condition holds; the exponent Q drops out of the selection rate.
- Signature-Logistic gives consistent latent-score estimates under d_K^2/n → 0 and d_K K^{-Q} → 0, so binary classification with path covariates inherits the same bias-variance logic.
- For analytic coefficient functions, the truncation error becomes exponential, which makes a logarithmic choice of K (roughly log n / (2 log ρ)) theoretically optimal.
Where Pith is reading between the lines
- If the approximation rate extends to barrier, stopping-time, and occupation-time functionals—listed by the authors as open—the same consistency framework would transfer to simulation-based pricing and optimal stopping, where the truncation level would play the role of a basis dimension with known error decay.
- A testable consequence of the exponential-rate result is that for very smooth path functionals a logarithmic truncation schedule should outperform deeper signatures; a cross-validated comparison on synthetic analytic targets would settle how tight the constant ρ is.
- The paper's own remark after the OLS theorem shows that the full signature dictionary cannot satisfy the polynomial-rate conditions, so the applied value of the theory hinges on the spectral gap of a selected feature subset; checking λ_min for the reported 155- and 858-feature dictionaries would directly test whether the theorems govern those implementations.
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The paper develops a quantitative theory of truncated signatures for path regression. Its central approximation result, Theorem 1, states that for a class of smooth functionals of Itô diffusions — functionals built from time integrals and Stratonovich integrals with coefficient functions of smoothness γ — the squared L2 truncation error of the level-K signature is O(K^{-2γ}), and that this rate is minimax sharp over the class U_{γ,R}. The paper then propagates a generic residual rate E|ξ|^2=O(K^{-Q}) through three estimators: Signature-OLS, Signature-LASSO, and Signature-Logistic, giving consistency and, for OLS, asymptotic normality. Simulations illustrate the K^{-2γ} ordering and the approximation–estimation tradeoff. Three applications (FX volatility, battery end-of-life, EEG seizure detection) compare signature features with handcrafted benchmarks.
Significance. If correct, the paper supplies a genuinely useful quantitative complement to the qualitative signature universal approximation theorem, and the explicit residual-propagation framework is a step forward for the statistical signature literature. The paper is also well served by its technical apparatus: Lemma A.4 gives a concrete, admissible sub-Weibull constant, Theorem 1 is proved with a localization-plus-Jackson argument, and the appendices contain full proofs and substantial simulation/application detail. The empirical sections honestly report overfitting (notably Signature-OLS in the battery application), which strengthens credibility. The central approximation theorem appears sound. However, I find the skeptic's concern about Lemma A.3(i) valid and load-bearing: the covariance-concentration step that underpins the OLS and logistic theorems ignores the K-dependence of the signature-coordinate tail constants, so the stated sufficient conditions do not prove the claimed statistical rates. This is repairable, but it is more than a local edit.
major comments (3)
- [Appendix B.1, Lemma A.3(i); Theorem 2] The proof asserts E‖s_i‖^4 = O(d_K^2) from "bounded fourth moment" of each coordinate. This is inconsistent with the paper's own Lemma A.4, where a level-k coordinate satisfies ‖S^I−ES^I‖_{ψ_{2/k}} ≤ C M_ψ^k with M_ψ ≥ 32e > 1. Consequently E[(S^I)^4] grows like M_ψ^{4k}, and summing over the d^k words of level k gives E‖s_K‖^4 = O((d M_ψ^4)^K), not O(d_K^2) with K-independent constants. The Frobenius bound (A.68) therefore needs an additional factor of order M_ψ^{4K}, and consistency in Lemma A.3(i) requires a condition such as n/(M_ψ^{4K} d_K^2)→∞ (up to polynomial factors), strictly stronger than d_K^2/n→0. Theorem 2(i)–(ii) relies on this step; as written, the displayed rate conditions do not establish the OLS consistency or asymptotic normality claims.
- [Sec. 4.1 Assumption 1 and Remark after Theorem 2; Sec. 6.1.2, 6.2.2, 6.3.2] Assumption 1(i) postulates a uniformly well-conditioned Gram matrix for all K, but the full time-augmented signature is exactly singular because of pure-time coordinates, as the paper itself notes. The statistical theorems therefore apply only to a preselected feature subset whose spectral properties are assumed rather than proved, and the relation between the full-signature rate in Theorem 1 and the residual of the selected subset is not established. In the applications, no eigenvalue diagnostics are reported, and the feature counts/sample sizes do not satisfy the displayed theorem conditions: 858 features with n≈2000 in Sec. 6.1.2 gives d_K^2/n≈368; 155 features with n=166 in Sec. 6.2.2 gives d_K^2/n≈145; the level-2 EEG features are even larger. Thus the consistency theorems are not demonstrated for the estimators actually implemented.
- [Appendix B.4, Theorem 4] The proof of Theorem 4 localizes the empirical logistic objective by asserting sup_{‖L−L0K‖≤r} ‖∇²L_n(L) − H_K(L)‖ = o_P(1) "by the same concentration argument as in the least-squares case." Since the Hessian involves rank-one matrices s_i s_i^T, the same missing M_ψ^{4K} factor from Lemma A.3(i) enters here. Without an additional condition such as n/(M_ψ^{4K} d_K^2)→∞, the localization and the displayed rate in Eq. (26) are unproved. The asymmetry with Theorem 3, which explicitly tracks M_ψ^{2K} in its rate condition, underscores that this is not just a stylistic omission.
minor comments (4)
- [Appendix B.3, Step 2] The proof of Theorem 3 writes "∥Y∥_{L_p} ≤ C p^{K/2} for every K≥1". Since Y is the target functional and does not depend on K, this should be p^{1/2} (Y is sub-Gaussian under the bounded-coefficient assumptions). As written it is confusing and technically wrong, although the subsequent bound can absorb the correct factor.
- [Sec. 2.1, Eq. (1)] The iterated integral in Eq. (1) is written without specifying Itô vs Stratonovich convention, while the text says Stratonovich is used throughout. Please state the convention at the definition or immediately after it, especially because Eq. (1) is also used for Itô integrals in Appendix A.2.
- [Sec. 6.1.2] The formula log n/(2 log(d+1)) with d+1=3 gives the heuristic K=3 for a single day's signature, but the actual feature vector concatenates 22 daily signatures and has 858 features. The heuristic should be stated as applying to the per-day dictionary only, otherwise the reader may infer a feature dimension much smaller than 858.
- [Fig. 1] The horizontal axis is labeled "full signature depth K (log scale)" but the axis is discrete (2 to 9). A linear depth axis or a clear note that only integer depths are shown would be clearer.
Circularity Check
No significant circularity; the main approximation rate and the statistical consistency theorems are modular, with the truncation-rate condition supplied by an independent proof rather than fitted or assumed from self-citation.
full rationale
The central derivation is self-contained. Theorem 1 proves the L2 truncation bound E|ξ_X^K(Y)|^2 ≤ C H_γ(h)^2 K^{-2γ} for the smooth diffusion-functional class U_{γ,R} using localization, Jackson polynomial approximation, the shuffle identity, and BDG/Itô-isometry estimates; the minimax lower bound is obtained from an explicit Brownian first-chaos subclass, not from the upper bound or from any fitted quantity. The statistical theorems are stated under explicit assumptions: Assumption 1(ii) posits E[ξ^2]=O(K^{-Q}), and the paper identifies this rate as supplied by Theorem 1 for its class (‘The second condition is the approximation-rate input supplied by Section 3’), which is a modular input, not a circular one. The OLS, LASSO, and logistic proofs carry the projection residual through standard concentration and optimization arguments; no fitted parameter is renamed as a prediction. The self-citations involving overlapping authors (Guo et al. 2025 for irrepresentability of signature dictionaries, Bayer et al. 2026 for background) are not load-bearing: the theorems treat irrepresentability and spectral nondegeneracy as assumptions rather than deriving them from those citations. The reviewer-flagged issue about K-dependent sub-Weibull constants in Lemma A.3 is a possible correctness/rate-condition gap, not a circular reduction, and therefore does not affect the circularity score.
Axiom & Free-Parameter Ledger
free parameters (3)
- EEG decision threshold =
0.10
- LASSO/elastic-net regularization strength λ =
5-fold CV (FX); not fully specified (battery, EEG)
- Application truncation levels K =
K=3 (FX, battery), K=2 (EEG)
axioms (7)
- domain assumption b and σ globally Lipschitz and uniformly bounded, Eq. (9); X solves the Itô diffusion SDE, Eq. (8)
- domain assumption Functional class (Eq. 10): Y = Σ_a ∫ h_a ∘ dX̃^a with smooth h; mixed smoothness norm H_γ (Eq. 11) requiring 2γ+1 / 2γ+2 spatial derivatives
- standard math Jackson polynomial approximation bounds on [0,T]×[-m,m]^d with explicit scaling, and classical L2 polynomial approximation lower bounds for Sobolev functions
- standard math Burkholder-Davis-Gundy inequality (Schachermayer-Stebegg) and sub-Weibull tail bounds for signature coordinates (Lemma A.4)
- standard math Wiener-Itô chaos expansion and Lemmas A.1/A.2 (Itô-Stratonovich span equality; signature-chaos polynomial identification)
- domain assumption LASSO irrepresentability condition, Assumption 2(i), inherited from Zhao-Yu (2006)/Wainwright (2009); applicability for Brownian signatures deferred to Guo et al. (2025)
- domain assumption Spectral uniformity: c < λ_min(Σ_K) ≤ λ_max(Σ_K) < C (Assumption 1(i)) and local logistic Hessian non-degeneracy (Assumption 3(i))
read the original abstract
Many prediction and decision-making problems in operations research involve path-valued covariates -- data that evolve over time -- for which path signatures have become a canonical feature representation. Their use is justified by a universal approximation theorem, but this is an existence result: it guarantees that a finite-level signature can approximate any continuous path functional, without quantifying how fast the approximation error decreases as the truncation level grows. This paper develops approximation and statistical theory for signature-based path regression. We establish an \(L^2\) approximation rate for smooth functionals of It\^{o} diffusions and show that it is minimax optimal. We then propagate the truncation error through three statistical learning procedures -- Signature-OLS, Signature-LASSO, and Signature-Logistic -- and establish their consistency. Three real-data applications show that signatures provide informative finite-dimensional representations of path-valued covariates and can improve prediction relative to handcrafted features, in the context of finance -- foreign exchange realized volatility forecasting from intraday price paths; energy -- battery end-of-life prediction from early diagnostic current-voltage pulse paths; and medicine -- epileptic seizure detection from short electroencephalogram windows.
Figures
Reference graph
Works this paper leans on
-
[3]
Let bAdenote the selected support
Results are averaged over 100 independent replications. Let bAdenote the selected support. Prediction is evaluated by independent test MSE. The selection metric in the main text is precision for the OU-relevant family, Precision = | bA∩A OU K | | bA| ,(A.141) with the convention that precision is zero when bA=∅. For the selection diagnostic, the LASSO pen...
2094
-
[4]
We first prove several technical lemmas
Throughout the appendix, we maintain the standing assumption that the underlying path is an Itˆ o diffusion satisfying the conditions in A-14 Section 3, and that the target variableYis generated from the pair (X,h)∈ U γ,R. We first prove several technical lemmas. Lemma A.3 establishes the convergence rate of the sample covariance matrix and the convergenc...
2021
-
[5]
doi: 10.1111/1467-9868.00336. O. E. Barndorff-Nielsen and N. Shephard. Power and bipower variation with stochastic volatility and jumps.Journal of Financial Econometrics, 2(1):1–37,
-
[10]
1109/TNSRE.2015.2505238. M. Zabihi, S. Kiranyaz, V. J ¨antti, T. Lipping, and M. Gabbouj. Patient-specific seizure detection using nonlinear dynamics and nullclines.IEEE Journal of Biomedical and Health Informatics, 24(2):543–555,
arXiv 2015
-
[13]
doi: 10.48550/arXiv.1603.03788. I. Chevyrev and T. Lyons. Characteristic functions of measures on geometric rough paths.The Annals of Probability, 44(6):4049–4082,
-
[14]
doi: 10.1214/15-AOP1068. I. Chevyrev and H. Oberhauser. Signature moments to characterize laws of stochastic processes. Journal of Machine Learning Research, 23(176):1–42,
-
[17]
doi: 10.1214/11-AOP721. F. Corsi. A simple approximate long-memory model of realized volatility.Journal of Financial Econometrics, 7(2):174–196,
-
[21]
doi: 10.1016/j.jmva.2022.105031. G. Flint, B. Hambly, and T. Lyons. Discretely sampled signals and the rough hoff process.Stochastic Processes and their Applications, 126(9):2593–2614,
arXiv 2022
-
[22]
doi: 10.1016/j.spa.2016.02.011. P. K. Friz and N. B. Victoir.Multidimensional Stochastic Processes as Rough Paths: Theory and Applications. Cambridge University Press,
-
[23]
doi: 10.1017/CBO9780511845079. M. Fujita, N. Sugiura, and S. Kouketsu. Prediction of atmospheric profiles with machine learn- ing using the signature method.Geophysical Research Letters, 51(6),
-
[24]
doi: 10.1287/opre.2024.1133. B. Hambly and T. Lyons. Uniqueness for the signature of a path of bounded variation and the reduced path group.Annals of Mathematics, 171(1):109–167,
arXiv 2024
-
[25]
doi: 10.4007/annals.2010. 171.109. S. H¨ormann and P. Kokoszka. Weakly dependent functional data.The Annals of Statistics, 38(3): 1845–1884,
-
[26]
doi: 10.1214/09-AOS768. R. Ibraheem, P. Dechent, and G. dos Reis. Path signature-based life prognostics of li-ion battery using pulse test data.Applied Energy, 378:124820,
-
[27]
doi: 10.1016/j.apenergy.2024.124820. 36 P. Kidger, P. Bonnier, I. Perez Arribas, C. Salvi, and T. Lyons. Deep signature transforms.Advances in Neural Information Processing Systems, 32,
arXiv 2024
-
[28]
doi: 10.48550/arXiv.1905.08494. F. J. Kir´ aly and H. Oberhauser. Kernels for sequentially ordered data.Journal of Machine Learning Research, 20(31):1–45,
-
[29]
doi: 10.5555/3322706.3361972. S. Kiranyaz, T. Ince, M. Zabihi, and D. Ince. Automated patient-specific classification of long-term electroencephalography.Journal of Biomedical Informatics, 49:16–31,
-
[31]
doi: 10.48550/arXiv.1309
-
[33]
doi: 10.1080/1350486X.2021.1891555. T. J. Lyons. Differential equations driven by rough signals.Revista Matem´ atica Iberoamericana, 14(2):215–310,
arXiv 2021
-
[38]
doi: 10.1016/j.jpowsour.2022.231127. 37 W. Schachermayer and F. Stebegg. The sharp constant for the burkholder–davis–gundy inequality and non-smooth pasting.Bernoulli, 24(4A):3032–3051,
arXiv 2022
-
[40]
doi: 10.1007/s10916-019-1234-4. K. A. Severson, P. M. Attia, N. Jin, N. Perkins, B. Jiang, Z. Yang, M. H. Chen, M. Aykol, P. K. Herring, D. Fraggedakis, M. Z. Bazant, S. J. Harris, W. C. Chueh, and R. D. Braatz. Data-driven prediction of battery cycle life before capacity degradation.Nature Energy, 4(5):383–391,
-
[41]
doi: 10.1038/s41560-019-0356-8. A. H. Shoeb and J. V. Guttag. Application of machine learning to epileptic seizure detection. In Proceedings of the 27th International Conference on Machine Learning, pages 975–982,
-
[42]
doi: 10.5555/3104322.3104446. S. A. van de Geer. High-dimensional generalized linear models and the lasso.The Annals of Statistics, 36(2):614–645,
-
[44]
doi: 10.1109/TIT.2009.2016018. M. Zabihi, S. Kiranyaz, A. B. Rad, A. K. Katsaggelos, M. Gabbouj, and T. Ince. Analysis of High- Dimensional Phase Space via Poincar´ e Section for Patient-Specific Seizure Detection.IEEE Transactions on Neural Systems and Rehabilitation Engineering, 24(3):386–398,
arXiv 2009
-
[46]
doi: 10.1109/JBHI.2019.2906400. H. Zhang and S. X. Chen. Concentration inequalities for statistical inference.Communications in Mathematical Research, 37(1):1–85,
arXiv 2019
-
[47]
doi: 10.4208/cmr.2020-0041. P. Zhao and B. Yu. On model selection consistency of lasso.Journal of Machine Learning Research, 7:2541–2563,
-
[1957]
doi: 10.2307/1969671. I. Chevyrev and A. Kormilitzin. A primer on the signature method in machine learning.arXiv preprint arXiv:1603.03788,
-
[1998]
doi: 10.4171/RMI/240. J. Morrill, C. Salvi, P. Kidger, and J. Foster. Neural rough differential equations for long time series. InProceedings of the 38th International Conference on Machine Learning, pages 7829–7838,
-
[2001]
doi: 10.1093/rfs/14.1.113. T. Lyons, S. Nejad, and I. Perez Arribas. Non-parametric pricing and hedging of exotic derivatives. Applied Mathematical Finance, 27(6):457–494,
-
[2002]
doi: 10.1111/1468-0262.00274. O. E. Barndorff-Nielsen and N. Shephard. Econometric analysis of realised volatility and its use 34 in estimating stochastic volatility models.Journal of the Royal Statistical Society: Series B (Statistical Methodology), 64(2):253–280,
-
[2003]
doi: 10.1111/1468-0262.00418. T. G. Andersen, T. Bollerslev, and F. X. Diebold. Roughing it up: Including jump components in the measurement, modeling, and forecasting of return volatility.The Review of Economics and Statistics, 89(4):701–720,
-
[2004]
doi: 10.1093/jjfinec/nbh001. C. Bayer, L. Pelizzari, and J. Schoenmakers. Primal and dual optimal stopping with signatures. Finance and Stochastics, 29:981–1014,
-
[2006]
doi: 10.5555/1248547.1248637. 38 Supplementary Appendices (Electronic Companion) Appendix Contents A Theoretical Details for Section 3 A-1 A.1 Technical Lemmas . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . . A-2 A.2 Signature-Chaos Relation and Brownian Signature Approximation Rate . . . . . . . A-3 A.3 Proof of Theorem 1 . . . ....
-
[2007]
doi: 10.1162/rest.89.4.701. P. M. Attia, A. Grover, N. Jin, K. A. Severson, T. M. Markov, Y.-H. Liao, M. H. Chen, B. Cheong, N. Perkins, Z. Yang, P. K. Herring, M. Aykol, S. J. Harris, R. D. Braatz, S. Ermon, and W. C. Chueh. Closed-loop optimization of fast-charging protocols for batteries with machine learning. Nature, 578:397–402,
-
[2008]
doi: 10.1214/009053607000000929. M. J. Wainwright. Sharp thresholds for high-dimensional and noisy sparsity recovery usingℓ 1- constrained quadratic programming.IEEE Transactions on Information Theory, 55(5):2183– 2202,
-
[2009]
doi: 10.1093/jjfinec/nbp001. C. Cuchiero, G. Gazzani, and S. Svaluto-Ferro. Signature-based models: Theory and calibration. SIAM Journal on Financial Mathematics, 14(3):910–957,
-
[2010]
doi: 10.1016/j.jfa.2010.04.017. 35 R. Cont and D.-A. Fourni´ e. Functional itˆ o calculus and stochastic integral representation of mar- tingales.The Annals of Probability, 41(1):109–133,
-
[2011]
doi: 10.1016/j.jeconom.2010.03.034. A. J. Patton and K. Sheppard. Good volatility, bad volatility: Signed jumps and the persistence of volatility.The Review of Economics and Statistics, 97(3):683–697,
-
[2013]
doi: 10.3150/11-BEJ410. H. Boedihardjo, X. Geng, T. Lyons, and D. Yang. The signature of a rough path: Uniqueness. Advances in Mathematics, 293:720–737,
-
[2014]
doi: 10.1016/j.jbi. 2014.02.005. D. Levin, T. Lyons, and H. Ni. Learning from the past, predicting the statistics for the future, learning an evolving system.arXiv preprint arXiv:1309.0260,
Pith/arXiv arXiv 2014
-
[2015]
doi: 10.1162/REST a 00503. N. H. Paulson, J. Kubal, L. Ward, S. Saxena, W. Lu, and S. J. Babinec. Feature engineering for machine learning enabled early prediction of battery lifetime.Journal of Power Sources, 527: 231127,
-
[2016]
doi: 10.1016/j.aim.2016.02.011. K.-T. Chen. Integration of paths, geometric invariants and a generalized baker-hausdorff formula. Annals of Mathematics, 65(1):163–178,
-
[2018]
doi: 10.3150/17-BEJ935. R. S. Selvakumari, M. Mahalakshmi, and P. Prashalee. Patient-Specific Seizure Detection Method using Hybrid Classifier with Optimized Electrodes.Journal of Medical Systems, 43(5):121,
-
[2019]
doi: 10.1080/ 14697688.2019.1575974. A. Fermanian. Functional linear regression with truncated signatures.Journal of Multivariate Analysis, 192:105031,
arXiv 2019
-
[2020]
doi: 10.1038/s41586-020-1994-5. Y. A¨ıt-Sahalia. Maximum likelihood estimation of discretely sampled diffusions: a closed-form approximation approach.Econometrica, 70(1):223–262,
-
[2021]
doi: 10.48550/arXiv.2009.08295. A. J. Patton. Volatility forecast comparison using imperfect volatility proxies.Journal of Econo- metrics, 160(1):246–256,
-
[2022]
doi: 10.48550/arXiv.1810.10971. R. Cont and D.-A. Fourni´ e. Change of variable formulas for non-anticipative functionals on path space.Journal of Functional Analysis, 259(4):1043–1072,
work page internal anchor Pith review Pith/arXiv arXiv doi:10.48550/arxiv.1810.10971
-
[2023]
doi: 10.1137/22M1512338. B. Dupire. Functional itˆ o calculus.Quantitative Finance, 19(5):721–729,
-
[2024]
doi: 10.1137/23M1571563. A. Belloni and V. Chernozhukov. Least squares after model selection in high-dimensional sparse models.Bernoulli, 19(2):521–547,
-
[2025]
doi: 10.1007/s00780-025-00570-8. C. Bayer, G. dos Reis, B. Horvath, and H. Oberhauser, editors.Signature Methods in Finance: An Introduction with Computational Applications. Springer Finance Lecture Notes. Springer,
-
[2026]
doi: 10.1007/978-3-031-97239-3. E. Bayraktar, Q. Feng, and Z. Zhang. Deep signature algorithm for multidimensional path- dependent options.SIAM Journal on Financial Mathematics, 15(1):194–214,
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.