REVIEW 3 major objections 4 minor 1 cited by
A weighted residual process test checks the parametric regression mean even when the number of predictors grows with the sample size, avoiding the collapse of classical integrated conditional moment (ICM) tests.
Reviewed by Pith at T0; open to challenge. T0 means a machine referee read the full paper against a public rubric. the ladder, T0–T4 →
A weighted-residual-process test with smoothed bootstrap avoids the degeneracy of classical ICM tests in regression models with a diverging number of predictors.
T0 review reviewed 2026-08-02 challenge →
load-bearing objection Worth a referee, but the test tests error–covariate independence, not the conditional-mean null it advertises. the 3 major comments →
Model Checking for Regressions Based on Weighted Residual Processes with Diverging Number of Predictors
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
Core claim
The paper's central claim is that parametric regression misspecification can be tested in diverging-p settings by studying a one-dimensional weighted residual process. Under the null, the residual equals the error ε, which is assumed independent of X, so the centered process has zero mean; under alternatives, the mean shift m(X) - m(X, β~0) makes the expectation nonzero. The main theorem shows that, under regularity conditions and p^3 log n / n → 0, the statistic WICM_n converges weakly to the nondegenerate integral ∫|U∞(t)|²φ(t)dt of a Gaussian process, and that the same limit is reproduced by a smooth residual bootstrap. Against fixed alternatives the statistic grows linearly in n; against
What carries the argument
The central object is the weighted residual empirical process Û_n(t) = n^{-1/2} Σ_{i=1}^n (g(X_i) - ḡ){cos(tê_i) + sin(tê_i)}. The test statistic is WICM_n = ∫|Û_n(t)|²φ(t)dt with an even integrable weight φ; with standard normal φ it collapses to a pairwise sum of covariate-centered weights times exp(-(ê_i-ê_j)²/2). This one-dimensional trigonometric construction separates the covariates from the residual differences, avoiding the interpoint distance concentration that makes ICM statistics degenerate in high dimensions. The smooth residual bootstrap adds smoothed noise v_n z_{i,j} to centered resampled residuals, refits the model, and recomputes WICM; Theorem 3 shows the bootstrap and the
Load-bearing premise
The construction assumes the error ε is independent of X under the null; if only E(ε|X) = 0 holds, the weighted characteristic moment E{g0(X) exp(itε)} need not vanish, and a correctly specified mean could be rejected.
What would settle it
Simulate a correctly specified linear mean with heteroskedastic noise, for example Y = Xβ + (1 + |X_1|)ε with ε standard normal independent of X, at n = 400 and p = 10, and run the proposed test at the 5% level. If the rejection rate substantially exceeds 5%, the independence-of-error assumption is doing load-bearing work.
If this is right
- Under the null, WICM_n has a nondegenerate Gaussian limit instead of collapsing to a constant, so the test can hold nominal size as the predictor dimension grows.
- Under any fixed misspecification, the test rejects with probability tending to 1, and under local alternatives within 1/√n of the null it retains nontrivial power.
- The smooth residual bootstrap is asymptotically valid for size and power, replacing the wild bootstrap that fails in the diverging-p regime.
- The statistic is computationally light: a closed-form pairwise kernel on residuals, with no d-dimensional numerical integration.
- The dimension condition p^3 log n / n → 0 is weaker than previously available conditions for such empirical-process tests.
Where Pith is reading between the lines
- The null theory needs ε independent of X, not just E(ε|X)=0; under heteroskedastic errors with a correctly specified mean, E{g0(X) exp(itε)} need not vanish, so the test may over-reject. The paper states this independence assumption explicitly, but practitioners should verify it.
- The asymptotics fix g while the implementation estimates g (directional via a working alternative family, or nonparametric via a Fourier/dimension-reduction basis); a formal proof that estimated g preserves the null limit would close a gap between theory and code.
- Because the test targets the conditional mean only, it will not detect variance misspecification or other distributional departures; a different residual transformation would be needed for those.
- The weight-choice analysis implies a practical extension: if g can be chosen close to the projected departure m(X, β~0) - m(X), the test becomes nearly adaptive, but guarding against its null degeneracy is exactly the paper's concern.
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The paper proposes a new specification test for parametric regression models when the number of predictors p diverges with the sample size n. The test is based on a weighted residual process U_n(t) = n^{-1/2} Σ (g(X_i)-gbar)(cos(t e_i)+sin(t e_i)) for a real-valued weight function g, integrated over t to form WICM_n. The authors claim a nondegenerate Gaussian limit under the null (Theorem 1), consistency against fixed alternatives and nontrivial power at the n^{-1/2} local rate (Theorem 2), and validity of a smooth residual bootstrap (Theorem 3 and Corollary 1), under the rate condition p^3 log n / n → 0. They also propose data-driven choices of g for directional and nonparametric alternatives (Section 4.3). Simulations and a real-data application are reported.
Significance. If valid, this would be a useful addition to the diverging-p model-checking literature, avoiding the degeneracy of classical ICM statistics and the failure of the wild bootstrap. The asymptotic rate condition p^3 log n / n → 0 is weaker than earlier conditions in Tan and Zhu (2019) and Tan et al. (2025), and the bootstrap justification is an important contribution. However, the paper's central claim—that the test checks the conditional mean (1.2)—is compromised by the reliance on full independence of errors and covariates, and the theoretical results do not cover the estimated weight functions actually used in implementation. These are load-bearing issues that undermine the validity of the test as advertised.
major comments (3)
- [Section 1 (after Eq. (1.1)) and Section 3 (Eq. (3.1))] The test is constructed on the identity E{g0(X) exp(ite)} = E{g0(X)}E{exp(ite)} = 0 under H0, which requires ε to be independent of X. The null (1.2) only specifies the conditional mean; the paper even states that the error distribution 'remains unrestricted.' If ε is heteroskedastic but E(ε|X)=0, then for small t, E{g0(X)(cos(tε)+sin(tε))} ≈ -(t²/2)E{g0(X)ε²} ≠ 0, so WICM_n diverges and the test rejects a correctly specified mean with probability tending to 1. Theorems 1–3 condition on independence, which is not part of the stated H0. The test is actually one of independence between ε and X, not of the parametric form of the conditional mean.
- [Section 4.3, Eqs. (4.4)–(4.5)] All theoretical results treat g as a fixed, nonrandom weight function (Assumption 1). In practice, g is estimated from the same data using (4.4) or (4.5). The theorems do not cover the estimated g. Consequently, the size and power simulations and the real-data analysis rely on a procedure without asymptotic justification. This is particularly problematic because the local-power argument in Section 4.3 depends on g being aligned with S; with an estimated g, the expansion (4.1) and the power claims may fail. The Fourier truncation level l and the structural dimension s are additional user-chosen quantities not covered by the theory.
- [Section 5, Tables 1–4] The numerical studies are conducted in regimes where the rate condition p^3 log n / n → 0 is far from satisfied. For example, p=10, n=100 gives p^3 log n / n ≈ 46; the real-data example with p=68, n=1059 gives ≈ 2078. The paper does not discuss this discrepancy. It is unclear whether the reported control of size and power reflects the asymptotic theory or some other mechanism. The real-data application is thus not supported by the paper's own theoretical conditions.
minor comments (4)
- [Throughout] There is a notation inconsistency: the introduction uses d and p for predictor and parameter dimensions, but Section 2 uses p for the covariate dimension and later p becomes the parameter dimension. Please standardize.
- [Section 4, Assumption 8] The assumption states ∫ t^4 φ(t)dt < ∞, but φ was not introduced in Assumption 8; it should be l(t). Also, 'phi' in the bootstrap step should be consistent with φ(t).
- [Section 4.3, p. 22] The sentence 'This choice, however, is not practically useful...' correctly notes the degeneracy of the optimal weight g ∝ m(X,β0)−m(X). But the replacement using the projection onto the score space is ad hoc and no property is proved for it. Please clarify or cite a justification.
- [Section 5.2] The reported p-value 'approximately equal to 0' is vague; provide a numerical upper bound. Also, the claim that the scatter plots 'suggest' nonlinearity is informal; the test result itself is enough.
Circularity Check
No significant circularity: the derivation is self-contained given its modeling assumptions; the main concerns are a null-hypothesis mismatch and unsupported estimated-g asymptotics, not circular reductions.
full rationale
No circular step is exhibited. The test statistic WICM_n in (3.2) is an integral of the squared weighted residual process (3.1), and it is not defined as the fitted value of the quantity it claims to predict; no fitted parameter is renamed as a prediction and no equation reduces to its input by construction. The null moment identity E{g0(X) exp(ite)} = E{g0(X)}E{exp(itε)} = 0 used in Section 3 relies on the independence assumption stated after Eq. (1.1). This is a real correctness/validity concern: the paper's advertised null in (1.2) concerns only the conditional mean, while the construction effectively tests independence between the error and X, so under heteroskedastic errors with E(ε|X)=0 the test can reject a correctly specified mean. That is a hypothesis mismatch rather than circularity, because the independence assumption is an input to the derivation, not the conclusion being derived. Likewise, the estimation of the weight function g in Section 4.3, Eqs. (4.4)-(4.5), is not covered by Theorems 1-3, which treat g as fixed under Assumption 1; this is an omitted proof or unsupported extension, but it does not make the local-power claim equivalent to the fitted input. The citations to Tan et al. (2025), Dette et al. (2007), Neumeyer (2009), and others are external rather than self-citations, and the bootstrap approximation is a standard residual bootstrap. Accordingly, the circularity score is 0.
Axiom & Free-Parameter Ledger
free parameters (4)
- bootstrap smoothing parameter v_n =
0.2 in simulations
- weight function phi(t) =
standard normal density
- Fourier truncation level l for nonparametric weight =
unspecified
- structural dimension s / central subspace B for WICM^(2) =
estimated via CSE and MERE
axioms (6)
- domain assumption epsilon is independent of X
- domain assumption Under H0, Y - m(X, beta_tilde_0) = epsilon and the model contains the true regression function
- ad hoc to paper The weight function g(X) is fixed and satisfies an envelope condition; estimated g is ignored in the proofs
- standard math Standard high-dimensional M-estimation regularity: Assumptions 1-7, p^3 log n / n -> 0, Gaussian-process tightness
- domain assumption Local alternative perturbation S satisfies E{S(X) ḏm(X, beta_tilde_0)} = 0 and g is aligned with S
- domain assumption Smooth residual bootstrap kernel l and smoothing parameter v_n allow uniform density estimation
Cite this review
Pith. "Pith review of Model Checking for Regressions Based on Weighted Residual Processes with Diverging Number of Predictors." pith.science (2026). https://pith.science/paper/PSEJT6MQ
@misc{pith2026260414649,
author = {Pith},
title = {Pith review of: Model Checking for Regressions Based on Weighted Residual Processes with Diverging Number of Predictors},
year = {2026},
howpublished = {\url{https://pith.science/paper/PSEJT6MQ}},
note = {Machine review of arXiv:2604.14649}
}
abstract
The integrated conditional moment (ICM) test is a classical and widely used method for assessing the adequacy of regression models. Although it performs well in fixed-dimension settings, its behavior changes dramatically when the predictor dimension diverges: in such regimes, the limiting null and alternative distributions of the ICM statistic degenerate to fixed constants. Moreover, when the number of predictors diverges, the commonly used wild bootstrap no longer approximates the null distribution of the ICM statistic well, leading to size distortion and substantial power loss. To address these challenges, we propose a new specification test based on weighted residual processes for evaluating the parametric form of the regression mean function in high-dimensional settings where the number of predictors increases with the sample size. We establish the asymptotic properties of the test statistic under the null hypothesis and under global and local alternatives. The proposed test maintains the nominal significance level and can detect local alternatives that deviate from the null hypothesis at the parametric rate $1/\sqrt{n}$. Furthermore, we propose a smooth residual bootstrap to approximate the limiting null distribution and establish its validity in high-dimensional settings. Two simulation studies and a real-data example are conducted to evaluate the finite-sample performance of the proposed test.
Figures
Forward citations
Cited by 1 Pith paper
-
Testing for correct model specification in copula regression models
A kernel-based test of the weighted L2 distance between the true regression function and the copula-regression approximation is consistent and asymptotically normal, with pivotal self-normalized confidence intervals f...
Reference graph
Works this paper leans on
-
[1]
Bierens, H. J. (1982). Consistent model specification tests.Journal of Econometrics, 20(1):105–134. Bierens, H. J. (1990). A consistent conditional moment test of functional form.Economet- rica, 58(6):1443–1458. Bierens, H. J. and Ploberger, W. (1997). Asymptotic theory of integrated conditional moment tests.Econometrica, 65(5):1129–1151. Cook, R. D. (200...
1982
-
[3]
Van Keilegom, I., Gonz´ alez Manteiga, W., and S´ anchez Sellero, C
Cambridge university press. Van Keilegom, I., Gonz´ alez Manteiga, W., and S´ anchez Sellero, C. (2008). Goodness-of-fit tests in parametric regression based on the estimation of the error distribution.Test, 17:401–415. Zheng, J. X. (1996). A consistent test of functional form via nonparametric estimation techniques.Journal of Econometrics, 75(2):263–289....
2008
-
[961]
and Lavergne, P
Guerre, E. and Lavergne, P. (2005). Data-driven rate-optimal specification testing in re- gression models.The Annals of Statistics, 33(2):840–870. Guo, X., Wang, T., and Zhu, L. (2016). Model checking for parametric single-index models: a dimension reduction model-adaptive approach.Journal of the Royal Statistical Society Series B: Statistical Methodology...
2005
-
[1947]
Hastie, T., Tibshirani, R., and Wainwright, M. (2015). Statistical learning with sparsity. Monographs on statistics and applied probability, 143(143):8. 33 Horowitz, J. L. and H¨ ardle, W. (1994). Testing a parametric model against a semipara- metric alternative.Econometric theory, 10(5):821–848. Khmaladze, E. V. and Koul, H. L. (2004). Martingale transfo...
2015
This paper was first reviewed by deepseek-v4-flash on August 2, 2026.
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.