Pith. sign in

REVIEW 3 cited by

An Equivalence Principle for the Spectrum of Random Inner-Product Kernel Matrices with Polynomial Scalings

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2205.06308 v2 pith:4KZCBYDC submitted 2022-05-12 math.PR stat.ML

classification math.PRstat.ML
keywords kernelmatrixrandommatriceslinearpolynomialspectrumcombination
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

We investigate random matrices whose entries are obtained by applying a nonlinear kernel function to pairwise inner products between $n$ independent data vectors, drawn uniformly from the unit sphere in $\mathbb{R}^d$. This study is motivated by applications in machine learning and statistics, where these kernel random matrices and their spectral properties play significant roles. We establish the weak limit of the empirical spectral distribution of these matrices in a polynomial scaling regime, where $d, n \to \infty$ such that $n / d^\ell \to \kappa$, for some fixed $\ell \in \mathbb{N}$ and $\kappa \in (0, \infty)$. Our findings generalize an earlier result by Cheng and Singer, who examined the same model in the linear scaling regime (with $\ell = 1$). Our work reveals an equivalence principle: the spectrum of the random kernel matrix is asymptotically equivalent to that of a simpler matrix model, constructed as a linear combination of a (shifted) Wishart matrix and an independent matrix sampled from the Gaussian orthogonal ensemble. The aspect ratio of the Wishart matrix and the coefficients of the linear combination are determined by $\ell$ and the expansion of the kernel function in the orthogonal Hermite polynomial basis. Consequently, the limiting spectrum of the random kernel matrix can be characterized as the free additive convolution between a Marchenko-Pastur law and a semicircle law. We also extend our results to cases with data vectors sampled from isotropic Gaussian distributions instead of spherical distributions.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. On the edge eigenvalues of sparse random geometric graphs

    math.PR 2025-09 conditional novelty 8.0 of 10

    For Gaussian-sampled random geometric graphs in the sparse regime, the first nontrivial edge eigenvalues of the scaled random-walk Laplacian converge in probability to 2(k-1)/sigma^2-type eigenvalues of a weighted Lap...

  2. Statistical Limits for Finite-Rank Tensor Estimation

    cs.IT 2025-06 conditional novelty 7.0 of 10

    A general q-wise interaction model yields asymptotically exact free energy and MMSE formulas, unifying and extending prior results for heteroskedastic tensors and higher-order assignment problems.

  3. Models of Heavy-Tailed Mechanistic Universality

    stat.ML 2025-06 conditional novelty 6.0 of 10

    A new random matrix model with one structure parameter explains heavy-tailed spectra in trained networks, and yields scaling laws, optimizer-tail behavior, and a description of the five-plus-one phases of training.

Pith tools