Pith. sign in

REVIEW 2 cited by

Frequency Bias in Neural Networks for Input of Non-Uniform Density

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2003.04560 v1 pith:HRGX2CU7 submitted 2020-03-10 cs.LG stat.ML

classification cs.LGstat.ML
keywords networksfrequencydensityneuralbiasconvergencedatadeep
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

Recent works have partly attributed the generalization ability of over-parameterized neural networks to frequency bias -- networks trained with gradient descent on data drawn from a uniform distribution find a low frequency fit before high frequency ones. As realistic training sets are not drawn from a uniform distribution, we here use the Neural Tangent Kernel (NTK) model to explore the effect of variable density on training dynamics. Our results, which combine analytic and empirical observations, show that when learning a pure harmonic function of frequency $\kappa$, convergence at a point $\x \in \Sphere^{d-1}$ occurs in time $O(\kappa^d/p(\x))$ where $p(\x)$ denotes the local density at $\x$. Specifically, for data in $\Sphere^1$ we analytically derive the eigenfunctions of the kernel associated with the NTK for two-layer networks. We further prove convergence results for deep, fully connected networks with respect to the spectral decomposition of the NTK. Our empirical study highlights similarities and differences between deep and shallow networks in this model.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Parity Supervision as a Driver of Generalization in Quantum Generative Modeling

    quant-ph 2026-05 unverdicted novelty 6.0 of 10

    Parity supervision improves exact KL fit and recovery of unseen high-value states in IQP Born machines beyond MSE training or max-entropy controls via parity-moment evidence transfer.

  2. Performance of Krotov, PRONTO and PINN for optimal control of quantum gates

    quant-ph 2026-07 conditional novelty 4.0 of 10

    An enhanced PINN with Fourier features, per-epoch normalization, and pretraining designs high-fidelity quantum gates comparable to Krotov and PRONTO, but more slowly and with overclaimed headline results.

Pith tools