Pith. sign in

Wide neural networks of any depth evolve as linear models under gradient descent * , volume=

1 Pith paper cite this work, alongside 283 external citations. Polarity classification is still indexing.

1 Pith paper citing it
283 external citations · OpenAlex

fields

cs.LG 1

years

2026 1

verdicts

ACCEPT 1

representative citing papers

Quantitative Gaussian-Process limits of Tensor Programs

cs.LG · 2026-07-07 · accept · novelty 6.0

Finite-width executions of Netsor tensor programs converge to their infinite-width Gaussian-process limits in Wasserstein distance at rate O(1/√n) per hidden width, covering weight-sharing architectures including RNNs and attention.

citing papers explorer

Showing 1 of 1 citing paper.

  • Quantitative Gaussian-Process limits of Tensor Programs cs.LG · 2026-07-07 · accept · none · ref 5

    Finite-width executions of Netsor tensor programs converge to their infinite-width Gaussian-process limits in Wasserstein distance at rate O(1/√n) per hidden width, covering weight-sharing architectures including RNNs and attention.