Pith. sign in

Exact Gaussian Processes on a Million Data Points

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Gaussian processes (GPs) are flexible non-parametric models, with a capacity that grows with the available data. However, computational constraints with standard inference procedures have limited exact GPs to problems with fewer than about ten thousand training points, necessitating approximations for larger datasets. In this paper, we develop a scalable approach for exact GPs that leverages multi-GPU parallelization and methods like linear conjugate gradients, accessing the kernel matrix only through matrix multiplication. By partitioning and distributing kernel matrix multiplies, we demonstrate that an exact GP can be trained on over a million points, a task previously thought to be impossible with current computing hardware, in less than 2 hours. Moreover, our approach is generally applicable, without constraints to grid data or specific kernel classes. Enabled by this scalability, we perform the first-ever comparison of exact GPs against scalable GP approximations on datasets with $10^4 \!-\! 10^6$ data points, showing dramatic performance improvements.

fields

cs.LG 1

years

2025 1

verdicts

ACCEPT 1

representative citing papers

New Bounds for Sparse Variational Gaussian Processes

cs.LG · 2025-02-12 · accept · novelty 5.0

Replacing the conditional GP prior inside sparse variational inference with a Gaussian sharing its mean but using diagonal, analytically optimal variance corrections yields a strictly tighter evidence lower bound at unchanged O(NM^2) cost.

citing papers explorer

Showing 1 of 1 citing paper.

  • New Bounds for Sparse Variational Gaussian Processes cs.LG · 2025-02-12 · accept · none · ref 2021 · internal anchor

    Replacing the conditional GP prior inside sparse variational inference with a Gaussian sharing its mean but using diagonal, analytically optimal variance corrections yields a strictly tighter evidence lower bound at unchanged O(NM^2) cost.