Pith. sign in

REVIEW 1 cited by

Preconditioning for Scalable Gaussian Process Hyperparameter Optimization

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2107.00243 v5 pith:XXWBDMXJ submitted 2021-07-01 cs.LG cs.NAmath.NA

classification cs.LGcs.NAmath.NA
keywords log-determinantoptimizationpreconditioningconvergencegaussianhyperparameterkernellinear
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Gaussian process hyperparameter optimization requires linear solves with, and log-determinants of, large kernel matrices. Iterative numerical techniques are becoming popular to scale to larger datasets, relying on the conjugate gradient method (CG) for the linear solves and stochastic trace estimation for the log-determinant. This work introduces new algorithmic and theoretical insights for preconditioning these computations. While preconditioning is well understood in the context of CG, we demonstrate that it can also accelerate convergence and reduce variance of the estimates for the log-determinant and its derivative. We prove general probabilistic error bounds for the preconditioned computation of the log-determinant, log-marginal likelihood and its derivatives. Additionally, we derive specific rates for a range of kernel-preconditioner combinations, showing that up to exponential convergence can be achieved. Our theoretical results enable provably efficient optimization of kernel hyperparameters, which we validate empirically on large-scale benchmark problems. There our approach accelerates training by up to an order of magnitude.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Assessing Quantum Advantage for Gaussian Process Regression

    quant-ph 2025-05 accept novelty 5.0 of 10

    Quantum algorithms for Gaussian process regression lose their exponential speedup because kernel matrix condition numbers grow at least linearly with dataset size.

Pith tools