A hybrid VI-HMC method runs expensive Hamiltonian Monte Carlo only on the neural network parameters most responsible for predictive uncertainty, cutting sampling cost while approximating full HMC posteriors.
Sequential Monte Carlo with active subspaces
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
Monte Carlo methods, such as Markov chain Monte Carlo (MCMC), remain the most regularly-used approach for implementing Bayesian inference. However, the computational cost of these approaches usually scales worse than linearly with the dimension of the parameter space, limiting their use for models with a large number of parameters. However, it is not uncommon for models to have identifiability problems. In this case, the likelihood is not informative about some subspaces of the parameter space, and hence the model effectively has a dimension that is lower than the number of parameters. Constantine et al. (2016) and Schuster et al. (2017) introduced the concept of directly using a Metropolis-Hastings (MH) MCMC on the subspaces of the parameter space that are informed by the likelihood as a means to reduce the dimension of the parameter space that needs to be explored. The same paper introduces an approach for identifying such a subspace in the case where it is a linear transformation of the parameter space: this subspace is known as the active subspace. This paper introduces sequential Monte Carlo (SMC) methods that make use of an active subspace. As well as an SMC counterpart to MH approach of Schuster et al. (2017), we introduce an approach to learn the active subspace adaptively, and an SMC$^{2}$ approach that is more robust to the linearity assumptions made when using active subspaces.
citation-role summary
citation-polarity summary
fields
stat.ML 1years
2025 1verdicts
CONDITIONAL 1roles
other 1polarities
unclear 1representative citing papers
citing papers explorer
-
Accelerating Hamiltonian Monte Carlo for Bayesian Inference in Neural Networks and Neural Operators
A hybrid VI-HMC method runs expensive Hamiltonian Monte Carlo only on the neural network parameters most responsible for predictive uncertainty, cutting sampling cost while approximating full HMC posteriors.