REVIEW 2 cited by
Bayesian Neural Network Priors Revisited
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Isotropic Gaussian priors are the de facto standard for modern Bayesian neural network inference. However, it is unclear whether these priors accurately reflect our true beliefs about the weight distributions or give optimal performance. To find better priors, we study summary statistics of neural network weights in networks trained using stochastic gradient descent (SGD). We find that convolutional neural network (CNN) and ResNet weights display strong spatial correlations, while fully connected networks (FCNNs) display heavy-tailed weight distributions. We show that building these observations into priors can lead to improved performance on a variety of image classification datasets. Surprisingly, these priors mitigate the cold posterior effect in FCNNs, but slightly increase the cold posterior effect in ResNets.
Forward citations
Cited by 2 Pith papers
-
Rethinking Likelihood distributions: Student's t Likelihood Boosts Bayesian Neural Network Performance
Student's t likelihood (ν=5) is a robust default for VI-trained BNNs, improving CRPS in most tested settings while occasionally losing on MSE to Gaussian under lognormal noise.
-
ALAS: Additive Learnable Alpha-Stable Kernels for Flexible Bayesian Optimization
ALAS learns the spectral tail exponent of a GP kernel, adapting from Gaussian to heavy-tailed behavior, with a per-dimension additive variant for high-dimensional Bayesian optimization.
Discussion (0). Continue with ORCID to comment.