Pith. sign in

REVIEW 2 cited by

Bayesian Low-Rank LeArning (Bella): A Practical Approach to Bayesian Neural Networks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2407.20891 v5 pith:VXNTHV2W submitted 2024-07-30 cs.LG cs.AIcs.CV

classification cs.LGcs.AIcs.CV
keywords bayesianlearningbellapracticallow-rankneuralapproachcomputational
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Computational complexity of Bayesian learning is impeding its adoption in practical, large-scale tasks. Despite demonstrations of significant merits such as improved robustness and resilience to unseen or out-of-distribution inputs over their non- Bayesian counterparts, their practical use has faded to near insignificance. In this study, we introduce an innovative framework to mitigate the computational burden of Bayesian neural networks (BNNs). Our approach follows the principle of Bayesian techniques based on deep ensembles, but significantly reduces their cost via multiple low-rank perturbations of parameters arising from a pre-trained neural network. Both vanilla version of ensembles as well as more sophisticated schemes such as Bayesian learning with Stein Variational Gradient Descent (SVGD), previously deemed impractical for large models, can be seamlessly implemented within the proposed framework, called Bayesian Low-Rank LeArning (Bella). In a nutshell, i) Bella achieves a dramatic reduction in the number of trainable parameters required to approximate a Bayesian posterior; and ii) it not only maintains, but in some instances, surpasses the performance of conventional Bayesian learning methods and non-Bayesian baselines. Our results with large-scale tasks such as ImageNet, CAMELYON17, DomainNet, VQA with CLIP, LLaVA demonstrate the effectiveness and versatility of Bella in building highly scalable and practical Bayesian deep models for real-world applications.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Stochastic Weight Sharing for Bayesian Neural Networks

    cs.LG 2025-05 conditional novelty 6.0 of 10

    2DGBNN compresses Bayesian neural networks by clustering weight means and variances into shared 2D Gaussians, reducing parameter counts by up to 99% on ImageNet-scale models with small accuracy losses.

  2. Enhancing Monte Carlo Dropout Performance for Uncertainty Quantification

    cs.CV 2025-05 reject novelty 3.0 of 10

    Tuning Monte Carlo Dropout hyperparameters with GWO, BO, or PSO and adding a predictive-entropy loss term reportedly improves accuracy, uncertainty accuracy, and calibration by 2-3% over vanilla MCD.

Pith tools