Computer Science > Machine Learning

arXiv:2210.12957 (cs)

[Submitted on 24 Oct 2022]

Title:On the optimization and pruning for Bayesian deep learning

View PDF

Abstract:The goal of Bayesian deep learning is to provide uncertainty quantification via the posterior distribution. However, exact inference over the weight space is computationally intractable due to the ultra-high dimensions of the neural network. Variational inference (VI) is a promising approach, but naive application on weight space does not scale well and often underperform on predictive accuracy. In this paper, we propose a new adaptive variational Bayesian algorithm to train neural networks on weight space that achieves high predictive accuracy. By showing that there is an equivalence to Stochastic Gradient Hamiltonian Monte Carlo(SGHMC) with preconditioning matrix, we then propose an MCMC within EM algorithm, which incorporates the spike-and-slab prior to capture the sparsity of the neural network. The EM-MCMC algorithm allows us to perform optimization and model pruning within one-shot. We evaluate our methods on CIFAR-10, CIFAR-100 and ImageNet datasets, and demonstrate that our dense model can reach the state-of-the-art performance and our sparse model perform very well compared to previously proposed pruning schemes.

Comments:	11 pages
Subjects:	Machine Learning (cs.LG); Computation (stat.CO)
Cite as:	arXiv:2210.12957 [cs.LG]
	(or arXiv:2210.12957v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2210.12957

Submission history

From: Xiongwen Ke [view email]
[v1] Mon, 24 Oct 2022 05:18:08 UTC (515 KB)

Computer Science > Machine Learning

Title:On the optimization and pruning for Bayesian deep learning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:On the optimization and pruning for Bayesian deep learning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators