Learning Expressive Priors for Generalization and Uncertainty Estimation in Neural Networks

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Schnaus, Dominik, Lee, Jongseok, Cremers, Daniel, Triebel, Rudolph
Format: Preprint
Published: 2023
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866911551580012544
author Schnaus, Dominik
Lee, Jongseok
Cremers, Daniel
Triebel, Rudolph
author_facet Schnaus, Dominik
Lee, Jongseok
Cremers, Daniel
Triebel, Rudolph
contents In this work, we propose a novel prior learning method for advancing generalization and uncertainty estimation in deep neural networks. The key idea is to exploit scalable and structured posteriors of neural networks as informative priors with generalization guarantees. Our learned priors provide expressive probabilistic representations at large scale, like Bayesian counterparts of pre-trained models on ImageNet, and further produce non-vacuous generalization bounds. We also extend this idea to a continual learning framework, where the favorable properties of our priors are desirable. Major enablers are our technical contributions: (1) the sums-of-Kronecker-product computations, and (2) the derivations and optimizations of tractable objectives that lead to improved generalization bounds. Empirically, we exhaustively show the effectiveness of this method for uncertainty estimation and generalization.
format Preprint
id arxiv_https___arxiv_org_abs_2307_07753
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle Learning Expressive Priors for Generalization and Uncertainty Estimation in Neural Networks
Schnaus, Dominik
Lee, Jongseok
Cremers, Daniel
Triebel, Rudolph
Machine Learning
Artificial Intelligence
In this work, we propose a novel prior learning method for advancing generalization and uncertainty estimation in deep neural networks. The key idea is to exploit scalable and structured posteriors of neural networks as informative priors with generalization guarantees. Our learned priors provide expressive probabilistic representations at large scale, like Bayesian counterparts of pre-trained models on ImageNet, and further produce non-vacuous generalization bounds. We also extend this idea to a continual learning framework, where the favorable properties of our priors are desirable. Major enablers are our technical contributions: (1) the sums-of-Kronecker-product computations, and (2) the derivations and optimizations of tractable objectives that lead to improved generalization bounds. Empirically, we exhaustively show the effectiveness of this method for uncertainty estimation and generalization.
title Learning Expressive Priors for Generalization and Uncertainty Estimation in Neural Networks
topic Machine Learning
Artificial Intelligence
url https://arxiv.org/abs/2307.07753