论文信息 - Empirical Frequentist Coverage of Deep Learning Uncertainty Quantification Procedures

Empirical Frequentist Coverage of Deep Learning Uncertainty Quantification Procedures

Uncertainty quantification for complex deep learning models is increasingly important as these techniques see growing use in high-stakes, real-world settings. Currently, the quality of a model's uncertainty is evaluated using point-prediction metrics such as negative log-likelihood or the Brier score on heldout data. In this study, we provide the first large scale evaluation of the empirical frequentist coverage properties of well known uncertainty quantification techniques on a suite of regression and classification tasks. We find that, in general, some methods do achieve desirable coverage properties on in distribution samples, but that coverage is not maintained on out-of-distribution data. Our results demonstrate the failings of current uncertainty quantification techniques as dataset shift increases and establish coverage as an important metric in developing models for real-world applications.

Jasper Snoek | Benjamin Kompa | Andrew Beam

[1] Sebastian Nowozin,et al. How Good is the Bayes Posterior in Deep Neural Networks Really? , 2020, ICML.

[2] Dustin Tran,et al. Flipout: Efficient Pseudo-Independent Weight Perturbations on Mini-Batches , 2018, ICLR.

[3] Larry Wasserman,et al. All of Statistics: A Concise Course in Statistical Inference , 2004 .

[4] Dustin Tran,et al. Simple and Principled Uncertainty Estimation with Deterministic Deep Learning via Distance Awareness , 2020, NeurIPS.

[5] Geoffrey E. Hinton,et al. Bayesian Learning for Neural Networks , 1995 .

[6] Ben Glocker,et al. Implicit Weight Uncertainty in Neural Networks. , 2017 .

[7] Andrew Gordon Wilson,et al. A Simple Baseline for Bayesian Uncertainty in Deep Learning , 2019, NeurIPS.

[8] Kevin Gimpel,et al. A Baseline for Detecting Misclassified and Out-of-Distribution Examples in Neural Networks , 2016, ICLR.

[9] Richard E. Turner,et al. Black-box α-divergence minimization , 2016, ICML 2016.

[10] Van Der Vaart,et al. The Bernstein-Von-Mises theorem under misspecification , 2012 .

[11] Alex Graves,et al. Practical Variational Inference for Neural Networks , 2011, NIPS.