论文信息 - Bayesian model selection for high-dimensional data

Bayesian model selection for high-dimensional data

Abstract High-dimensional data, where the number of features or covariates can even be larger than the number of independent samples, are ubiquitous and are encountered on a regular basis by statistical scientists both in academia and in industry. A majority of the classical research in statistics dealt with the settings where there is a small number of covariates. Due to the modern advancements in data storage and computational power, the high-dimensional data revolution has significantly occupied mainstream statistical research. In gene expression datasets, for instance, it is not uncommon to encounter datasets with observations on at most a few hundred independent samples (subjects) and with information on tens or hundreds of thousands of genes per each sample. An important and common question that arises quickly is—“which of the available covariates are relevant to the outcome of interest?” This concerns the problem of variable selection (and more generally model selection) in statistics and data science. This chapter will provide an overview of some of the most well-known model selection methods along with some of the more recent methods. While frequentist methods will be discussed, Bayesian approaches will be given a more elaborate treatment. The frequentist framework for model selection is primarily based on penalization, whereas the Bayesian framework relies on prior distributions for inducing shrinkage and sparsity. The chapter treats the Bayesian framework in the light of objective and empirical Bayesian viewpoints as the priors in the high-dimensional setting are typically not completely based subjective prior beliefs. An important practical aspect of high-dimensional model selection methods is computational scalability which will also be discussed.

Naveennaidu Narisetty | N. Narisetty

[1] A. V. D. Vaart,et al. BAYESIAN LINEAR REGRESSION WITH SPARSE PRIORS , 2014, 1403.0735.

[2] Martin J. Wainwright,et al. On the Computational Complexity of High-Dimensional Bayesian Variable Selection , 2015, ArXiv.

[3] E. George,et al. Journal of the American Statistical Association is currently published by American Statistical Association. , 2007 .

[4] Zoubin Ghahramani,et al. Deep Bayesian Active Learning with Image Data , 2017, ICML.

[5] H. Zou. The Adaptive Lasso and Its Oracle Properties , 2006 .

[6] V. Rocková,et al. Bayesian estimation of sparse signals with a continuous spike-and-slab prior , 2018 .

[7] Lan Wang,et al. Quantile-adaptive model-free variable screening for high-dimensional heterogeneous data , 2013, 1304.2186.

[8] G. Casella,et al. Consistency of Bayesian procedures for variable selection , 2009, 0904.2978.

[9] A. V. D. Vaart,et al. Convergence rates of posterior distributions , 2000 .

[10] D. Rubin,et al. Maximum likelihood from incomplete data via the EM - algorithm plus discussions on the paper , 1977 .

[11] E. George,et al. APPROACHES FOR BAYESIAN VARIABLE SELECTION , 1997 .