论文信息 - On Random Matrices Arising in Deep Neural Networks. Gaussian Case

On Random Matrices Arising in Deep Neural Networks. Gaussian Case

The paper deals with distribution of singular values of product of random matrices arising in the analysis of deep neural networks. The matrices resemble the product analogs of the sample covariance matrices, however, an important difference is that the population covariance matrices, which are assumed to be non-random in the standard setting of statistics and random matrix theory, are now random, moreover, are certain functions of random data matrices. The problem has been considered in recent work [21] by using the techniques of free probability theory. Since, however, free probability theory deals with population matrices which are independent of the data matrices, its applicability in this case requires an additional justification. We present this justification by using a version of the standard techniques of random matrix theory under the assumption that the entries of data matrices are independent Gaussian random variables. In the subsequent paper [18] we extend our results to the case where the entries of data matrices are just independent identically distributed random variables with several finite moments. This, in particular, extends the property of the so-called macroscopic universality on the considered random matrices.

L. Pastur

[1] Ausif Mahmood,et al. Review of Deep Learning Algorithms and Architectures , 2019, IEEE Access.

[2] Robert C. Qiu,et al. Spectrum Concentration in Deep Residual Learning: A Free Probability Approach , 2018, IEEE Access.

[3] Roman Vershynin,et al. High-Dimensional Probability , 2018 .

[4] Dong Eui Chang,et al. Deep Neural Networks in a Mathematical Framework , 2018, SpringerBriefs in Computer Science.

[5] Surya Ganguli,et al. The Emergence of Spectral Universality in Deep Networks , 2018, AISTATS.

[6] Richard E. Turner,et al. Gaussian Process Behaviour in Wide Deep Neural Networks , 2018, ICLR.

[7] A. Chakrabarty,et al. A note on the folklore of free independence , 2018, 1802.00952.

[8] Michael W. Mahoney,et al. Rethinking generalization requires revisiting old ideas: statistical mechanics approaches and complex learning behavior , 2017, ArXiv.

[9] Jeffrey Pennington,et al. Geometry of Neural Network Loss Surfaces via Random Matrix Theory , 2017, ICML.

[10] Dianhui Wang,et al. Randomness in neural networks: an overview , 2017, WIREs Data Mining Knowl. Discov..

[11] Surya Ganguli,et al. Deep Information Propagation , 2016, ICLR.