Statistical Mechanics of Deep Learning

Bahri, Yasaman; Kadmon, Jonathan; Pennington, Jeffrey; Schoenholz, Sam S.; Sohl-Dickstein, Jascha; Ganguli, Surya

doi:10.1146/annurev-conmatphys-031119-050745

Citation Details

Statistical Mechanics of Deep Learning

The recent striking success of deep neural networks in machine learning raises profound questions about the theoretical principles underlying their success. For example, what can such deep networks compute? How can we train them? How does information propagate through them? Why can they generalize? And how can we teach them to imagine? We review recent work in which methods of physical analysis rooted in statistical mechanics have begun to provide conceptual insights into these questions. These insights yield connections between deep learning and diverse physical and mathematical topics, including random landscapes, spin glasses, jamming, dynamical phase transitions, chaos, Riemannian geometry, random matrix theory, free probability, and nonequilibrium statistical mechanics. Indeed, the fields of statistical mechanics and machine learning have long enjoyed a rich history of strongly coupled interactions, and recent advances at the intersection of statistical mechanics and deep learning suggest these interactions will only deepen going forward. more »

Award ID(s):: 1845166

PAR ID:: 10291285

Author(s) / Creator(s):: Bahri, Yasaman; Kadmon, Jonathan; Pennington, Jeffrey; Schoenholz, Sam S.; Sohl-Dickstein, Jascha; Ganguli, Surya

Date Published:: 2020-03-10

Journal Name:: Annual Review of Condensed Matter Physics

Volume:: 11

Issue:: 1

ISSN:: 1947-5454

Page Range / eLocation ID:: 501 to 528

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Journal Article:
https://doi.org/10.1146/annurev-conmatphys-031119-050745

More Like this