From complex to simple : hierarchical free-energy landscape renormalized in deep neural networks

Yoshino, Hajime

doi:10.21468/SciPostPhysCore.2.2.005

SciPost Physics Core

From complex to simple : hierarchical free-energy landscape renormalized in deep neural networks

Hajime Yoshino

SciPost Phys. Core 2, 005 (2020) · published 15 April 2020

doi: 10.21468/SciPostPhysCore.2.2.005
pdf
Submissions/Reports

Abstract

We develop a statistical mechanical approach based on the replica method to study the design space of deep and wide neural networks constrained to meet a large number of training data. Specifically, we analyze the configuration space of the synaptic weights and neurons in the hidden layers in a simple feed-forward perceptron network for two scenarios: a setting with random inputs/outputs and a teacher-student setting. By increasing the strength of constraints,~i.e. increasing the number of training data, successive 2nd order glass transition (random inputs/outputs) or 2nd order crystalline transition (teacher-student setting) take place layer-by-layer starting next to the inputs/outputs boundaries going deeper into the bulk with the thickness of the solid phase growing logarithmically with the data size. This implies the typical storage capacity of the network grows exponentially fast with the depth. In a deep enough network, the central part remains in the liquid phase. We argue that in systems of finite width N, the weak bias field can remain in the center and plays the role of a symmetry-breaking field that connects the opposite sides of the system. The successive glass transitions bring about a hierarchical free-energy landscape with ultrametricity, which evolves in space: it is most complex close to the boundaries but becomes renormalized into progressively simpler ones in deeper layers. These observations provide clues to understand why deep neural networks operate efficiently. Finally, we present some numerical simulations of learning which reveal spatially heterogeneous glassy dynamics truncated by a finite width $N$ effect.

TY  - JOUR
PB  - SciPost Foundation
DO  - 10.21468/SciPostPhysCore.2.2.005
TI  - From complex to simple : hierarchical free-energy landscape renormalized in deep neural networks
PY  - 2020/04/15
UR  - https://scipost.org/SciPostPhysCore.2.2.005
JF  - SciPost Physics Core
JA  - SciPost Phys. Core
VL  - 2
IS  - 2
SP  - 005
A1  - Yoshino, Hajime
AB  - We develop a statistical mechanical approach based on the replica method to study the design space of deep and wide neural networks constrained to meet a large number of training data. Specifically, we analyze the configuration space of the synaptic weights and neurons in the hidden layers in a simple feed-forward perceptron network for two scenarios: a setting with random inputs/outputs and a teacher-student setting. By increasing the strength of constraints,~i.e. increasing the number of training data, successive 2nd order glass transition (random inputs/outputs) or 2nd order crystalline transition (teacher-student setting) take place layer-by-layer starting next to the inputs/outputs boundaries going deeper into the bulk with the thickness of the solid phase growing logarithmically with the data size. This implies the typical storage capacity of the network grows exponentially fast with the depth. In a deep enough network, the central part remains in the liquid phase. We argue that in systems of finite width N, the weak bias field can remain in the center and plays the role of a symmetry-breaking field that connects the opposite sides of the system. The successive glass transitions bring about a hierarchical free-energy landscape with ultrametricity, which evolves in space: it is most complex close to the boundaries but becomes renormalized into progressively simpler ones in deeper layers. These observations provide clues to understand why deep neural networks operate efficiently. Finally, we present some numerical simulations of learning which reveal spatially heterogeneous glassy dynamics truncated by a finite width $N$ effect.
ER  -

@Article{10.21468/SciPostPhysCore.2.2.005,
	title={{From complex to simple : hierarchical free-energy landscape renormalized in deep neural networks}},
	author={Hajime Yoshino},
	journal={SciPost Phys. Core},
	volume={2},
	pages={005},
	year={2020},
	publisher={SciPost},
	doi={10.21468/SciPostPhysCore.2.2.005},
	url={https://scipost.org/10.21468/SciPostPhysCore.2.2.005},
}

Cited by 5

Ontology / Topics

See full Ontology or Topics database.

Neural networks

Author / Affiliation: mappings to Contributors and Organizations

See all Organizations.

¹ Hajime Yoshino

¹ 大阪大学 / Osaka University

Funder for the research work leading to this publication

文部科学省 Monbu-kagaku-shō / Ministry of Education, Culture, Sports, Science and Technology [MEXT]