Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Michael Westdickenberg

Learning deep linear neural networks: Riemannian gradient flows and convergence to global minimizers

Oct 12, 2019

Bubacarr Bah, Holger Rauhut, Ulrich Terstiege, Michael Westdickenberg

Figure 1 for Learning deep linear neural networks: Riemannian gradient flows and convergence to global minimizers

Figure 2 for Learning deep linear neural networks: Riemannian gradient flows and convergence to global minimizers

Figure 3 for Learning deep linear neural networks: Riemannian gradient flows and convergence to global minimizers

Figure 4 for Learning deep linear neural networks: Riemannian gradient flows and convergence to global minimizers

Abstract:We study the convergence of gradient flows related to learning deep linear neural networks from data (i.e., the activation function is the identity map). In this case, the composition of the network layers amounts to simply multiplying the weight matrices of all layers together, resulting in an overparameterized problem. We show that the gradient flow with respect to these factors can be re-interpreted as a Riemannian gradient flow on the manifold of rank-$r$ matrices endowed with a suitable Riemannian metric. We show that the flow always converges to a critical point of the underlying functional. Moreover, in the special case of an autoencoder, we show that the flow converges to a global minimum for almost all initializations.

* 21 pages

Via

Access Paper or Ask Questions