Weight dynamics of learning networks
Nahal Sharafi1, Christoph Martin1, Sarah Hallerberg1
1Hamburg University of Applied Sciences, Berliner Tor 21, 20099 Hamburg, Germany.
Abstract:
Neural networks have become a widely adopted tool for tackling a variety of problems in machine learning and artificial intelligence. In this contribution, we use the mathematical framework of local stability analysis to gain a deeper understanding of the learning dynamics of feedforward neural networks. We derive equations for the tangent operator of the learning dynamics of three-layer networks learning regression tasks. The results are valid for an arbitrary number of nodes and arbitrary choices of activation functions. Applying the results to a network learning a regression task, we investigate numerically how stability indicators relate to the final training loss. Although the specific results vary with different choices of initial conditions and activation functions, we demonstrate that it is possible to predict the final training loss by monitoring finite-time Lyapunov exponents during the training process.
Related Concept Videos
Observational Learning
First Law: Particles in Two-dimensional Equilibrium
Newton's first law tells us about...
Weighted Mean
For example, consider the number of goals scored in the matches of a tournament. While computing the average number of goals scored in the tournament, it may be more important to...
Dynamic Equilibrium
Introduction to Learning
In contrast to learned behaviors, unlearned behaviors such as crying, sexual...
Rigid Body Equilibrium Problems - II
Consider two children sitting on a seesaw, which has negligible mass. The first child has a mass (m1) of 26 kg and sits at point A, which is 1.6 meters (r1) from the pivot point B; the second child has a mass (m2) of 32 kg and sits at point C. How far from the pivot point B should the second child sit (r2) to balance the seesaw?


