Related Experiment Video
Updated: Dec 12, 2025

Deep Neural Networks for Image-Based Dietary Assessment
Published on: March 13, 2021
Tangent-space gradient optimization of tensor network for machine learning
Zheng-Zhi Sun1, Shi-Ju Ran2, Gang Su1,3
1School of Physical Sciences, University of Chinese Academy of Sciences, P.O. Box 4588, Beijing 100049, China.
Abstract:
The gradient-based optimization method for deep machine learning models suffers from gradient vanishing and exploding problems, particularly when the computational graph becomes deep. In this work, we propose the tangent-space gradient optimization (TSGO) for probabilistic models to keep the gradients from vanishing or exploding. The central idea is to guarantee the orthogonality between variational parameters and gradients. The optimization is then implemented by rotating the parameter vector towards the direction of gradient. We explain and test TSGO in tensor network (TN) machine learning, where TN describes the joint probability distribution as a normalized state |ψ〉 in Hilbert space. We show that the gradient can be restricted in tangent space of 〈ψ|ψ〉=1 hypersphere. Instead of additional adaptive methods to control the learning rate η in deep learning, the learning rate of TSGO is naturally determined by rotation angle θ as η=tanθ. Our numerical results reveal better convergence of TSGO in comparison to the off-the-shelf Adam.
Related Concept Videos
Gradient and Del Operator
Second Derivatives and Laplace Operator
Consider a scalar function. The curl of its...
Tangent to a Curve
Curvilinear Motion: Normal and Tangential Components
The positive direction of the t-axis aligns with the increasing position of the car along the curved path, denoted by the unit vector ut. Simultaneously, the n-axis, perpendicular to the t-axis, dissects the curved path into differential arc segments, each forming the arc of a circle with a radius of...
Gauss's Law: Problem-Solving
Poisson's And Laplace's Equation