Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Activated Gradients for Deep Neural Networks

Jul 09, 2021

Mei Liu, Liangming Chen, Xiaohao Du, Long Jin, Mingsheng Shang

Figure 1 for Activated Gradients for Deep Neural Networks

Figure 2 for Activated Gradients for Deep Neural Networks

Figure 3 for Activated Gradients for Deep Neural Networks

Figure 4 for Activated Gradients for Deep Neural Networks

Share this with someone who'll enjoy it:

Abstract:Deep neural networks often suffer from poor performance or even training failure due to the ill-conditioned problem, the vanishing/exploding gradient problem, and the saddle point problem. In this paper, a novel method by acting the gradient activation function (GAF) on the gradient is proposed to handle these challenges. Intuitively, the GAF enlarges the tiny gradients and restricts the large gradient. Theoretically, this paper gives conditions that the GAF needs to meet, and on this basis, proves that the GAF alleviates the problems mentioned above. In addition, this paper proves that the convergence rate of SGD with the GAF is faster than that without the GAF under some assumptions. Furthermore, experiments on CIFAR, ImageNet, and PASCAL visual object classes confirm the GAF's effectiveness. The experimental results also demonstrate that the proposed method is able to be adopted in various deep neural networks to improve their performance. The source code is publicly available at https://github.com/LongJin-lab/Activated-Gradients-for-Deep-Neural-Networks.

View paper on

Share this with someone who'll enjoy it:

Title:Activated Gradients for Deep Neural Networks

Paper and Code