AdaSAM: Boosting sharpness-aware minimization with adaptive learning rate and momentum for training deep neural

Hao Sun1, Li Shen2, Qihuang Zhong3

  • 1School of Computer Science, University of Science and Technology of China, Hefei, 230026, Anhui, China.

Summary

Sharpness Aware Minimization (SAM) with adaptive learning rate and momentum (AdaSAM) offers better generalization for deep learning. This study theoretically proves AdaSAM

Related Concept Videos