NobleBlocks
Public

Implicit Gradient Regularization

Published in arXiv (Cornell University) • Sep 23, 2020
Authors:
David G. T. Barrett
,
Benoît Dherin

Abstract

Gradient descent can be surprisingly good at optimizing deep neural networks without overfitting and without explicit regularization. We find that the discrete steps of gradient descent implicitly regularize models by penalizing gradient descent trajectories that have large loss gradients. We call t...

Finding related papers...

Discussions

(0)

No comments yet

Be the first to share your thoughts!