Efforts to understand the generalization mystery in deep learning have led to the belief that gradient-based optimization induces a form of implicit regularization, a bias towards models of low "complexity." We study the implicit regularization of gradient descent over deep linear neural networks fo...
Research Assistant
AI chat, annotations, notes & similar papers
No comments yet
Be the first to share your thoughts!