Skip to content

Opening book details…

Can I read SGD on EtoBox?

SGD by Andy is a document available to read on EtoBox.

What is SGD about?

This document discusses the effectiveness of Stochastic Gradient Descent (SGD) in escaping local minima during the training of neural networks. It presents an alternative perspective that SGD operates on a smoothed version of the loss function, which allows it to avoid sharp local minima and converge to better solutions. The authors provide theoretical insights and empirical observations supporting their claims, highlighting the role of gradient noise and step size in the optimization process.

Author
Andy
Language
EN