arXiv:2606. 06722v1 Announce Type: new Abstract: The training of neural networks often entails objective functions that are not globally $L$-smooth.
Paper
Flatland: The Adventures of Gradient Descent with Large Step Sizes
Unreadunread
Paper
Unreadunread
arXiv:2606. 06722v1 Announce Type: new Abstract: The training of neural networks often entails objective functions that are not globally $L$-smooth.