Machine Learning · Statistics
The Implicit Bias of Gradient Descent on Separable Data
Daniel Soudry, Elad Hoffer, Mor Shpigel Nacson, Suriya Gunasekar +1
2024-10-29
Machine Learning · Statistics
The Marginal Value of Adaptive Gradient Methods in Machine Learning
Ashia C. Wilson, Rebecca Roelofs, Mitchell Stern, Nathan Srebro +1
2018-05-23
Machine Learning · Computer Science
Towards The Implicit Bias on Multiclass Separable Data Under Norm Constraints
Shengping Xie, Zekun Wu, Quan Chen, Kaixu Tang
2026-03-25
Machine Learning · Computer Science
Private Adaptive Gradient Methods for Convex Optimization
Hilal Asi, John Duchi, Alireza Fallah, Omid Javidbakht +1
2021-06-28
Machine Learning · Computer Science
Adaptivity without Compromise: A Momentumized, Adaptive, Dual Averaged Gradient Method for Stochastic Optimization
Aaron Defazio, Samy Jelassi
2021-08-27
Machine Learning · Statistics
Implicit differentiation of Lasso-type models for hyperparameter optimization
Quentin Bertrand, Quentin Klopfenstein, Mathieu Blondel, Samuel Vaiter +2
2020-09-04
Machine Learning · Statistics
Scalable Adaptive Stochastic Optimization Using Random Projections
Gabriel Krummenacher, Brian McWilliams, Yannic Kilcher, Joachim M. Buhmann +1
2016-11-22
Machine Learning · Computer Science
Domain-independent Dominance of Adaptive Methods
Pedro Savarese, David McAllester, Sudarshan Babu, Michael Maire
2020-03-18
Machine Learning · Computer Science
When Will Gradient Methods Converge to Max-margin Classifier under ReLU Models?
Tengyu Xu, Yi Zhou, Kaiyi Ji, Yingbin Liang
2018-10-17
Machine Learning · Computer Science
On the Convergence of AdaGrad(Norm) on $\R^{d}$: Beyond Convexity, Non-Asymptotic Rate and Acceleration
Zijian Liu, Ta Duy Nguyen, Alina Ene, Huy L. Nguyen
2023-10-05
Machine Learning · Statistics
Implicit differentiation for fast hyperparameter selection in non-smooth convex learning
Quentin Bertrand, Quentin Klopfenstein, Mathurin Massias, Mathieu Blondel +3
2022-08-10
Machine Learning · Computer Science
Scalable Bayesian Meta-Learning through Generalized Implicit Gradients
Yilang Zhang, Bingcong Li, Shijian Gao, Georgios B. Giannakis
2023-12-22
Machine Learning · Statistics
Non-asymptotic Analysis of Biased Adaptive Stochastic Approximation
Sobihan Surendran, Antoine Godichon-Baggioni, Adeline Fermanian, Sylvain Le Corff
2025-03-17