Nonsmooth Nonconvex Stochastic Heavy Ball

My bibliography Save this article

Nonsmooth Nonconvex Stochastic Heavy Ball

Author

Listed:

Tam Le
(University of Toulouse)

Registered:

Abstract

Motivated by the conspicuous use of momentum-based algorithms in deep learning, we study a nonsmooth nonconvex stochastic heavy ball method and show its convergence. Our approach builds upon semialgebraic (definable) assumptions commonly met in practical situations and combines a nonsmooth calculus with a differential inclusion method. Additionally, we provide general conditions for the sample distribution to ensure the convergence of the objective function. Our results are general enough to justify the use of subgradient sampling in modern implementations that heuristically apply rules of differential calculus on nonsmooth functions, such as backpropagation or implicit differentiation. As for the stochastic subgradient method, our analysis highlights that subgradient sampling can make the stochastic heavy ball method converge to artificial critical points. Thanks to the semialgebraic setting, we address this concern showing that these artifacts are almost surely avoided when initializations are randomized, leading the method to converge to Clarke critical points.

Suggested Citation

Tam Le, 2024. "Nonsmooth Nonconvex Stochastic Heavy Ball," Journal of Optimization Theory and Applications, Springer, vol. 201(2), pages 699-719, May.

Handle: RePEc:spr:joptap:v:201:y:2024:i:2:d:10.1007_s10957-024-02408-3
DOI: 10.1007/s10957-024-02408-3

Download full text from publisher

As the access to this document is restricted, you may want to

for a different version of it.

References listed on IDEAS

Michel Benaim & Josef Hofbauer & Sylvain Sorin, 2005. "Stochastic Approximations and Differential Inclusions II: Applications," Levine's Bibliography 784828000000000098, UCLA Department of Economics.
Le, Tam & Bolte, Jérôme & Pauwels, Edouard, 2022. "Subgradient sampling for nonsmooth nonconvex minimization," TSE Working Papers 22-1310, Toulouse School of Economics (TSE).

Full references (including those not matched with items on IDEAS)

Citations

Citations are extracted by the CitEc Project, subscribe to its RSS feed for this item.

Cited by:

Kuangyu Ding & Kim-Chuan Toh, 2025. "Stochastic Bregman Subgradient Methods for Nonsmooth Nonconvex Optimization Problems," Journal of Optimization Theory and Applications, Springer, vol. 206(3), pages 1-36, September.

Most related items

These are the items that most often cite the same works as this one and are cited by the same works as this one.

Giacomo Lanzani, 2025. "Dynamic Concern for Misspecification," Econometrica, Econometric Society, vol. 93(4), pages 1333-1370, July.
Benaïm, Michel & Hofbauer, Josef & Hopkins, Ed, 2009. "Learning in games with unstable equilibria," Journal of Economic Theory, Elsevier, vol. 144(4), pages 1694-1709, July.
- Ed Hopkins & Josef Hofbauer & Michel Benaim, 2005. "Learning in Games with Unstable Equilibria," Edinburgh School of Economics Discussion Paper Series 135, Edinburgh School of Economics, University of Edinburgh.
- Michel Benaim & Josef Hofbauer & Ed Hopkins, 2006. "Learning in Games with Unstable Equilibria," Levine's Bibliography 321307000000000547, UCLA Department of Economics.
- Michel Benaim & Josef Hofbauer & Ed Hopkins, 2005. "Learning in Games with Unstable Equilibria," Levine's Bibliography 784828000000000609, UCLA Department of Economics.
Dai Zusai, 2023. "Evolutionary dynamics in heterogeneous populations: a general framework for an arbitrary type distribution," International Journal of Game Theory, Springer;Game Theory Society, vol. 52(4), pages 1215-1260, December.
Akimoto, Youhei & Auger, Anne & Hansen, Nikolaus, 2022. "An ODE method to prove the geometric convergence of adaptive stochastic algorithms," Stochastic Processes and their Applications, Elsevier, vol. 145(C), pages 269-307.
Josef Hofbauer & Sylvain Sorin & Yannick Viossat, 2009. "Time Average Replicator and Best Reply Dynamics," Post-Print hal-00360767, HAL.
Bolte, Jérôme & Le, Tam & Pauwels, Edouard & Silveti-Falls, Antonio, 2022. "Nonsmooth Implicit Differentiation for Machine Learning and Optimization," TSE Working Papers 22-1314, Toulouse School of Economics (TSE).
- Bolte, Jérôme & Le, Tam & Pauwels, Edouard & Silveti-Falls, Antonio, 2022. "Nonsmooth Implicit Differentiation for Machine Learning and Optimization," TSE Working Papers 126768, Toulouse School of Economics (TSE).
Ziv Gorodeisky, 2008. "Stochastic Approximation of Discontinuous Dynamics," Discussion Paper Series dp496, The Federmann Center for the Study of Rationality, the Hebrew University, Jerusalem.
Michel Benaïm & Josef Hofbauer & Sylvain Sorin, 2006. "Stochastic Approximations and Differential Inclusions, Part II: Applications," Mathematics of Operations Research, INFORMS, vol. 31(4), pages 673-695, November.
- Michel Benaïm & Josef Hofbauer & Sylvain Sorin, 2005. "Stochastic Approximations and Differential Inclusions; Part II: Applications," Working Papers hal-00242974, HAL.
Vinayaka G. Yaji & Shalabh Bhatnagar, 2020. "Stochastic Recursive Inclusions in Two Timescales with Nonadditive Iterate-Dependent Markov Noise," Mathematics of Operations Research, INFORMS, vol. 45(4), pages 1405-1444, November.
Swenson, Brian & Murray, Ryan & Kar, Soummya, 2020. "Regular potential games," Games and Economic Behavior, Elsevier, vol. 124(C), pages 432-453.
Bervoets, Sebastian & Faure, Mathieu, 2020. "Convergence in games with continua of equilibria," Journal of Mathematical Economics, Elsevier, vol. 90(C), pages 25-30.
- Sebastian Bervoets & Mathieu Faure, 2020. "Convergence in games with continua of equilibria," Post-Print hal-02964989, HAL.
Bervoets, Sebastian & Faure, Mathieu, 2019. "Stability in games with continua of equilibria," Journal of Economic Theory, Elsevier, vol. 179(C), pages 131-162.
- Sebastian Bervoets & Mathieu Faure, 2019. "Stability in games with continua of equilibria," Post-Print hal-02021221, HAL.
Andrés Contreras & Juan Peypouquet, 2019. "Asymptotic Equivalence of Evolution Equations Governed by Cocoercive Operators and Their Forward Discretizations," Journal of Optimization Theory and Applications, Springer, vol. 182(1), pages 30-48, July.
Michel Benaïm & Mathieu Faure, 2013. "Consistency of Vanishingly Smooth Fictitious Play," Mathematics of Operations Research, INFORMS, vol. 38(3), pages 437-450, August.
- Michel Benaïm & Mathieu Faure, 2013. "Consistency of Vanishingly Smooth Fictitious Play," Post-Print hal-01498243, HAL.
Michel Benaim & Olivier Raimond, 2007. "Simulated Annealing, Vertex-Reinforced Random Walks and Learning in Games," Levine's Bibliography 122247000000001702, UCLA Department of Economics.
Josef Hofbauer & Sylvain Sorin & Yannick Viossat, 2009. "Time Average Replicator and Best-Reply Dynamics," Mathematics of Operations Research, INFORMS, vol. 34(2), pages 263-269, May.
- Josef Hofbauer & Sylvain Sorin & Yannick Viossat, 2009. "Time Average Replicator and Best Reply Dynamics," Post-Print hal-00360767, HAL.
Esponda, Ignacio & Pouzo, Demian & Yamamoto, Yuichi, 2021. "Asymptotic behavior of Bayesian learners with misspecified models," Journal of Economic Theory, Elsevier, vol. 195(C).
- Ignacio Esponda & Demian Pouzo & Yuichi Yamamoto, 2019. "Asymptotic Behavior of Bayesian Learners with Misspecified Models," Papers 1904.08551, arXiv.org, revised Oct 2019.
Bolte, Jérôme & Pauwels, Edouard, 2021. "A mathematical model for automatic differentiation in machine learning," TSE Working Papers 21-1184, Toulouse School of Economics (TSE).
Arunselvan Ramaswamy & Shalabh Bhatnagar, 2017. "A Generalization of the Borkar-Meyn Theorem for Stochastic Recursive Inclusions," Mathematics of Operations Research, INFORMS, vol. 42(3), pages 648-661, August.
Candogan, Ozan & Ozdaglar, Asuman & Parrilo, Pablo A., 2013. "Dynamics in near-potential games," Games and Economic Behavior, Elsevier, vol. 82(C), pages 66-90.

More about this item

Keywords

; ; ; ; ;

Statistics

Access and download statistics

Corrections

All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:spr:joptap:v:201:y:2024:i:2:d:10.1007_s10957-024-02408-3. See general information about how to correct material in RePEc.

If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Sonal Shukla or Springer Nature Abstracting and Indexing (email available below). General contact details of provider: http://www.springer.com .

Please note that corrections may take a couple of weeks to filter through the various RePEc services.

IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.

Browse Econ Literature

More features

Nonsmooth Nonconvex Stochastic Heavy Ball

Author

Abstract

Suggested Citation

Download full text from publisher

References listed on IDEAS

Citations

Most related items

More about this item

Keywords

Statistics

Corrections

More services and features

MyIDEAS

Author registration

Rankings

RePEc Genealogy

RePEc Biblio

MPRA

New papers by email

EconAcademics

Plagiarism

About RePEc

RePEc home

Blog

Help/FAQ

RePEc team

Participating archives

Privacy statement

Help us

Corrections

Volunteers

Get papers listed

Open a RePEc archive

Get RePEc data