arXiv Science⌕ Search

arXiv · 2610.09364

Nonmonotone Coderivative-Based Newton Methods for Nonsmooth Optimization with Machine Learning Applications

Abstract

Newton-type algorithms are among the most effective methods for solving optimization problems because of their rapid local convergence. However, extending these methods to nonsmooth optimization is challenging due to the difficulty of incorporating second-order information. To address this issue, we propose a coderivative-based Newton framework for solving both unconstrained and constrained nonsmooth optimization problems using tools from variational analysis and generalized differentiation. The proposed method employs generalized Hessians, defined as coderivatives of the subgradient mapping, and is applicable to both $C ^{1,1}$ functions and convex composite optimization problems with extended-real-valued components. To enhance robustness, the algorithm incorporates an adaptive regularization strategy that enables it to handle problems whose generalized Hessians are positive semidefinite. In addition, a Hybrid Adaptive Nonmonotone (HAN) line search scheme is developed as a globalization technique to improve global convergence and practical performance. Under standard assumptions, both the exact and inexact versions of the algorithm are globally convergent and achieve local superlinear convergence when the associated subgradient mapping satisfies the semismooth* property. The proposed framework is further extended to convex composite optimization through the forward-backward envelope, allowing it to handle problems with or without strong convexity in the smooth component of the objective function. Numerical experiments on Lasso, logistic Lasso, and support vector machine (SVM) problems demonstrate the efficiency and robustness of the proposed algorithm. Comparisons with several well-established first-order and second-order methods for nonsmooth optimization show that the proposed approach is computationally competitive while maintaining strong convergence properties.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Sina Kazemdehbashi, Yanchao Liu, Boris S. Mordukhovich. 2026-10-07. Nonmonotone Coderivative-Based Newton Methods for Nonsmooth Optimization with Machine Learning Applications. https://arxiv.org/abs/2610.09364

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Constrained portfolio game with heterogeneous agents

We investigate stochastic utility maximization games under relative performance concerns in both finite-agent and infinite-agent (graphon) settings. An incomplete market model is considered where agents with power (CRRA) utility functions trade in a common risk-free bond and individual stocks driven by both common and idiosyncratic noise. The Nash equilibrium for both settings is characterized by forward-backward stochastic differential equations (FBSDEs) with a quadratic growth generator, where the solution of the graphon game leads to a novel form of infinite-dimensional McKean-Vlasov FBSDEs. Under mild conditions, we prove the existence of Nash equilibrium for both the graphon game and the $n$-agent game without common noise. Furthermore, we establish a convergence result showing that, with modest assumptions on the sensitivity matrix, as the number of agents increases, the Nash equilibrium and associated equilibrium value of the finite-agent game converge to those of the graphon game.

math.OC↗

A Model-Based Derivative-Free Optimization Algorithm for Partially Separable Problems

We propose UPOQA, a derivative-free optimization algorithm for partially separable unconstrained problems, leveraging quadratic interpolation and a structured trust-region framework. By decomposing the objective into element functions, UPOQA constructs underdetermined element models and solves subproblems efficiently via a modified projected gradient method. Innovations include an approximate projection operator for structured trust regions, improved management of elemental radii and models, a starting point search mechanism, and support for hybrid black-white-box optimization, etc. Numerical experiments on 85 CUTEst problems demonstrate that \texttt{UPOQA} can significantly reduce the number of function evaluations. To quantify the impact of exploiting partial separability, we introduce the speed-up profile to further evaluate the acceleration effect. Results show that the speed-up of UPOQA over baselines is less significant in low-precision scenarios but becomes more pronounced in high-precision scenarios. Applications to quantum variational problems further validate its practical utility.

math.OC↗

Convergence Analysis of Noisy Distributed Gradient Descent for Non-convex Optimization -- Saddle Point Escape

This paper studies noisy distributed gradient descent (\textbf{NDGD}) for smooth non-convex finite-sum optimization over networks. Random perturbations enable saddle-point escape while preserving distributed implementation and consensus. Under suitable regularity conditions, \textbf{NDGD} converges with high probability to a neighborhood of a common local minimizer. Its convergence complexity is comparable to centralized first-order saddle-point escape methods, reducing exponential dependence on problem dimension to polynomial dependence. Numerical experiments demonstrate improved saddle-point escape over standard \textbf{DGD}.

math.OC↗