Skip to main navigation Skip to search Skip to main content

Nonconvex Nonsmooth Multicomposite Optimization and Its Applications to Recurrent Neural Networks

Research output: Journal article publicationJournal articleAcademic researchpeer-review

Abstract

We consider a class of nonconvex nonsmooth multicomposite optimization problems where the objective function consists of a Tikhonov regularizer and a composition of multiple nonconvex nonsmooth component functions. Such optimization problems arise from tangible applications in machine learning and beyond. To define and compute its first-order and second-order
d(irectional)-stationary points effectively, we first derive the closed-form expression of the tangent cone for the feasible region of its constrained reformulation. Building on this, we establish its equivalence with the corresponding constrained and l1-penalty reformulations in terms of global optimality and d-stationarity. The equivalence offers indirect methods to attain the first-order and second-order d-stationary points of the original problem in certain cases. We apply our results to the training process of recurrent neural networks (RNNs).
Original languageEnglish
Pages (from-to)2343-2371
JournalSIAM Journal on Optimization
Volume35
Issue number4
DOIs
Publication statusPublished - 22 Oct 2025

Keywords

  • multicomposite optimization
  • tangent cone
  • first-order d-stationarity
  • second-order d-stationarity
  • recurrent neural network

Fingerprint

Dive into the research topics of 'Nonconvex Nonsmooth Multicomposite Optimization and Its Applications to Recurrent Neural Networks'. Together they form a unique fingerprint.

Cite this