PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 2, 20260 citations

From Convergence to Generalization: Stability of Stationary-Point Learning Algorithms.

View Full Paper
YLYunwen LeiZWZimeng WangXYXiaoming Yuan

Key Points

  • This research aims to explore the stability and generalization properties of learning algorithms beyond the constraints of convexity.
  • Established bounds on stability and generalization for various algorithms in nonconvex settings under a differentiability assumption.
  • Introduced an algorithm-dependent quantity based on the training dataset and the algorithm's output.
  • Applied stability analyses to gradient descent, linear models, and shallow neural networks.
  • Demonstrated new stability and generalization bounds that incorporate optimization error and the algorithm-dependent quantity.
  • Findings are valid even when algorithms do not converge to global or local minimizers.
  • Empirical studies confirmed the effectiveness of these stability analyses.

Abstract

Algorithmic stability is a fundamental concept in learning theory for studying the generalization guarantees of learning algorithms. A notable limitation of classical stability analyses is that they often require convexity assumptions to obtain nontrivial bounds. In this paper, we investigate the stability and generalization properties of learning algorithms in nonconvex settings. We introduce an algorithm-dependent quantity that depends only on the training dataset and the algorithm's output. Under a mild differentiability assumption, we establish stability and generalization bounds that apply to almost any algorithm. Our bounds explicitly involve the optimization error and the algorithm-dependent quantity, thereby capturing the local curvature of the objective function around the learned model. A key feature of our analysis is that it remains valid even when the algorithm does not converge to a global or local minimizer. We further apply our general framework to gradient descent and demonstrate its implications for both linear models and shallow neural networks. Empirical studies verify the effectiveness of our stability analyses.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Lei et al. (2026) studied this question.

synapsesocial.com/papers/69f594ca71405d493afffa32https://doi.org/10.1109/tpami.2026.3688491
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1On the Stability and Generalization of Triplet Learning2023 · 2 citations
  2. 2LIBSVM2011 · 41,487 citations
  3. 3Toward Understanding Generalization and Stability Gaps Between Centralized and Decentralized Federated Learning2025 · 1 citations
  4. 4Linear Convergence of Gradient and Proximal-Gradient Methods Under the Polyak-Łojasiewicz Condition2016 · 856 citations
  5. 5On the Stability and Generalization of Meta-Learning2024 · 2 citations