Papers
Topics
Authors
Recent
Search
2000 character limit reached

Escaping Saddle Points via Curvature-Calibrated Perturbations: A Complete Analysis with Explicit Constants and Empirical Validation

Published 22 Aug 2025 in cs.LG, math.OC, and stat.ML | (2508.16540v1)

Abstract: We present a comprehensive theoretical analysis of first-order methods for escaping strict saddle points in smooth non-convex optimization. Our main contribution is a Perturbed Saddle-escape Descent (PSD) algorithm with fully explicit constants and a rigorous separation between gradient-descent and saddle-escape phases. For a function f:R<sup>dRf:\mathbb{R}<sup>d\to\mathbb{R} with \ell-Lipschitz gradient and ρ\rho-Lipschitz Hessian, we prove that PSD finds an (ϵ,ρϵ)(\epsilon,\sqrt{\rho\epsilon})-approximate second-order stationary point with high probability using at most O(Δf/ϵ<sup>2)O(\ell\Delta_f/\epsilon<sup>2) gradient evaluations for the descent phase plus O((/ρϵ)log(d/δ))O((\ell/\sqrt{\rho\epsilon})\log(d/\delta)) evaluations per escape episode, with at most O(Δf/ϵ<sup>2)O(\ell\Delta_f/\epsilon<sup>2) episodes needed. We validate our theoretical predictions through extensive experiments across both synthetic functions and practical machine learning tasks, confirming the logarithmic dimension dependence and the predicted per-episode function decrease. We also provide complete algorithmic specifications including a finite-difference variant (PSD-Probe) and a stochastic extension (PSGD) with robust mini-batch sizing.

Authors (2)

Summary

No one has generated a summary of this paper yet.

Paper to Video (Beta)

No one has generated a video about this paper yet.

Whiteboard

No one has generated a whiteboard explanation for this paper yet.

Open Problems

We haven't generated a list of open problems mentioned in this paper yet.

Continue Learning

We haven't generated follow-up questions for this paper yet.

Tweets

Sign up for free to view the 2 tweets with 2 likes about this paper.