Near-Optimal Stochastic Approximation for Online Principal Component Estimation (1603.05305v4)

Published 16 Mar 2016 in math.OC and stat.ML

Abstract: Principal component analysis (PCA) has been a prominent tool for high-dimensional data analysis. Online algorithms that estimate the principal component by processing streaming data are of tremendous practical and theoretical interests. Despite its rich applications, theoretical convergence analysis remains largely open. In this paper, we cast online PCA into a stochastic nonconvex optimization problem, and we analyze the online PCA algorithm as a stochastic approximation iteration. The stochastic approximation iteration processes data points incrementally and maintains a running estimate of the principal component. We prove for the first time a nearly optimal finite-sample error bound for the online PCA algorithm. Under the subgaussian assumption, we show that the finite-sample error bound closely matches the minimax information lower bound.

Citations (64)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Near-Optimal Stochastic Approximation for Online Principal Component Estimation (1603.05305v4)

Summary

Related Papers