On the Computational Complexity of Private High-dimensional Model Selection (2310.07852v5)

Published 11 Oct 2023 in stat.ML, cs.LG, stat.CO, and stat.ME

Abstract: We consider the problem of model selection in a high-dimensional sparse linear regression model under privacy constraints. We propose a differentially private (DP) best subset selection method with strong statistical utility properties by adopting the well-known exponential mechanism for selecting the best model. To achieve computational expediency, we propose an efficient Metropolis-Hastings algorithm and under certain regularity conditions, we establish that it enjoys polynomial mixing time to its stationary distribution. As a result, we also establish both approximate differential privacy and statistical utility for the estimates of the mixed Metropolis-Hastings chain. Finally, we perform some illustrative experiments on simulated data showing that our algorithm can quickly identify active features under reasonable privacy budget constraints.

References (55)

Summary

We haven't generated a summary for this paper yet.

Summarize Now

GitHub

GitHub - roysaptaumich/DP-BSS: Metropolis-Hastings algorithm for differentially private best subset selection algorithms proposed in this paper. (2 stars)

On the Computational Complexity of Private High-dimensional Model Selection (2310.07852v5)

Summary

Related Papers

GitHub

Tweets