Do subsampled newton methods work for high-dimensional data?

Xiang Li; Shusen Wang; Zhihua Zhang

Conference ProceedingsOPEN ACCESS

Do subsampled newton methods work for high-dimensional data?

AAAI 2020 - 34th AAAI Conference on Artificial Intelligence (2020) 4723-4730

DOI: 10.1609/aaai.v34i04.5905

6Citations

19Readers

Abstract

Subsampled Newton methods approximate Hessian matrices through subsampling techniques to alleviate the per-iteration cost. Previous results require Ω(d) samples to approximate Hessians, where d is the dimension of data points, making it less practical for high-dimensional data. The situation is deteriorated when d is comparably as large as the number of data points n, which requires to take the whole dataset into account, making subsampling not useful. This paper theoretically justifies the effectiveness of subsampled Newton methods on strongly convex empirical risk minimization with high dimensional data. Specifically, we provably require only Θ(~ dγeff) samples for approximating the Hessian matrices, where dγeff is the γ-ridge leverage and can be much smaller than d as long as nγ 1. Our theories work for three types of Newton methods: subsampled Netwon, distributed Newton, and proximal Newton.

Cite

CITATION STYLE

APA

Li, X., Wang, S., & Zhang, Z. (2020). Do subsampled newton methods work for high-dimensional data? In AAAI 2020 - 34th AAAI Conference on Artificial Intelligence (pp. 4723–4730). AAAI press. https://doi.org/10.1609/aaai.v34i04.5905

Do subsampled newton methods work for high-dimensional data?

Abstract

Cite

Register to see more suggestions