Average Sensitivity of Spectral Clustering

26Citations
Citations of this article
22Readers
Mendeley users who have this article in their library.
Get full text

Abstract

Spectral clustering is one of the most popular clustering methods for finding clusters in a graph, which has found many applications in data mining. However, the input graph in those applications may have many missing edges due to error in measurement, withholding for a privacy reason, or arbitrariness in data conversion. To make reliable and efficient decisions based on spectral clustering, we assess the stability of spectral clustering against edge perturbations in the input graph using the notion of average sensitivity, which is the expected size of the symmetric difference of the output clusters before and after we randomly remove edges. We first prove that the average sensitivity of spectral clustering is proportional to $łambda-2/łambda-3 2$, where $łambda-i$ is the i-th smallest eigenvalue of the (normalized) Laplacian. We also prove an analogous bound for k-way spectral clustering, which partitions the graph into k clusters. Then, we empirically confirm our theoretical bounds by conducting experiments on synthetic and real networks. Our results suggest that spectral clustering is stable against edge perturbations when there is a cluster structure in the input graph.

Cite

CITATION STYLE

APA

Peng, P., & Yoshida, Y. (2020). Average Sensitivity of Spectral Clustering. In Proceedings of the ACM SIGKDD International Conference on Knowledge Discovery and Data Mining (pp. 1132–1140). Association for Computing Machinery. https://doi.org/10.1145/3394486.3403166

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free