Our method requires only linear time in input size, and still matches the information theoretical optimal sample complexity up to a data distribution dependent condition number factor.
Clustering is an important technique for identifying structural information in large-scale data analysis, where the underlying dataset may be too large to store.