2022/05/22 by Usman Mahmood, Mahmood, Usman, Daniel Pimentel-Alarcón +1
Computer Science · Mathematics · #Advanced Clustering Algorithms Research #Anomaly Detection Techniques and Applications #FOS: Computer and information sciences #Face and Expression Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.LG #stat.ML
paper · pdf · doi:10.48550/arxiv.2205.10872
Accepted at IJCNN 2022. arXiv admin note: substantial text overlap with arXiv:1808.00628
arxiv created 2022/05/22 · openalex publication_date 2022/05/22 · arxiv updated 2022/05/24 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28
This paper introduces \em fusion subspace clustering, a novel method to learn low-dimensional structures that approximate large scale yet highly incomplete data. The main idea is to assign each datum to a subspace of its own, and minimize the distance between the subspaces of all data, so that subspaces of the same cluster get \em fused together. Our method allows low, high, and even full-rank data; it directly accounts for noise, and its sample complexity approaches the information-theoretic limit. In addition, our approach provides a natural model selection \em clusterpath, and a direct completion method. We give convergence guarantees, analyze computational complexity, and show through extensive experiments on real and synthetic data that our approach performs comparably to the state-of-the-art with complete data, and dramatically better if data is missing.