2018/08/20 by Arthur Zimek, Peter Filzmoser · 1 citation
Computer Science · Mathematics · #Anomaly Detection Techniques and Applications #Advanced Statistical Methods and Models #Imbalanced Data Classification Techniques #Computer science #Anomaly detection #Outlier #Data mining #Artificial intelligence #Data science #Machine learning #Algorithm
paper · doi:10.1002/widm.1280
openalex publication_date 2018/08/20 · openalex created_date 2025/10/10 · openalex updated_date 2026/08/01
Outlier detection has been a topic in statistics for centuries. Over mainly the last two decades, there has been also an increasing interest in the database and data mining community to develop scalable methods for outlier detection. Initially based on statistical reasoning, however, these methods soon lost the direct probabilistic interpretability of the derived outlier scores. Here, we detail from a joint point of view of data mining and statistics the roots and the path of development of statistical outlier detection and of database‐related data mining methods for outlier detection. We discuss their inherent meaning, review approaches to again find a statistically meaningful interpretation of outlier scores, and sketch related current research topics. This article is categorized under: Algorithmic Development > Statistics Algorithmic Development > Scalable Statistical Methods Technologies > Machine Learning