2013/08/22 by Longbing Cao · 4 citations
Computer Science · Business, Management and Accounting · #Data Mining Algorithms and Applications #Imbalanced Data Classification Techniques #Customer churn and segmentation
paper · doi:10.1093/comjnl/bxt084
openalex publication_date 2013/08/22 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/22
Most of the classic theoretical systems and tools in statistics, data mining and machine learning are built on the fundamental assumption of IIDness, which assumes the independence and identical distribution of underlying objects, attributes and/or values. However, complex behavioral and social problems often exhibit strong couplings and heterogeneity between values, attributes and objects (i.e., non-IIDness). This fundamentally challenges the IIDness-based learning methodologies and techniques. This paper presents a high-level overview of the needs, challenges and opportunities of non-IIDness learning for handling complex behavioral and social problems. By reviewing the nature and issues of classic IIDness-based algorithms in frequent pattern mining, clustering and classification to complex behavioral and social applications, concepts, structures, frameworks and exemplar techniques are discussed for non-IIDness learning. Case studies, relatedwork and prospects of non-IIDness learning are presented. Non-IIDness learning is also a fundamental issue in big data analytics. © The British Computer Society 2013.