2019/09/16 by Madjid Khalilian, Khalilian, Madjid, Shiva Hassanzadeh +1
Computer Science · #Advanced Text Analysis Techniques #FOS: Computer and information sciences #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Spam and Phishing Detection #Text and Document Classification Technologies
paper · pdf · doi:10.48550/arxiv.1909.07368
openalex publication_date 2019/09/16 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28
Information on different fields which are collected by users requires appropriate management and organization to be structured in a standard way and retrieved fast and more easily. Document classification is a conventional method to separate text based on their subjects among scientific text, web pages and digital library. Different methods and techniques are proposed for document classifications that have advantages and deficiencies. In this paper, several unsupervised and supervised document classification methods are studied and compared.