vix.ing · top · new · best · stats · spec

LSHTC: A Benchmark for Large-Scale Text Classification

2015/03/30 by Ioannis Partalas, Aris Kosmopoulos, Partalas, Ioannis +14 · 5 citations
Computer Science · #Algorithms and Data Compression #Computation and Language (cs.CL) #FOS: Computer and information sciences #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Spam and Phishing Detection #Text and Document Classification Technologies

paper · pdf · doi:10.48550/arxiv.1503.08581

openalex publication_date 2015/03/30 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28

Abstract

LSHTC is a series of challenges which aims to assess the performance of classification systems in large-scale classification in a a large number of classes (up to hundreds of thousands). This paper describes the dataset that have been released along the LSHTC series. The paper details the construction of the datsets and the design of the tracks as well as the evaluation measures that we implemented and a quick overview of the results. All of these datasets are available online and runs may still be submitted on the online server of the challenges.

Cited by

Related