2022/05/12 by Manikandan Ravikiran, Ravikiran, Manikandan, Bharathi Raja Chakravarthi +13
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection #Text Readability and Simplification
paper · pdf · doi:10.48550/arxiv.2205.06118
openalex publication_date 2022/05/12 · openalex created_date 2022/05/22 · openalex updated_date 2026/07/28
Offensive content moderation is vital in social media platforms to support healthy online discussions. However, their prevalence in codemixed Dravidian languages is limited to classifying whole comments without identifying part of it contributing to offensiveness. Such limitation is primarily due to the lack of annotated data for offensive spans. Accordingly, in this shared task, we provide Tamil-English code-mixed social comments with offensive spans. This paper outlines the dataset so released, methods, and results of the submitted systems