2025/06/23 by Holli Sargeant, Hannes Waldetoft, Måns Magnusson · 1 voice
Computer Science · Social Sciences · #Hate Speech and Cyberbullying Detection #Law in Society and Culture
paper · pdf · doi:10.1145/3715275.3732016
openalex publication_date 2025/06/23 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/29
Hate crimes, driven by biases against specific demographic groups, harm not only individuals but undermine the security, trust, and cohesion of entire communities.Accurately identifying such crimes remains a significant challenge due to under-reporting, limited training, and the complexity of determining bias motivations.In this paper, we analyze the results of a text classification model developed to improve the precision of hate crime statistics and identification in Sweden.Empirical results indicate the model outperforms traditional manual police classification of hate crimes, achieving higher precision across various crime types and regions.We further disaggregate performance to pinpoint persistent challenges and highlight categories where both human and machine decision-makers struggle.While the model focuses on statistical estimation rather than direct case-level decision-making, we discuss the broader implications of algorithmic transparency, accountability, and explainability.Ultimately, this research illustrates how transformer-based neural networks can responsibly bolster the detection and understanding of hate crimes, informing policies to better protect vulnerable communities.