arXiv · 2603.19279
Multilingual Hate Speech Detection and Counterspeech Generation: A Comprehensive Survey and Practical Guide
Abstract
Combating online hate speech in multilingual settings requires approaches that go beyond English-centric models and capture the cultural and linguistic diversity of global online discourse. This paper presents a comprehensive survey and practical guide to multilingual hate speech detection and counterspeech generation, integrating recent advances in natural language processing. We analyze why monolingual systems often fail in non-English and code-mixed contexts, missing implicit hate and culturally specific expressions. To address these challenges, we outline a structured three-phase framework - task design, data curation, and evaluation - drawing on state-of-the-art datasets, models, and metrics. The survey consolidates progress in multilingual resources and techniques while highlighting persistent obstacles, including data scarcity in low-resource languages, fairness and bias in system development, and the need for multimodal solutions. By bridging technical progress with ethical and cultural considerations, we provide researchers, practitioners, and policymakers with scalable guidelines for building context-aware, inclusive systems. Our roadmap contributes to advancing online safety through fairer, more effective detection and counterspeech generation across diverse linguistic environments.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Zahra Safdari Fesaghandis, Suman Kalyan Maity. 2026-03-01. Multilingual Hate Speech Detection and Counterspeech Generation: A Comprehensive Survey and Practical Guide. https://arxiv.org/abs/2603.19279
Cite the original work for its findings. Save a collection to share your selection of sources.