arXiv · 2605.22380
Multi-Stage Training for Abusive Comment Detection in Indic Languages
Abstract
In recent years social media has become an increasingly popular tool for communication. People use it to share their ideas, exchange information, and discuss thoughts. Given its prevalence and widespread reach, social media must remain a safe space for people. Content generated on social media can be abusive and it has become increasingly important to detect such content. In this paper, we use a language-based preprocessing and an ensemble of several models and analyze their performance of abusive comment detection. Through extensive experimentation, we propose a pipeline that minimizes the false-positive rate (marking non-abusive as abusive) so that these systems can detect abusive comments without undermining the freedom of expression.
Explore related subjects
Keep this discovery
Pranshu Rastogi, Madhav Mathur, Ramaneswaran S, Kshitij Mohan. 2026-05-21. Multi-Stage Training for Abusive Comment Detection in Indic Languages. https://arxiv.org/abs/2605.22380
Cite the original work for its findings. Save a collection to share your selection of sources.