Multi-Stage Training for Abusive Comment Detection in Indic Languages
Summary
This paper proposes a multi-stage training pipeline using language-based preprocessing and an ensemble of models to detect abusive comments in Indic languages, aiming to minimize false positives while preserving freedom of expression.
View Cached Full Text
Cached at: 05/22/26, 08:46 AM
# Multi-Stage Training for Abusive Comment Detection in Indic Languages Source: [https://arxiv.org/abs/2605.22380](https://arxiv.org/abs/2605.22380) [View PDF](https://arxiv.org/pdf/2605.22380) > Abstract:In recent years social media has become an increasingly popular tool for communication\. People use it to share their ideas, exchange information, and discuss thoughts\. Given its prevalence and widespread reach, social media must remain a safe space for people\. Content generated on social media can be abusive and it has become increasingly important to detect such content\. In this paper, we use a language\-based preprocessing and an ensemble of several models and analyze their performance of abusive comment detection\. Through extensive experimentation, we propose a pipeline that minimizes the false\-positive rate \(marking non\-abusive as abusive\) so that these systems can detect abusive comments without undermining the freedom of expression\. ## Submission history From: Pranshu Rastogi \[[view email](https://arxiv.org/show-email/a923246b/2605.22380)\] **\[v1\]**Thu, 21 May 2026 12:09:53 UTC \(486 KB\)
Similar Articles
Cross-Platform Chinese Offensive Comment Detection via Dual-Threshold Hard Example Mining
This paper proposes a dual-threshold hard example mining strategy for cross-platform Chinese offensive comment detection, addressing performance degradation due to domain shift. The method fine-tunes a RoBERTa model on the COLD dataset and adapts it to four Chinese social media platforms with minimal labeled data.
Inspect India Evals: An Open Benchmarking Framework for Evaluating Large Language Models in the Indian Linguistic and Cultural Context
Introduces Inspect India Evals, an open-source framework for evaluating LLMs in Indian linguistic and cultural contexts, with six benchmarks testing multilingual ability, bias, safety, and cultural knowledge. Tests on five models show Sarvam-M 24B and Gemma 2 27B lead.
Who and What? Using Linguistic Features and Annotator Characteristics to Analyze Annotation Variation
This paper presents a large-scale analysis of four harmful language detection datasets, examining how annotator characteristics and linguistic features interact to influence annotation variation. It highlights intersectional effects and warns against generalizing findings across different datasets.
IndicTalk: A Large-Scale Persona-Based Multilingual Conversational Corpus for Indic Languages
IndicTalk is a large-scale multilingual conversational corpus covering 9 Indic languages with code-mixed dialogues, generated via an automated pipeline with news grounding and persona conditioning, aimed at advancing conversational AI for underrepresented languages.
A Survey of Toxicity Detection and Mitigation Strategies for Multilingual Language Models
This survey synthesizes research on toxicity detection and detoxification for multilingual large language models, cataloging threat models, task formulations, detection approaches, and mitigation strategies, while identifying persistent challenges such as uneven language coverage and culturally contingent definitions of harm.