MRI
MRI India Journals Vol. 15 No. 1S (2026): Special Issue on Cognition, Human and Artificial Intelligence

From Text to Toxicity: Exploring the Challenges in Marathi Hate Speech Detection

Authors

  • Preeti V. Sarode Department of Computer Science and Information Technology K. M. Agrawal College, Kalyan, Maharashtra, India.
  • Harshali B. Patil Department of Computer Science Dr. Annasaheb G. D. Bendale Mahila Mahavidyalaya, Jalgaon, Maharashtra, India.

DOI:

https://doi.org/10.65521/ijacte.v15i1S.1318

Keywords:

Machine Learning Deep Learning Hate Speech

Abstract

Hate speech on social media has become a serious problem in society. There is a increasing need to create systems that can automatically detect such hateful content. Marathi, an Indo-Aryan language widely spoken in India, remains under-represented in natural language processing research due to limited linguistic resources and annotated datasets. This study focuses on the major challenges faced in detecting hate speech in the Marathi language. Marathi is a low-resource and complex language. It does not have enough large and balanced datasets for proper model training. The presence of code-mixing, transliteration between Devanagari and Roman scripts, and various dialects increases the complexity of text processing. Ambiguity in meaning, sarcasm, and context-dependent expressions make automatic detection more difficult. This paper systematically reviews the challenges in Marathi hate speech detection arising from data scarcity, code-mixing, transliteration, dialectal variation, annotation ambiguity, and context-dependent expressions.

Downloads

Published

2026-01-18

How to Cite

Sarode, P. V., & Patil, H. B. (2026). From Text to Toxicity: Exploring the Challenges in Marathi Hate Speech Detection. International Journal on Advanced Computer Theory and Engineering, 15(1S), 201–210. https://doi.org/10.65521/ijacte.v15i1S.1318

Similar Articles

1 2 3 4 5 6 7 8 9 10 > >> 

You may also start an advanced similarity search for this article.