Submitted:
05 August 2025
Posted:
06 August 2025
You are already at the latest version
Abstract
Keywords:
Introduction
Background
Problem Statement
Objectives
Literature Review
Model and Methodology
Data
4.1. Structure and Composition
4.2. Label Encoding
4.3. Shuffling and Splitting Strategy
4.4. Key Features
4.5. Challenges
Model Architecture

Model Training

Results and Discussion
Conclusion
Future Scope
References
- Alammar, J. The illustrated transformer. 2021. Available online: https://jalammar.github.io/illustrated-transformer/.
- Bojanowski, P.; Grave, E.; Joulin, A.; Mikolov, T. Enriching word vectors with subword information. Transactions of the Association for Computational Linguistics 2017, 5, 135–146. [Google Scholar] [CrossRef]
- Ba, J.L.; Kiros, J.R.; Hinton, G.E. Layer normalization. arXiv arXiv:1607.06450, 2016. [PubMed]
- Bilal, M.; Khan, A.; Jan, S.; Musa, S.; Ali, S. Roman urdu hate speech detection using transformer-based model for cyber security applications. Sensors 2023, 23, 3909. [Google Scholar] [CrossRef] [PubMed]
- Sanh, V.; Debut, L.; Chaumond, J.; Wolf, T. Distilbert, a distilled version of bert: Smaller, faster, cheaper and lighter. arXiv 2019. [CrossRef]
- Chaudhary, A. A visual guide to fasttext word embeddings. Available online: https://amitness.com/2020/06/fasttext-embeddings/ (accessed on 30 March 2022).
- Das, M.; Saha, P.; Mathew, B.; Mukherjee, A. Hatecheckhin: Evaluating hindi hate speech detection models. arXiv arXiv:2205.00328, 2022.
- Das, M.; Banerjee, S.; Saha, P. Abusive and threatening language detection in urdu using boosting based and bert based models: A comparative approach. arXiv 2021. [CrossRef]
- Devlin, J.; Chang, M.-W.; Lee, K.; Toutanova, K. Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv 2018. [CrossRef]
- Devjyoti Chakrobarty, A.R.G. ‘An introduction to the global vectors (glove) algorithm. 2021. Available online: https://wandb.ai/authors/embeddings-2/reports/GloVe--VmlldzozNDg2NTQ.
- ‘‘Distillation of bert-like models: The theory. Available online: https://towardsdatascience.com/distillation-of-bert-like-models-the-theory-32e19a02641f (accessed on 30 March 2023).
- Venugopal, K. Mathematical introduction to glove word embedding. 2021. Available online: https://becominghuman.ai/mathematical-introduction-to-glove-wordembedding-60f24154e54c.
- Jadhav, I.; Kanade, A.; Waghmare, V.; Chaudhari, D. Hate and offensive speech detection in hindi twitter corpus. in Forum for Information Retrieval Evaluation (Working Notes)(FIRE), CEUR-WS. org, 2021.
- Jha, A. Vectorization techniques in nlp [guide]. 2023. Available online: https://neptune.ai/blog/vectorization-techniques-in-nlp-guide.
- Jitender. Implement sigmoid function using numpy. 2023. Available online: https://www.geeksforgeeks.org/implement-sigmoid-function-using-numpy/#article-meta-div.
- Kakwani, D.; Kunchukuttan, A.; Golla, S.; Gokul, N.; Bhattacharyya, A.; Khapra, M.M.; Kumar, P. Indicnlpsuite: Monolingual corpora, evaluation benchmarks and pre-trained multilingual language models for indian languages. in Findings of the Association for Computational Linguistics: EMNLP 2020, 2020, pp. 4948--4961.
- Khan, M.M.; Shahzad, K.; Malik, M.K. Hate speech detection in roman urdu. ACM Transactions on Asian and Low-Resource Language Information Processing (TALLIP) 2021, 20, 1–19. [Google Scholar] [CrossRef]
- Khanuja, S.; Bansal, D.; Mehtani, S.; Khosla, S.; Dey, A.; Gopalan, B.; Margam, D.K.; Aggarwal, P.; Nagipogu, R.T.; Dave, S.; et al. Muril: Multilingual representations for indian languages. arXiv 2021. [CrossRef]
- Mandl, T.; Modha, S.; Shahi, G.K.; Madhu, H.; Satapara, S.; Majumder, P.; Schäfer, J.; Ranasinghe, T.; Zampieri, M.; Nandini, D.; et al. Overview of the hasoc subtrack at fire 2021: Hate speech and offensive content identification in english and indo-aryan languages. arXiv 2021. [CrossRef]
- Mehmood, K.; Essam, D.; Shafi, K.; Malik, M.K. An unsupervised lexical normalization for roman hindi and urdu sentiment analysis. Information Processing & Management 2020, 57, 102368. [Google Scholar] [CrossRef]
- Mikolov, T.; Chen, K.; Corrado, G.; Dean, J. Efficient estimation of word representations in vector space. arXiv 2013. [CrossRef]
- Mahabert. Available online: https://huggingface.co/l3cube-pune/marathi-bert (accessed on 30 March 2022).
- Mishra, R. devanagari-to-roman-script-transliteration. 2019. Available online: https://github.com/ritwikmishra/devanagari-to-roman-script-transliteration.
- Weng, L. Learning word embedding. 2017. Available online: https://lilianweng.github.io/posts/2017-10-15-word-embedding/.
- Pennington, J.; Socher, R.; Manning, C.D. Glove: Global vectors for word representation. In Proceedings of the 2014 conference on empirical methods in natural language processing (EMNLP); 2014; pp. 1532–1543. [Google Scholar]
- Sharma, A.; Kabra, A.; Jain, M. Ceasing hate with moh: Hate speech detection in hindi--english code-switched language. Information Processing & Management 2022, 59, 102760. [Google Scholar]
- Shukla, S.; Nagpal, S.; Sabharwal, S. Hate speech detection in hindi language using bert and convolution neural network. in 2022 International Conference on Computing, Communication, and Intelligent Systems (ICCCIS). IEEE, 2022, pp. 642--647.
- Srivastava, N.; Hinton, G.; Krizhevsky, A.; Sutskever, I.; Salakhutdinov, R. Dropout: A simple way to prevent neural networks from overfitting. The Journal of Machine Learning Research 2014, 15, 1929–1958. [Google Scholar]
- Velankar, A.; Patil, H.; Gore, A.; Salunke, S.; Joshi, R. Hate and offensive speech detection in hindi and marathi. arXiv 2021. [CrossRef]
- Velankar, A.; Patil, H.; Gore, A.; Salunke, S.; Joshi, R. L3cube-mahahate: A tweet-based marathi hate speech detection dataset and bert models. arXiv arXiv:2203.13778, 2022.
- Phi, M. Illustrated guide to transformers- step by step explanation. Available online: https://towardsdatascience.com/illustrated-guide-to-transformers-stepby-step-explanation-f74876522bc0 (accessed on 30 March 2022).
- Navlani, A. Decision tree classification in python tutorial. 2023. Available online: https://www.datacamp.com/tutorial/decision-tree-classification-python.
Disclaimer/Publisher’s Note: The statements, opinions and data contained in all publications are solely those of the individual author(s) and contributor(s) and not of MDPI and/or the editor(s). MDPI and/or the editor(s) disclaim responsibility for any injury to people or property resulting from any ideas, methods, instructions or products referred to in the content. |
© 2025 by the authors. Licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license (http://creativecommons.org/licenses/by/4.0/).