Design of a Real-Time, Multilingual, Emotion-Aware Cyberbullying Detection System using Multi-Teacher Knowledge Distillation and Explainable AI
Keywords:
: Explainable AI, Cyberbullying, Real-Time NLP, Multi-Teacher Knowledge Distillation, XGBoost, SHAP, Emotion Detection, Sarcasm Detection, Multilingual NLPAbstract
Social media cyberbullying has propagated rapidly and is being experienced by individuals worldwide. It tends to be expressed using sarcasm, emotional language, and multiple languages, making it difficult to determine the identity of the perpetrator. Although automated detection systems are becoming increasingly prevalent, the majority of existing systems suffer from language issues, function only in offline batch mode, and are black-box models that cannot be interpreted. These constraints make it more difficult to intervene with speed and transparency.
This paper offers a real-time, multilingual system for detecting cyberbullying, using explainable AI, emotion and sarcasm detection, and Multi-Teacher Knowledge Distillation (MTKD) to address shortcomings.
The system leverages an ensemble of transformer-based teacher models, like mBERT, XLM-R, and IndicBERT, to capture language-specific features. Then, the models collaborate to produce a lightweight XGBoost classifier. To assist with the interpretation of context, additional layers are incorporated to identify sarcasm and emotion. SHAP (SHapley Additive Explanations) is employed to provide each prediction token-level interpretability. Algorithmic and architectural design of a system that would form a transparent, efficient, and deployable solution to detect cyberbullying in different emotional and linguistic contexts is the focus of this study.
References
(2018) Languages: SAIL Code-Mixed Shared Task. https://arxiv.org/abs/1803.06745
S. Mehendale, D. Dodia, H. Palshetkar (2022) A Review on Cyberbullying Detection Using Machine Learning in English and Hinglish. https://papers.ssrn.com/sol3/papers.cfm?abstract_id=4116153
S. Prasomphan (2025) Enhance Social Network Bullying Detection Using Multi-Teacher Knowledge Distillation With XGBoost Classifier. https://www.researchgate.net/publication/392212453
D. Demszky, D. Movshovitz-Attias, J. Ko, A. Cowen, G. Nemade, S. Ravi (2020) GoEmotions: A Dataset of Fine-Grained Emotions. https://arxiv.org/abs/2005.00547
S. M. Lundberg, S.-I. Lee (2017) A Unified Approach to Interpreting Model Predictions. 30. https://proceedings.neurips.cc/paper_files/paper/2017/file/8a20a8621978632d76c43dfd28b67767-Paper.pdf
A. Mishra, A. Jain, P. Bhattacharyya (2016) A Deep Learning Approach to Sarcasm Detection in Social Media. https://arxiv.org/abs/1605.01159
Y. Liu, M. Ott, N. Goyal (2019) RoBERTa: A Robustly Optimized BERT Pretraining Approach. https://arxiv.org/abs/1907.11692
(2021) UNICEF calls for concerted action to prevent online bullying. https://www.unicef.org/india/press-releases/safer-internet-day-unicef-calls-concerted-action-prevent-bullying-and-harassment
V. Srivastava, M. Singh (2021) Challenges and Considerations with Code-Mixed NLP for Multilingual Societies. https://arxiv.org/abs/2106.07823
K. Maity, R. Jain, P. Jha, S. Saha (2024) Explainable Cyberbullying Detection in Hinglish: A Generative Approach. 11(3), 3338–3347. https://doi.org/10.1109/TCSS.2023.3333675
C. Felbo, A. Mislove, A. Søgaard, I. Rahwan, S. Lehmann (2017) Using millions of emoji occurrences to learn any-domain representations for detecting sentiment, emotion and sarcasm. 1615–1625. https://aclanthology.org/D17-1169
A. Bapna, et al. (2022) IndicTrans: An effective transformer-based model for English–Indic translation. https://aclanthology.org/2022.naacl-main.58
J. Devlin, M. Chang, K. Lee, K. Toutanova (2019) BERT: Pre-training of deep bidirectional transformers for language understanding. 4171–4186. https://aclanthology.org/N19-1423
A. Conneau, et al. (2020) Unsupervised cross-lingual representation learning at scale. 8440–8451. https://aclanthology.org/2020.acl-main.747
K. Khanuja, et al. (2021) MuRIL: Multilingual representations for Indian languages. https://arxiv.org/abs/2103.10730
G. Hinton, O. Vinyals, J. Dean (2015) Distilling the knowledge in a neural network. https://arxiv.org/abs/1503.02531
T. Chen, C. Guestrin (2016) XGBoost: A scalable tree boosting system. 785–794. https://doi.org/10.1145/2939672.2939785
MongoDB Atlas. https://www.mongodb.com/cloud/atlas
Firebase Firestore. https://firebase.google.com/docs/firestore
PRAW (Python Reddit API Wrapper). https://praw.readthedocs.io/
Pushshift Reddit API. https://pushshift.io/
(2025) YouTube Data API v3. https://developers.google.com/youtube/v3
(2023) Community Standards Enforcement Report. https://transparency.fb.com
T. Fawcett (2006) An introduction to ROC analysis. 27(8), 861–874. https://doi.org/10.1016/j.patrec.2005.10.010
M. Caselli, V. Basile, E. Mitrović, B. Nissim (2020) HateBERT: Retraining BERT for abusive language detection in English. https://arxiv.org/abs/2010.12472
(2025) Heroku: Free cloud application hosting. https://www.heroku.com/free
(2022) Ending cyberbullying: Promoting safe and inclusive digital spaces for children. https://unesdoc.unesco.org/ark:/48223/pf0000381649
Downloads
Published
Issue
Section
License
Copyright (c) 2025 Authors and Global Journals Private Limited

This work is licensed under a Creative Commons Attribution 4.0 International License.
