Skip to main navigation Skip to search Skip to main content

Cyberbullying Detection in Low Resource Code-Mixed Languages Using ML and NLP Techniques

Research output: Chapter in Book/Report/Conference proceedingConference contribution

Abstract

Cyberbullying has become a pervasive issue on social media platforms, necessitating effective detection methods across diverse linguistic contexts. This study proposes a methodology for cyberbullying detection in Kannada, Kannada written in English, and English texts. Leveraging machine learning and natural language processing techniques, we develop and evaluate three types of models: traditional (Random Forest, Support Vector Machine), neural network (Single-Layer Perceptron, Multi-layer Perceptron), and deep neural network (Long Short-Term Memory, BERT, IndicBERT). Results show that deep learning models, particularly those leveraging pre-trained language models, exhibit superior performance in detecting offensive language, contributing to the ongoing efforts to combat cyberbullying in contemporary discourse. The model BERT outperforms the Random Forest model by 5.40 in accuracy and 13.79 in F1-score. Similarly, the Semi-Supervised model surpasses the Logistic Regression model by 10.62 in accuracy and 19.37 in F1-score, highlighting significant improvements. Additionally, an analysis of the underlying mechanisms highlights the inherent strengths of deep learning architectures, such as their ability to capture intricate linguistic patterns and contextual nuances, as well as model long-range dependencies within textual data, thereby enhancing their discriminatory capabilities in identifying cyberbullying behavior across multilingual and cross-script contexts.

Original languageEnglish
Title of host publicationArtificial Intelligence
Subtitle of host publicationTheory and Applications - Proceedings of AITA 2024
EditorsHarish Sharma, Antorweep Chakravorty, Shahid Hussain, Rajani Kumari
PublisherSpringer Science and Business Media Deutschland GmbH
Pages191-203
Number of pages13
ISBN (Print)9789819616862
DOIs
Publication statusPublished - 2025
Event2nd International Conference on Artificial Intelligence: Theory and Applications, AITA 2024 - Bengaluru, India
Duration: 09-08-202410-08-2024

Publication series

NameLecture Notes in Networks and Systems
Volume5589 LNNS
ISSN (Print)2367-3370
ISSN (Electronic)2367-3389

Conference

Conference2nd International Conference on Artificial Intelligence: Theory and Applications, AITA 2024
Country/TerritoryIndia
CityBengaluru
Period09-08-2410-08-24

All Science Journal Classification (ASJC) codes

  • Control and Systems Engineering
  • Signal Processing
  • Computer Networks and Communications

Fingerprint

Dive into the research topics of 'Cyberbullying Detection in Low Resource Code-Mixed Languages Using ML and NLP Techniques'. Together they form a unique fingerprint.

Cite this