摘要
The most critical challenge in cybersecurity is dealing with cyber-attacks. Since it is imperative to act quickly to lessen the harm caused by cyberbullying. The complicated dynamics of social media, which are marked by their complexity, variety, subjectivity, and multimodal nature, provide obstacles to the identification of cyberbullying. The complexity, diversity, subjective nature, and multimodal aspects of social media have significantly increased. This has led to the need for automated mechanisms that can identify these harmful behaviors. This study aims to assess how well various categorization methods detect cyber bullying. For training and testing, this study uses a cybersecurity-related data. The models that have been selected include the Linear SVC, Random Forest, Decision Tree, Logistic Regression, and Stochastic Gradient classifiers. We use hyperparameter tuning to improve the model's performance, and then we show the results based on important metrics like accuracy, precision, recall, and F1 score. The results demonstrate the superiority of the stochastic gradient classifier, which has an F1 score of 94.39%, recall of 91.94%, accuracy of 92.81%, and precision of 96.97%. The investigation examines the advantages and disadvantages of each approach, offering insightful information for the cybersecurity field. In addition, suggestions for more studies are made to strengthen the resilience of cyber defenses. This work advances the effectiveness of cybersecurity measures by finding the best models for detecting threats and offering directions for improvement as cyber threats change. Other techniques that can be used are the three distinct feature extraction techniques—Bag of Words (BoW), Term Frequency-Inverse Document Frequency (TF-IDF), and Word2Vec—are merged with the algorithms of Logistic Regression (LR), Naïve Bayes (NB), Support Vector Machine (SVM), and Random Forest (RF) to build the model (Johari & Jaafar, 2022). Our objectives are to analyze the effectiveness of several classification techniques for identifying cyberbullying, such as Random Forest, Decision Tree, Linear SVC, Logistic Regression, and Stochastic Gradient classifiers, to improve the performance of the model by using hyperparameter tweaking methods and analyze the outcomes using the F1 score, accuracy, precision, recall, and other critical performance