CORTEXA
← Browse
crossrefAI2024-07-02Cited by 7

Arabic Spam Tweets Classification: A Comprehensive Machine Learning Approach

Wafa Hussain Hantom, Atta Rahman

Nowadays, one of the most common problems faced by Twitter (also known as X) users, including individuals as well as organizations, is dealing with spam tweets. The problem continues to proliferate due to the increasing popularity and number of users of social media platforms. Due to this overwhelming interest, spammers can post texts, images, and videos containing suspicious links that can be used to spread viruses, rumors, negative marketing, and sarcasm, and potentially hack the user’s information. Spam detection is among the hottest research areas in natural language processing (NLP) and cybersecurity. Several studies have been conducted in this regard, but they mainly focus on the English language. However, Arabic tweet spam detection still has a long way to go, especially emphasizing the diverse dialects other than modern standard Arabic (MSA), since, in the tweets, the standard dialect is seldom used. The situation demands an automated, robust, and efficient Arabic spam tweet detection approach. To address the issue, in this research, various machine learning and deep learning models have been investigated to detect spam tweets in Arabic, including Random Forest (RF), Support Vector Machine (SVM), Naive Bayes (NB) and Long-Short Term Memory (LSTM). In this regard, we have focused on the words as well as the meaning of the tweet text. Upon several experiments, the proposed models have produced promising results in contrast to the previous approaches for the same and diverse datasets. The results showed that the RF classifier achieved 96.78% and the LSTM classifier achieved 94.56%, followed by the SVM classifier that achieved 82% accuracy. Further, in terms of F1-score, there is an improvement of 21.38%, 19.16% and 5.2% using RF, LSTM and SVM classifiers compared to the schemes with same dataset.

View free PDFSource page

Related papers

crossrefAI2020-08-31Cited by 22

Maize Kernel Abortion Recognition and Classification Using Binary Classification Machine Learning Algorithms and Deep Convolutional Neural Networks

Lovemore Chipindu, Walter Mupangwa, Jihad Mtsilizah, Isaiah Nyagumbo, Mainassara Zaman-Allah

Maize kernel traits such as kernel length, kernel width, and kernel number determine the total kernel weight and, consequently, maize yield. Therefore, the measurement of kernel traits is important for maize breeding and the evaluation of maize yield. There are a few methods that…

View free PDFSource page
crossrefAI2025-09-21Cited by 1

Improving Remote Access Trojans Detection: A Comprehensive Approach Using Machine Learning and Hybrid Feature Engineering

AlsharifHasan Mohamad Aburbeian, Manuel Fernández-Veiga, Ahmad Hasasneh

Remote Access Trojans (RATs) pose a serious cybersecurity risk due to their stealthy control over compromised systems. This study presents a detection framework that integrates host, network, and newly engineered behavioral features to enhance the identification of RATs. Two sets…

View free PDFSource page
crossrefAI2025-02-08Cited by 14

Hybrid Machine Learning and Deep Learning Approaches for Insult Detection in Roman Urdu Text

Nisar Hussain, Amna Qasim, Gull Mehak, Olga Kolesnikova, Alexander Gelbukh, Grigori Sidorov

Thisstudy introduces a new model for detecting insults in Roman Urdu, filling an important gap in natural language processing (NLP) for low-resource languages. The transliterated nature of Roman Urdu also poses specific challenges from a computational linguistics perspective, inc…

View free PDFSource page
crossrefAI2024-05-10Cited by 4

From Eye Movements to Personality Traits: A Machine Learning Approach in Blood Donation Advertising

Stefanos Balaskas, Maria Koutroumani, Maria Rigou, Spiros Sirmakessis

Blood donation heavily depends on voluntary involvement, but the problem of motivating and retaining potential blood donors remains. Understanding the personality traits of donors can assist in this case, bridging communication gaps and increasing participation and retention. To…

View free PDFSource page
crossrefAI2024-09-24Cited by 31

Software Defect Prediction Based on Machine Learning and Deep Learning Techniques: An Empirical Approach

Waleed Albattah, Musaad Alzahrani

Software bug prediction is a software maintenance technique used to predict the occurrences of bugs in the early stages of the software development process. Early prediction of bugs can reduce the overall cost of software and increase its reliability. Machine learning approaches…

View free PDFSource page