Javascript must be enabled to continue!
Scalable Learning Framework for Detecting New Types of Twitter Spam with Misuse and Anomaly Detection
View through CrossRef
The growing popularity of social media has engendered the social problem of spam proliferation through this medium. New spam types that evade existing spam detection systems are being developed continually, necessitating corresponding countermeasures. This study proposes an anomaly detection-based framework to detect new Twitter spam, which works by modeling the characteristics of non-spam tweets and using anomaly detection to classify tweets deviating from this model as anomalies. However, because modeling varied non-spam tweets is challenging, the technique’s spam detection and false positive (FP) rates are low and high, respectively. To overcome this shortcoming, anomaly detection is performed on known spam tweets pre-detected using a trained decision tree while modeling normal tweets. A one-class support vector machine and an autoencoder with high detection rates are used for anomaly detection. The proposed framework exhibits superior detection rates for unknown spam compared to conventional techniques, while maintaining equivalent or improved detection and FP rates for known spam. Furthermore, the framework can be adapted to changes in spam conditions by adjusting the costs of detection errors.
Title: Scalable Learning Framework for Detecting New Types of Twitter Spam with Misuse and Anomaly Detection
Description:
The growing popularity of social media has engendered the social problem of spam proliferation through this medium.
New spam types that evade existing spam detection systems are being developed continually, necessitating corresponding countermeasures.
This study proposes an anomaly detection-based framework to detect new Twitter spam, which works by modeling the characteristics of non-spam tweets and using anomaly detection to classify tweets deviating from this model as anomalies.
However, because modeling varied non-spam tweets is challenging, the technique’s spam detection and false positive (FP) rates are low and high, respectively.
To overcome this shortcoming, anomaly detection is performed on known spam tweets pre-detected using a trained decision tree while modeling normal tweets.
A one-class support vector machine and an autoencoder with high detection rates are used for anomaly detection.
The proposed framework exhibits superior detection rates for unknown spam compared to conventional techniques, while maintaining equivalent or improved detection and FP rates for known spam.
Furthermore, the framework can be adapted to changes in spam conditions by adjusting the costs of detection errors.
Related Results
Faith Tweets: Ambient Religious Communication and Microblogging Rituals
Faith Tweets: Ambient Religious Communication and Microblogging Rituals
There’s no reason to think that Jesus wouldn’t have Facebooked or twittered if he came into the world now. Can you imagine his killer status updates? Reverend Schenck, New York, Al...
Spam Review Detection Techniques: A Systematic Literature Review
Spam Review Detection Techniques: A Systematic Literature Review
Online reviews about the purchase of products or services provided have become the main source of users’ opinions. In order to gain profit or fame, usually spam reviews are written...
SMS Spam Detection Using Machine Learning Approach
SMS Spam Detection Using Machine Learning Approach
Abstract
Currently, as the popularity of mobile phones has increased, Short Message Service (SMS) has grown tremendously. The minimal cost of messaging services has increased spam ...
Alts and Automediality: Compartmentalising the Self through Multiple Social Media Profiles
Alts and Automediality: Compartmentalising the Self through Multiple Social Media Profiles
IntroductionAlt, or alternative, accounts are secondary profiles people use in addition to a main account on a social media platform. They are a kind of automediation, a way of rep...
Perbandingan Kinerja Algoritma Naïve Bayes Dan C.45 Dalam Klasifikasi Spam Email
Perbandingan Kinerja Algoritma Naïve Bayes Dan C.45 Dalam Klasifikasi Spam Email
Antispam dengan algoritma tertentu yang dapat memisahkan antara spam-mail dengan non spam mail. Perbandingan kinerja antara algoritma naïve bayes, dan decision tree yang memakai al...
VNSED: Vietnamese spam email detection using multi deep learning models
VNSED: Vietnamese spam email detection using multi deep learning models
Email is one of the most popular communication methods today. However, a high percentage of spam emails are used for various purposes. Therefore, detecting spam emails and proposin...
Regulating Spam: Directive 2002/58 and Beyond
Regulating Spam: Directive 2002/58 and Beyond
This paper analyses the legal framework regulating unsolicited commercial communications or spam in the European Union. Our focus is on the Directive on privacy and electronic comm...
Analysis of Naıve Bayes Algorithm for Email Spam
Filtering
Analysis of Naıve Bayes Algorithm for Email Spam
Filtering
The upsurge in the volume of unwanted emails called spam has created an intense need for the
development of more dependable and robust antispam filters. Machine learning methods of...

