Inproceedings,

Don't follow me: Spam detection in Twitter

A. Wang.
Security and Cryptography (SECRYPT), Proceedings of the 2010 International Conference on, page 1--10. IEEE, (2010)

Abstract

The rapidly growing social network Twitter has been inﬁltrated by large amount of spam. In this paper, a spam detection prototype system is proposed to identify suspicious users on Twitter. A directed social graph model is proposed to explore the “follower” and “friend” relationships among Twitter. Based on Twitter’s spam policy, novel content-based features and graph-based features are also proposed to facilitate spam detection. A Web crawler is developed relying on API methods provided by Twitter. Around 25K users, 500K tweets, and 49M follower/friend relationships in total are collected from public available data on Twitter. Bayesian classiﬁcation algorithm is applied to distinguish the suspicious behaviors from normal ones. I analyze the data set and evaluate the performance of the detection system. Classic evaluation metrics are used to compare the performance of various traditional classiﬁcation methods. Experiment results show that the Bayesian classiﬁer has the best overall performance in term of F-measure. The trained classiﬁer is also applied to the entire data set. The result shows that the spam detection system can achieve 89% precision.

BibTeX key: wang2010don
entry type: inproceedings
booktitle: Security and Cryptography (SECRYPT), Proceedings of the 2010 International Conference on
year: 2010
organization: IEEE
pages: 1--10

BibSonomy

Don't follow me: Spam detection in Twitter

Abstract

Tags

Users

Comments and Reviewsshow / hide

Cite this publication

More citation styles

search on