首页    期刊浏览 2024年10月05日 星期六
登录注册

文章基本信息

  • 标题:Classifying Unsolicited Bulk Email (UBE) using Python Machine Learning Techniques
  • 本地全文:下载
  • 作者:Sabah Mohammed ; Osama Mohammed ; Jinan Fiaidhi
  • 期刊名称:International Journal of Hybrid Information Technology
  • 印刷版ISSN:1738-9968
  • 出版年度:2013
  • 卷号:6
  • 期号:1
  • 出版社:SERSC
  • 摘要:Email has become one of the fastest and most economical forms of communication. However, the increase of email users has resulted in the dramatic increase of spam emails during the past few years. As spammers always try to find a way to evade existing filters, new filters need to be developed to catch spam. Generally, the main tool for email filtering is based on text classification. A classifier then is a system that classifies incoming messages as spam or legitimate (ham) using classification methods. The most important methods of classification utilize machine learning techniques. There are a plethora of options when it comes to deciding how to add a machine learning component to a python email classification. This article describes an approach for spam filtering using Python where the interesting spam or ham words (spam-ham lexicon) are filtered first from the training dataset and then this lexicon is used to generate the training and testing tables that are used by variety of data mining algorithms. Our experimentation using one dataset reveals the affectivity of the Na.ve Bayes and the SVM classifiers for spam filtering
  • 关键词:Spam Filtering; Machine Learning; Python
国家哲学社会科学文献中心版权所有