LICENCE
Creative Commons CCZero (CC0-1.0), Open Data Commons Public Domain Dedication and Licence (PDDL-1.0)
DOMAIN
Science and Technology
COVERAGE
EU-wide
FORMATS
ARFF, CSV, JSON, RDF, XLS / XLSX
PERSONAL DATA PROTECTION
No personal data
* Please note that the classification is taken from the original source
One of the challenges faced by our research was the unavailability of reliable training datasets. In fact this challenge faces any researcher in the field. However, although plenty of articles about predicting phishing websites have been disseminated these days, no reliable training dataset has been published publically, may be because there is no agreement in literature on the definitive features that characterize phishing webpages, hence it is difficult to shape a dataset that covers all possible features. In this dataset, we shed light on the important features that have proved to be sound and effective in predicting phishing websites. In addition, we propose some new features.