Phishing Detection through Email Embeddings

12/28/2020
by   Luis Felipe Gutierrez, et al.
0

The problem of detecting phishing emails through machine learning techniques has been discussed extensively in the literature. Conventional and state-of-the-art machine learning algorithms have demonstrated the possibility of building classifiers with high accuracy. The existing research studies treat phishing and genuine emails through general indicators and thus it is not exactly clear what phishing features are contributing to variations of the classifiers. In this paper, we crafted a set of phishing and legitimate emails with similar indicators in order to investigate whether these cues are captured or disregarded by email embeddings, i.e., vectorizations. We then fed machine learning classifiers with the carefully crafted emails to find out about the performance of email embeddings developed. Our results show that using these indicators, email embeddings techniques is effective for classifying emails as phishing or legitimate.

READ FULL TEXT

page 7

page 8

research
03/27/2020

word2vec, node2vec, graph2vec, X2vec: Towards a Theory of Vector Embeddings of Structured Data

Vector representations of graphs and relational structures, whether hand...
research
01/08/2023

Prognosis and Treatment Prediction of Type-2 Diabetes Using Deep Neural Network and Machine Learning Classifiers

Type 2 Diabetes is a fast-growing, chronic metabolic disorder due to imb...
research
12/04/2020

Predicting Emotions Perceived from Sounds

Sonification is the science of communication of data and events to users...
research
12/05/2017

FlagIt: A System for Minimally Supervised Human Trafficking Indicator Mining

In this paper, we describe and study the indicator mining problem in the...
research
05/06/2018

Automatic Classification of Object Code Using Machine Learning

Recent research has repeatedly shown that machine learning techniques ca...
research
10/21/2019

A Single-MOSFET MAC for Confidence and Resolution (CORE) Driven Machine Learning Classification

Mixed-signal machine-learning classification has recently been demonstra...
research
11/05/2021

Toward Learning Human-aligned Cross-domain Robust Models by Countering Misaligned Features

Machine learning has demonstrated remarkable prediction accuracy over i....

Please sign up or login with your details

Forgot password? Click here to reset