Javascript must be enabled to continue!
Intelligent Deep Machine Learning Cyber Phishing URL Detection Based on BERT Features Extraction
View through CrossRef
Recently, phishing attacks have been a crucial threat to cyberspace security. Phishing is a form of fraud that attracts people and businesses to access malicious uniform resource locators (URLs) and submit their sensitive information such as passwords, credit card ids, and personal information. Enormous intelligent attacks are launched dynamically with the aim of tricking users into thinking they are accessing a reliable website or online application to acquire account information. Researchers in cyberspace are motivated to create intelligent models and offer secure services on the web as phishing grows more intelligent and malicious every day. In this paper, a novel URL phishing detection technique based on BERT feature extraction and a deep learning method is introduced. BERT was used to extract the URLs’ text from the Phishing Site Predict dataset. Then, the natural language processing (NLP) algorithm was applied to the unique data column and extracted a huge number of useful data features in terms of meaningful text information. Next, a deep convolutional neural network method was utilised to detect phishing URLs. It was used to constitute words or n-grams in order to extract higher-level features. Then, the data were classified into legitimate and phishing URLs. To evaluate the proposed method, a famous public phishing website URLs dataset was used, with a total of 549,346 entries. However, three scenarios were developed to compare the outcomes of the proposed method by using similar datasets. The feature extraction process depends on natural language processing techniques. The experiments showed that the proposed method had achieved 96.66% accuracy in the results, and then the obtained results were compared to other literature review works. The results showed that the proposed method was efficient and valid in detecting phishing websites’ URLs.
Title: Intelligent Deep Machine Learning Cyber Phishing URL Detection Based on BERT Features Extraction
Description:
Recently, phishing attacks have been a crucial threat to cyberspace security.
Phishing is a form of fraud that attracts people and businesses to access malicious uniform resource locators (URLs) and submit their sensitive information such as passwords, credit card ids, and personal information.
Enormous intelligent attacks are launched dynamically with the aim of tricking users into thinking they are accessing a reliable website or online application to acquire account information.
Researchers in cyberspace are motivated to create intelligent models and offer secure services on the web as phishing grows more intelligent and malicious every day.
In this paper, a novel URL phishing detection technique based on BERT feature extraction and a deep learning method is introduced.
BERT was used to extract the URLs’ text from the Phishing Site Predict dataset.
Then, the natural language processing (NLP) algorithm was applied to the unique data column and extracted a huge number of useful data features in terms of meaningful text information.
Next, a deep convolutional neural network method was utilised to detect phishing URLs.
It was used to constitute words or n-grams in order to extract higher-level features.
Then, the data were classified into legitimate and phishing URLs.
To evaluate the proposed method, a famous public phishing website URLs dataset was used, with a total of 549,346 entries.
However, three scenarios were developed to compare the outcomes of the proposed method by using similar datasets.
The feature extraction process depends on natural language processing techniques.
The experiments showed that the proposed method had achieved 96.
66% accuracy in the results, and then the obtained results were compared to other literature review works.
The results showed that the proposed method was efficient and valid in detecting phishing websites’ URLs.
Related Results
PUMMP: Phishing URL Detection using Machine Learning with Monomorphic and Polymorphic Treatment of Features
PUMMP: Phishing URL Detection using Machine Learning with Monomorphic and Polymorphic Treatment of Features
Phishing scams are increasing drastically, which affects Internet users in compromising personal credentials. This paper proposes a novel feature utilization method for phishing UR...
Phishing Cyber Security Threats
Phishing Cyber Security Threats
Phishing is a growing threat in the realm of cybersecurity, where cybercriminals use various phishing techniques to steal sensitive information from individuals and organizations. ...
Anti-Phishing Technologies and Tools
Anti-Phishing Technologies and Tools
Phishing continues to be one of the most common and effective forms of cyber security threats and involve deception of users thereby getting them provide unauthorized individuals w...
Intelligent Detection Designs of HTML URL Phishing Attacks
Intelligent Detection Designs of HTML URL Phishing Attacks
Phishing attacks are a type of cybercrime that has grown in recent years. It is part of social engineering attacks where an attacker deceives users by sending fake messages using s...
AI-powered phishing detection: Integrating natural language processing and deep learning for email security
AI-powered phishing detection: Integrating natural language processing and deep learning for email security
Phishing attacks are major threats to email security and pose challenges, while cyber attackers utilize increasingly sophisticated means to deceive the user and steal away importan...
Spear-Phishing in the Wild: A Real-World Study of Personality, Phishing Self-Efficacy and Vulnerability to Spear-Phishing Attacks
Spear-Phishing in the Wild: A Real-World Study of Personality, Phishing Self-Efficacy and Vulnerability to Spear-Phishing Attacks
Recent research has begun to focus on the factors that cause people to respond to phishing attacks. In this study a real-world spear-phishing attack was performed on employees in o...
Selection of Injectable Drug Product Composition using Machine Learning Models (Preprint)
Selection of Injectable Drug Product Composition using Machine Learning Models (Preprint)
BACKGROUND
As of July 2020, a Web of Science search of “machine learning (ML)” nested within the search of “pharmacokinetics or pharmacodynamics” yielded over 100...
Knowledge-Grounded LLM-Driven Augmentation via Graph RAG for Phishing URL Detection
Knowledge-Grounded LLM-Driven Augmentation via Graph RAG for Phishing URL Detection
Integrating prior rule knowledge into generative data augmentation remains an open problem when detectors must generalize under distribution shift. URL-based phishing illustrates t...

