PhishMatch: A Layered Approach for Effective Detection of Phishing URLs

TOP 文献データベース PhishMatch: A Layered Approach for Effective Detection of Phishing URLs

arxiv

AIセキュリティポータルbot

文献データベースの情報は、自動的に収集されています。

Source

https://arxiv.org/abs/2112.02226

PDF

https://arxiv.org/pdf/2112.02226

文献情報

作者: Harshal Tupsamudre;Sparsh Jain;Sachin Lodha
公開日: 2021-12-4
所属機関: TCS Research
所属の国: India
会議名: Computing Research Repository (CoRR)

AIにより推定されたラベル

フィッシング検出ユーザ行動分析メモリ管理手法

※ こちらのラベルはAIによって自動的に追加されました。そのため、正確でないことがあります。
詳細は文献データベースについてをご覧ください。

Abstract

Phishing attacks continue to be a significant threat on the Internet. Prior studies show that it is possible to determine whether a website is phishing or not just by analyzing its URL more carefully. A major advantage of the URL based approach is that it can identify a phishing website even before the web page is rendered in the browser, thus avoiding other potential problems such as cryptojacking and drive-by downloads. However, traditional URL based approaches have their limitations. Blacklist based approaches are prone to zero-hour phishing attacks, advanced machine learning based approaches consume high resources, and other approaches send the URL to a remote server which compromises user's privacy. In this paper, we present a layered anti-phishing defense, PhishMatch, which is robust, accurate, inexpensive, and client-side. We design a space-time efficient Aho-Corasick algorithm for exact string matching and n-gram based indexing technique for approximate string matching to detect various cybersquatting techniques in the phishing URL. To reduce false positives, we use a global whitelist and personalized user whitelists. We also determine the context in which the URL is visited and use that information to classify the input URL more accurately. The last component of PhishMatch involves a machine learning model and controlled search engine queries to classify the URL. A prototype plugin of PhishMatch, developed for the Chrome browser, was found to be fast and lightweight. Our evaluation shows that PhishMatch is both efficient and effective.

外部データセット

PhishTank

DMOZ

PhishTank (Train)

PhishTank (Test)

DMOZ (Train)

DMOZ (Test)

BH-Train-60

PhishMatch (New)

BH-Test-30