A hybrid phish detection approach by identity discovery and keywords retrieval

16 years 1 months ago

Download www2009.eprints.org

Phishing is a significant security threat to the Internet, which causes tremendous economic loss every year. In this paper, we proposed a novel hybrid phish detection method based on information extraction (IE) and information retrieval (IR) techniques. The identity-based component of our method detects phishing webpages by directly discovering the inconsistency between their identity and the identity they are imitating. The keywords-retrieval component utilizes IR algorithms exploiting the power of search engines to identify phish. Our method requires no training data, no prior knowledge of phishing signatures and specific implementations, and thus is able to adapt quickly to constantly appearing new phishing patterns. Comprehensive experiments over a diverse spectrum of data sources with 11449 pages show that both components have a low false positive rate and the stacked approach achieves a true positive rate

Guang Xiang, Jason I. Hong

Real-time Traffic

Component Utilizes Ir | False Positive Rate | Internet Technology | Phish Detection Method | WWW 2009 |

claim paper

Post Info
More Details (n/a)

Added	21 Nov 2009
Updated	21 Nov 2009
Type	Conference
Year	2009
Where	WWW
Authors	Guang Xiang, Jason I. Hong

Comments (0)

Sciweavers

A hybrid phish detection approach by identity discovery and keywords retrieval

Component Utilizes Ir | False Positive Rate | Internet Technology | Phish Detection Method | WWW 2009 |

Explore & Download

Productivity Tools

Document Tools

Image Tools

Sciweavers