Sciweavers

APWEB
2003
Springer

Mining "Hidden Phrase" Definitions from the Web

14 years 2 months ago
Mining "Hidden Phrase" Definitions from the Web
Keyword searching is the most common form of document search on the Web. Many Web publishers manually annotate the META tags and titles of their pages with frequently queried phrases in order to improve their placement and ranking. A "hidden phrase" is defined as a phrase that occurs in the META tag of a Web page but not in its body. In this paper we present an algorithm that mines the definitions of hidden phrases from the Web documents. Phrase definitions allow (i) publishers to find relevant phrases with high query frequency, and, (ii) search engines to test if the content of the body of a document matches the phrases. We use cooccurrence clustering and association rule mining algorithms to learn phrase definitions from high-dimensional data sets. We also provide experimental results.
Hung V. Nguyen, P. Velamuru, Deepak Kolippakkam, H
Added 23 Aug 2010
Updated 23 Aug 2010
Type Conference
Year 2003
Where APWEB
Authors Hung V. Nguyen, P. Velamuru, Deepak Kolippakkam, Hasan Davulcu, Huan Liu, M. Ates
Comments (0)