Learning to cluster web search results

15 years 12 months ago

Download research.microsoft.com

Organizing Web search results into clusters facilitates users' quick browsing through search results. Traditional clustering techniques are inadequate since they don't generate clusters with highly readable names. In this paper, we reformalize the clustering problem as a salient phrase ranking problem. Given a query and the ranked list of documents (typically a list of titles and snippets) returned by a certain Web search engine, our method first extracts and ranks salient phrases as candidate cluster names, based on a regression model learned from human labeled training data. The documents are assigned to relevant salient phrases to form candidate clusters, and the final clusters are generated by merging these candidate clusters. Experimental results verify our method's feasibility and effectiveness. Categories and Subject Descriptors H.3.3 [Information Storage and Retrieval]: Information Search and Retrieval - Search process, Clustering, Selection process; G.3 [Probab...

Hua-Jun Zeng, Qi-Cai He, Zheng Chen, Wei-Ying Ma,

Real-time Traffic

Candidate Clusters | Salient Phrases | SIGIR 2004 | Web Search |

claim paper

Related Content

» Grouping web image search result

» IGroup a web image search engine with semantic clustering of search results

» Clustering web search results using fuzzy ants

» Personal Name Disambiguation in Web Search Results Based on a Semisupervised Clustering Ap...

» Clustering web people search results using fuzzy ants

» STC and NMSTC Two Novel Online Results Clustering Methods for Web Searching

» Semantic Hierarchical Online Clustering of Web Search Results

» Link Based Clustering of Web Search Results

» Learning to aggregate vertical results into web search results

Post Info
More Details (n/a)

Added	30 Jun 2010
Updated	30 Jun 2010
Type	Conference
Year	2004
Where	SIGIR
Authors	Hua-Jun Zeng, Qi-Cai He, Zheng Chen, Wei-Ying Ma, Jinwen Ma

Comments (0)

Sciweavers

Learning to cluster web search results

Candidate Clusters | Salient Phrases | SIGIR 2004 | Web Search |

Explore & Download

Productivity Tools

Sciweavers