Automating the Construction of Internet Portals with Machine Learning

15 years 6 months ago

Download www.kamalnigam.com

Domain-specific internet portals are growing in popularity because they gather content from the Web and organize it for easy access, retrieval and search. For example, www.campsearch.com allows complex queries by age, location, cost and specialty over summer camps. This functionality is not possible with general, Web-wide search engines. Unfortunately these portals are difficult and time-consuming to maintain. This paper advocates the use of machine learning techniques to greatly automate the creation and maintenance of domain-specific Internet portals. We describe new research in reinforcement learning, information extraction and text classification that enables efficient spidering, the identification of informative text segments, and the population of topic hierarchies. Using these techniques, we have built a demonstration system: a portal for computer science research papers. It already contains over 50,000 papers and is publicly available at www.cora.justresearch.com. These techniq...

Andrew McCallum, Kamal Nigam, Jason Rennie, Kristi

Real-time Traffic

Domain-specific Internet Portals | Enables Efficient Spidering | IR 2000 | Natural Language Processing | Web-wide Search Engines |

claim paper

» ACAS automated construction of application signatures

» Automated construction of web accessibility models from transaction clickstreams

» Machine Learning Methods for OneSession Ahead Prediction of Accesses to Page Categories

» Discovering user communities on the Internet using unsupervised machine learning technique...

» SEMPL a semantic portal

» Design and implementation of contextual information portals

» Automated Traffic Classification and Application Identification using Machine Learning

» Automating StandardsBased Courseware Development Using UML

Post Info
More Details (n/a)

Added	18 Dec 2010
Updated	18 Dec 2010
Type	Journal
Year	2000
Where	IR
Authors	Andrew McCallum, Kamal Nigam, Jason Rennie, Kristie Seymore

Comments (0)

Sciweavers

Automating the Construction of Internet Portals with Machine Learning

Domain-specific Internet Portals | Enables Efficient Spidering | IR 2000 | Natural Language Processing | Web-wide Search Engines |

Explore & Download

Productivity Tools

Sciweavers