THESUS: Organizing Web document collections based on link semantics

16 years 6 months ago

Download www.db-net.aueb.gr

Abstract. The requirements for effective search and management of the WWW are stronger than ever. Currently Web documents are classified based on their content not taking into account the fact that these documents are connected to each other by links. We claim that a page's classification is enriched by the detection of its incoming links' semantics. This would enable effective browsing and enhance the validity of search results in the WWW context. Another aspect that is underaddressed and strictly related to the tasks of browsing and searching is the similarity of documents at the semantic level. The above observations lead us to the adoption of a hierarchy of concepts (ontology) and a thesaurus to exploit links and provide a better characterization of Web documents. The enhancement of document characterization makes operations such as clustering and labeling very interesting. To this end, we devised a system called THESUS. The system deals with an initial sets of Web docume...

Maria Halkidi, Benjamin Nguyen, Iraklis Varlamis,

Real-time Traffic

Characterization Makes Operations | Database | Incoming Links | VLDB 2003 | Web Documents |

claim paper

» Measuring Structural Similarity Among Web Documents Preliminary Results

» Conceptual Open Hypermedia The Semantic Web

» Corpus Linguistics for Establishing The Natural Language Content of Digital Library Docume...

» Semantic Web Techniques for Personalization of eGovernment Services

» HealthFinland Finnish Health Information on the Semantic Web

» Generalized SemanticsBased Service Composition

» Web Document Classification Managing Context Change

» VISION a Semantic Web Portal for Describing the Stateoftheart on European Knowledge Manag...

Post Info
More Details (n/a)

Added	05 Dec 2009
Updated	05 Dec 2009
Type	Conference
Year	2003
Where	VLDB
Authors	Maria Halkidi, Benjamin Nguyen, Iraklis Varlamis, Michalis Vazirgiannis

Comments (0)

Sciweavers

THESUS: Organizing Web document collections based on link semantics

Characterization Makes Operations | Database | Incoming Links | VLDB 2003 | Web Documents |

Explore & Download

Productivity Tools

Sciweavers