Sciweavers

Free Online Productivity Tools i2Speak i2Symbol i2OCR iTex2Img iWeb2Print iWeb2Shot i2Type iPdf2Split iPdf2Merge i2Bopomofo i2Arabic i2Style i2Image i2PDF iLatex2Rtf Sci2ools

111

ICWE
2007
Springer

favoriteEmaildiscussreport

114views Internet Technology» more ICWE 2007»

Fixing Weakly Annotated Web Data Using Relational Models

15 years 8 months ago

Fixing Weakly Annotated Web Data Using Relational Models

Download www.public.asu.edu

In this paper, we present a fast and scalable Bayesian model for improving weakly annotated data – which is typically generated by a (semi) automated information extraction (IE) system from Web documents. Weakly annotated data suﬀers from two major problems: they (i) might contain incorrect ontological role assignments, and (ii) might have many missing attributes. Our experimental evaluations with the TAP and RoadRunner data sets, and a collection of 20,000 home pages from university, shopping and sports Web sites, indicate that the model described here can improve the accuracy of role assignments from 40% to 85% for template driven sites, from 68% to 87% for non-template driven sites. The Bayesian model is also shown to be useful for improving the performance of IE systems by informing them with additional domain information.

Fatih Gelgi, Srinivas Vadrevu, Hasan Davulcu

Real-time Traffic

Bayesian Model | ICWE 2007 | Role Assignments | Scalable Bayesian Model |

claim paper

Related Content

» Using Web annotations for asynchronous collaboration around documents

» KnowledgeBased Weak Supervision for Information Extraction of Overlapping Relations

» MultiLevel Active Prediction of Useful Image Annotations for Recognition

» Annotating and Searching Web Tables Using Entities Types and Relationships

» Building Web Annotation Stickies based on Bidirectional Links

» Contentsensitive User Interfaces for Annotated Web Pages

» Webassisted annotation semantic indexing and search of television and radio news

» From XML Schema to Relations A CostBased Approach to XML Storage

» Sample Eigenvalue Based Detection of HighDimensional Signals in White Noise Using Relative...

Post Info
More Details (n/a)

Added	08 Jun 2010
Updated	08 Jun 2010
Type	Conference
Year	2007
Where	ICWE
Authors	Fatih Gelgi, Srinivas Vadrevu, Hasan Davulcu

Comments (0)