Sciweavers

EMNLP
2009

Segmenting Email Message Text into Zones

13 years 10 months ago
Segmenting Email Message Text into Zones
In the early days of email, widely-used conventions for indicating quoted reply content and email signatures made it easy to segment email messages into their functional parts. Today, the explosion of different email formats and styles, coupled with the ad hoc ways in which people vary the structure and layout of their messages, means that simple techniques for identifying quoted replies that used to yield 95% accuracy now find less than 10% of such content. In this paper, we describe Zebra, an SVM-based system for segmenting the body text of email messages into nine zone types based on graphic, orthographic and lexical cues. Zebra performs this task with an accuracy of 87.01%; when the numones is abstracted to two or three zone classes, this increases to 93.60% and
Andrew Lampert, Robert Dale, Cécile Paris
Added 17 Feb 2011
Updated 17 Feb 2011
Type Journal
Year 2009
Where EMNLP
Authors Andrew Lampert, Robert Dale, Cécile Paris
Comments (0)