Sciweavers

Free Online Productivity Tools i2Speak i2Symbol i2OCR iTex2Img iWeb2Print iWeb2Shot i2Type iPdf2Split iPdf2Merge i2Bopomofo i2Arabic i2Style i2Image i2PDF iLatex2Rtf Sci2ools

113

ACL
2003

favoriteEmaildiscussreport

106views Computational Linguistics» more ACL 2003»

Unsupervised Segmentation of Words Using Prior Distributions of Morph Length and Frequency

15 years 3 months ago

Unsupervised Segmentation of Words Using Prior Distributions of Morph Length and Frequency

Download www.aclweb.org

We present a language-independent and unsupervised algorithm for the segmentation of words into morphs. The algorithm is based on a new generative probabilistic model, which makes use of relevant prior information on the length and frequency distributions of morphs in a language. Our algorithm is shown to outperform two competing algorithms, when evaluated on data from a language with agglutinative morphology (Finnish), and to perform well also on English data.

Mathias Creutz

Real-time Traffic

ACL 2003 | ACL 2007 | Generative Probabilistic Model | Relevant Prior Information | Unsupervised Algorithm |

claim paper

Related Content

» Shared Segmentation of Natural Scenes Using Dependent PitmanYor Processes

» Supervised Hierarchical PitmanYor Process for Natural Scene Segmentation

» Genomic DNA kmer Spectra Models and Modalities

» Reconstructing Ancestral Haplotypes with a Dictionary Model

Post Info
More Details (n/a)

Added	31 Oct 2010
Updated	31 Oct 2010
Type	Conference
Year	2003
Where	ACL
Authors	Mathias Creutz

Comments (0)