Sciweavers

ACL
2006

Combination of Arabic Preprocessing Schemes for Statistical Machine Translation

14 years 1 months ago
Combination of Arabic Preprocessing Schemes for Statistical Machine Translation
Statistical machine translation is quite robust when it comes to the choice of input representation. It only requires consistency between training and testing. As a result, there is a wide range of possible preprocessing choices for data used in statistical machine translation. This is even more so for morphologically rich languages such as Arabic. In this paper, we study the effect of different word-level preprocessing schemes for Arabic on the quality of phrase-based statistical machine translation. We also present and evaluate different methods for combining preprocessing schemes resulting in improved translation quality.
Fatiha Sadat, Nizar Habash
Added 30 Oct 2010
Updated 30 Oct 2010
Type Conference
Year 2006
Where ACL
Authors Fatiha Sadat, Nizar Habash
Comments (0)