Showing 2017Show all
3 papers · 1 filter
cs.CL2017
Robust Tuning Datasets for Statistical Machine Translation
Preslav Nakov, Stephan Vogel
We explore the idea of automatically crafting a tuning dataset for Statistical Machine Translation (SMT) that makes the hyper-parameters of the SMT system more robust with respect…
cs.CL2017
Speech Recognition Challenge in the Wild: Arabic MGB-3
Ahmed Ali, Stephan Vogel, Steve Renals
This paper describes the Arabic MGB-3 Challenge - Arabic Speech Recognition in the Wild. Unlike last year's Arabic MGB-2 Challenge, for which the recognition task was based on more…
cs.CL2017
Challenging Language-Dependent Segmentation for Arabic: An Application to Machine Translation and Part-of-Speech Tagging
Hassan Sajjad, Fahim Dalvi, Nadir Durrani +3
Word segmentation plays a pivotal role in improving any Arabic NLP application. Therefore, a lot of research has been spent in improving its accuracy. Off-the-shelf tools, however,…