Exploiting Out-of-Domain Data Sources for Dialectal Arabic Statistical Machine Translation

CoRR, Volume abs/1509.01938, 2015.

Cited by: 0|Bibtex|Views0|Links
EI

Abstract:

Statistical machine translation for dialectal Arabic is characterized by a lack of data since data acquisition involves the transcription and translation of spoken language. In this study we develop techniques for extracting parallel data for one particular dialect of Arabic (Iraqi Arabic) from out-of-domain corpora in different dialect...More

Code:

Data:

Your rating :
0

 

Tags
Comments