Neural Machine Translation of Basque

Autor: Etchegoyhen, Thierry, Martínez Garcia, Eva, Azpeitia, Andoni, Labaka Intxauspe, Gorka, Alegría Loinaz, Iñaki, Cortés Etxabe, Itziar, Jauregi Carrera, Amaia, Ellakuria, Igor, Martin, Maite, Calonge, Eusebi
Rok vydání: 2018
Předmět:
Zdroj: RUA. Repositorio Institucional de la Universidad de Alicante
Universidad de Alicante (UA)
Popis: We describe the first experimental results in neural machine translation for Basque. As a synthetic language featuring agglutinative morphology, an extended case system, complex verbal morphology and relatively free word order, Basque presents a large number of challenging characteristics for machine translation in general, and for data-driven approaches such as attention-based encoder-decoder models in particular. We present our results on a large range of experiments in Basque-Spanish translation, comparing several neural machine translation system variants with both rule-based and statistical machine translation systems. We demonstrate that significant gains can be obtained with a neural network approach for this challenging language pair, and describe optimal configurations in terms of word segmentation and decoding parameters, measured against test sets that feature multiple references to account for word order variability. This work was supported by the Department of Economic Development and Competitiveness of the Basque Government via the MODELA project.
Databáze: OpenAIRE