Current browse context:
cs.CL
Change to browse by:
References & Citations
Computer Science > Computation and Language
Title: Morphology Generation for Statistical Machine Translation using Deep Learning Techniques
(Submitted on 7 Oct 2016 (v1), last revised 6 Feb 2017 (this version, v2))
Abstract: Morphology in unbalanced languages remains a big challenge in the context of machine translation. In this paper, we propose to de-couple machine translation from morphology generation in order to better deal with the problem. We investigate the morphology simplification with a reasonable trade-off between expected gain and generation complexity. For the Chinese-Spanish task, optimum morphological simplification is in gender and number. For this purpose, we design a new classification architecture which, compared to other standard machine learning techniques, obtains the best results. This proposed neural-based architecture consists of several layers: an embedding, a convolutional followed by a recurrent neural network and, finally, ends with sigmoid and softmax layers. We obtain classification results over 98% accuracy in gender classification, over 93% in number classification, and an overall translation improvement of 0.7 METEOR.
Submission history
From: Marta R. Costa-jussà [view email][v1] Fri, 7 Oct 2016 09:59:13 GMT (302kb,D)
[v2] Mon, 6 Feb 2017 15:15:40 GMT (933kb)
Link back to: arXiv, form interface, contact.