Learning to Reuse Translations: Guiding Neural Machine Translation with Examples

Cao, Qian; Kuang, Shaohui; Xiong, Deyi

Full-text links:

Download:

Current browse context:

cs.CL

< prev | next >

new | recent | 1911

Change to browse by:

Computer Science > Computation and Language

Title: Learning to Reuse Translations: Guiding Neural Machine Translation with Examples

Authors: Qian Cao, Shaohui Kuang, Deyi Xiong

(Submitted on 25 Nov 2019 (v1), last revised 28 Nov 2019 (this version, v2))

Abstract: In this paper, we study the problem of enabling neural machine translation (NMT) to reuse previous translations from similar examples in target prediction. Distinguishing reusable translations from noisy segments and learning to reuse them in NMT are non-trivial. To solve these challenges, we propose an Example-Guided NMT (EGNMT) framework with two models: (1) a noise-masked encoder model that masks out noisy words according to word alignments and encodes the noise-masked sentences with an additional example encoder and (2) an auxiliary decoder model that predicts reusable words via an auxiliary decoder sharing parameters with the primary decoder. We define and implement the two models with the state-of-the-art Transformer. Experiments show that the noise-masked encoder model allows NMT to learn useful information from examples with low fuzzy match scores (FMS) while the auxiliary decoder model is good for high-FMS examples. More experiments on Chinese-English, English-German and English-Spanish translation demonstrate that the combination of the two EGNMT models can achieve improvements of up to +9 BLEU points over the baseline system and +7 BLEU points over a two-encoder Transformer.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:1911.10732 [cs.CL]
	(or arXiv:1911.10732v2 [cs.CL] for this version)

Submission history

From: Qian Cao [view email]
[v1] Mon, 25 Nov 2019 07:22:47 GMT (565kb,D)
[v2] Thu, 28 Nov 2019 03:19:33 GMT (170kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:1911.10732

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Computation and Language

Title: Learning to Reuse Translations: Guiding Neural Machine Translation with Examples

Submission history