Interactive Attention for Neural Machine Translation

Meng, Fandong; Lu, Zhengdong; Li, Hang; Liu, Qun

Computer Science > Computation and Language

arXiv:1610.05011 (cs)

[Submitted on 17 Oct 2016]

Title:Interactive Attention for Neural Machine Translation

Authors:Fandong Meng, Zhengdong Lu, Hang Li, Qun Liu

View PDF

Abstract:Conventional attention-based Neural Machine Translation (NMT) conducts dynamic alignment in generating the target sentence. By repeatedly reading the representation of source sentence, which keeps fixed after generated by the encoder (Bahdanau et al., 2015), the attention mechanism has greatly enhanced state-of-the-art NMT. In this paper, we propose a new attention mechanism, called INTERACTIVE ATTENTION, which models the interaction between the decoder and the representation of source sentence during translation by both reading and writing operations. INTERACTIVE ATTENTION can keep track of the interaction history and therefore improve the translation performance. Experiments on NIST Chinese-English translation task show that INTERACTIVE ATTENTION can achieve significant improvements over both the previous attention-based NMT baseline and some state-of-the-art variants of attention-based NMT (i.e., coverage models (Tu et al., 2016)). And neural machine translator with our INTERACTIVE ATTENTION can outperform the open source attention-based NMT system Groundhog by 4.22 BLEU points and the open source phrase-based system Moses by 3.94 BLEU points averagely on multiple test sets.

Comments:	Accepted at COLING 2016
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:1610.05011 [cs.CL]
	(or arXiv:1610.05011v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.1610.05011

Submission history

From: Fandong Meng [view email]
[v1] Mon, 17 Oct 2016 08:33:20 UTC (562 KB)

Computer Science > Computation and Language

Title:Interactive Attention for Neural Machine Translation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Interactive Attention for Neural Machine Translation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators