June 1, 2017Open Access

Generating Alignments Using Target Foresight in Attention-Based Neural Machine Translation

Puntos clave

Los puntos clave no están disponibles para este artículo en este momento.

Resumen

Abstract Neural machine translation (NMT) has shown large improvements in recent years. The currently most successful approach in this area relies on the attention mechanism, which is often interpreted as an alignment, even though it is computed without explicit knowledge of the target word. This limitation is the most likely reason that the quality of attention-based alignments is inferior to the quality of traditional alignment methods. Guided alignment training has shown that alignments are still capable of improving translation quality. In this work, we propose an extension of the attention-based NMT model that introduces target information into the attention mechanism to produce high-quality alignments. In comparison to the conventional attention-based alignments, our model halves the A er with an absolute improvement of 19.1% A er . Compared to GIZA++ it shows an absolute improvement of 2.0% A er .

Me gusta

Guardar

Ver artículo completo