Fast Neural Machine Translation Implementation

Daniel Torregrosa; Hieu Hoang; Kenneth Heafield; Rihards Krislauks; Tomasz Dwojak

arxiv: 1805.09863 · v3 · pith:A7WUEQLDnew · submitted 2018-05-24 · 💻 cs.CL

Fast Neural Machine Translation Implementation

Hieu Hoang , Tomasz Dwojak , Rihards Krislauks , Daniel Torregrosa , Kenneth Heafield This is my paper

classification 💻 cs.CL

keywords machineneuraltranslationuniversityalgorithmamunefficiencyefficient

0 comments

read the original abstract

This paper describes the submissions to the efficiency track for GPUs at the Workshop for Neural Machine Translation and Generation by members of the University of Edinburgh, Adam Mickiewicz University, Tilde and University of Alicante. We focus on efficient implementation of the recurrent deep-learning model as implemented in Amun, the fast inference engine for neural machine translation. We improve the performance with an efficient mini-batching algorithm, and by fusing the softmax operation with the k-best extraction algorithm. Submissions using Amun were first, second and third fastest in the GPU efficiency track.

This paper has not been read by Pith yet.

Fast Neural Machine Translation Implementation

discussion (0)