pith. machine review for the scientific record. sign in

arxiv: 1610.06550 · v1 · submitted 2016-10-20 · 💻 cs.CL

Recognition: unknown

Neural Machine Translation with Characters and Hierarchical Encoding

Authors on Pith no claims yet
classification 💻 cs.CL
keywords charactercharactershierarchicaltranslationwordsencoderinputmachine
0
0 comments X
read the original abstract

Most existing Neural Machine Translation models use groups of characters or whole words as their unit of input and output. We propose a model with a hierarchical char2word encoder, that takes individual characters both as input and output. We first argue that this hierarchical representation of the character encoder reduces computational complexity, and show that it improves translation performance. Secondly, by qualitatively studying attention plots from the decoder we find that the model learns to compress common words into a single embedding whereas rare words, such as names and places, are represented character by character.

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.