Pith. sign in

REVIEW 1 cited by

Hybrid Self-Attention Network for Machine Translation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1811.00253 v3 pith:2GNVIOGQ submitted 2018-11-01 cs.CL

classification cs.CL
keywords self-attentiontranslationdifferentmachinenetworkcannotcurrentframework
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The encoder-decoder is the typical framework for Neural Machine Translation (NMT), and different structures have been developed for improving the translation performance. Transformer is one of the most promising structures, which can leverage the self-attention mechanism to capture the semantic dependency from global view. However, it cannot distinguish the relative position of different tokens very well, such as the tokens located at the left or right of the current token, and cannot focus on the local information around the current token either. To alleviate these problems, we propose a novel attention mechanism named Hybrid Self-Attention Network (HySAN) which accommodates some specific-designed masks for self-attention network to extract various semantic, such as the global/local information, the left/right part context. Finally, a squeeze gate is introduced to combine different kinds of SANs for fusion. Experimental results on three machine translation tasks show that our proposed framework outperforms the Transformer baseline significantly and achieves superior results over state-of-the-art NMT systems.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Hard but Robust, Easy but Sensitive: How Encoder and Decoder Perform in Neural Machine Translation

    cs.CL 2019-08 conditional novelty 6.0 of 10

    In neural machine translation, the decoder handles an easier but more noise-sensitive task than the encoder, because it depends strongly on the immediately preceding output words.

Pith tools