REVIEW 3 cited by
Insertion-based Decoding with automatically Inferred Generation Order
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Conventional neural autoregressive decoding commonly assumes a fixed left-to-right generation order, which may be sub-optimal. In this work, we propose a novel decoding algorithm -- InDIGO -- which supports flexible sequence generation in arbitrary orders through insertion operations. We extend Transformer, a state-of-the-art sequence generation model, to efficiently implement the proposed approach, enabling it to be trained with either a pre-defined generation order or adaptive orders obtained from beam-search. Experiments on four real-world tasks, including word order recovery, machine translation, image caption and code generation, demonstrate that our algorithm can generate sequences following arbitrary orders, while achieving competitive or even better performance compared to the conventional left-to-right generation. The generated sequences show that InDIGO adopts adaptive generation orders based on input information.
Forward citations
Cited by 3 Pith papers
-
FlowSeq: Non-Autoregressive Conditional Sequence Generation with Generative Flow
A flow-based latent variable model enables non-autoregressive neural machine translation with parallel decoding and near-constant time, reaching BLEU scores comparable to state-of-the-art non-autoregressive systems.
-
Latent-Variable Non-Autoregressive Neural Machine Translation with Deterministic Inference Using a Delta Posterior
A latent-variable non-autoregressive translation model with deterministic delta-posterior inference matches autoregressive quality within 2 BLEU points while decoding 12.5x faster.
-
Attending to Future Tokens For Bidirectional Sequence Generation
BISON uses placeholder tokens in a bidirectional Transformer to generate sequences, and fine-tuning BERT with this scheme beats GPT2 on two dialogue tasks.
Discussion (0). Continue with ORCID to comment.