Pith. sign in

REVIEW 1 cited by

Evaluating Structural Generalization in Neural Machine Translation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2406.13363 v1 pith:SO3UPHOQ submitted 2024-06-19 cs.CL

Evaluating Structural Generalization in Neural Machine Translation

classification cs.CL
keywords generalizationmachinetranslationcompositionalmodelsneuralstructuralstructures
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Compositional generalization refers to the ability to generalize to novel combinations of previously observed words and syntactic structures. Since it is regarded as a desired property of neural models, recent work has assessed compositional generalization in machine translation as well as semantic parsing. However, previous evaluations with machine translation have focused mostly on lexical generalization (i.e., generalization to unseen combinations of known words). Thus, it remains unclear to what extent models can translate sentences that require structural generalization (i.e., generalization to different sorts of syntactic structures). To address this question, we construct SGET, a machine translation dataset covering various types of compositional generalization with control of words and sentence structures. We evaluate neural machine translation models on SGET and show that they struggle more in structural generalization than in lexical generalization. We also find different performance trends in semantic parsing and machine translation, which indicates the importance of evaluations across various tasks.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Names Don't Matter: Symbol-Invariant Transformer for Open-Vocabulary Learning

    cs.LG 2026-01 conditional novelty 5.0

    A shared-parameter per-symbol stream design makes Transformers exactly invariant to symbol renaming and lets them handle unseen symbols at test time, with large gains on propositional-logic generalization and small on...