REVIEW 2 cited by
Making Transformers Solve Compositional Tasks
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Several studies have reported the inability of Transformer models to generalize compositionally, a key type of generalization in many NLP tasks such as semantic parsing. In this paper we explore the design space of Transformer models showing that the inductive biases given to the model by several design decisions significantly impact compositional generalization. Through this exploration, we identified Transformer configurations that generalize compositionally significantly better than previously reported in the literature in a diverse set of compositional tasks, and that achieve state-of-the-art results in a semantic parsing compositional generalization benchmark (COGS), and a string edit operation composition benchmark (PCFG).
Forward citations
Cited by 2 Pith papers
-
Generalized Locomotion in Out-of-distribution Conditions with Robust Transformer
A transformer with body tokenization and consistent dropout generalizes to unseen leg damages and sensor noise while trained on limited dynamics and clean observations.
-
Decomposing Complex Visual Comprehension into Atomic Visual Skills for Vision Language Models
VLMs score far below adult humans on a new 13,188-question benchmark of 36 atomic 2D geometry perception skills.
Discussion (0). Sign in to comment.