REVIEW 2 cited by
Colorless green recurrent networks dream hierarchically
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Recurrent neural networks (RNNs) have achieved impressive results in a variety of linguistic processing tasks, suggesting that they can induce non-trivial properties of language. We investigate here to what extent RNNs learn to track abstract hierarchical syntactic structure. We test whether RNNs trained with a generic language modeling objective in four languages (Italian, English, Hebrew, Russian) can predict long-distance number agreement in various constructions. We include in our evaluation nonsensical sentences where RNNs cannot rely on semantic or lexical cues ("The colorless green ideas I ate with the chair sleep furiously"), and, for Italian, we compare model performance to human intuitions. Our language-model-trained RNNs make reliable predictions about long-distance agreement, and do not lag much behind human performance. We thus bring support to the hypothesis that RNNs are not just shallow-pattern extractors, but they also acquire deeper grammatical competence.
Forward citations
Cited by 2 Pith papers
-
Learning from Impairment: Leveraging Insights from Clinical Linguistics in Language Modelling Research
Aphasia treatment protocols like CATE offer complexity hierarchies that the paper proposes to reuse for language model evaluation and curriculum learning, without providing empirical evidence.
-
Higher-order Comparisons of Sentence Encoder Representations
Sentences that take humans longer to read also show larger disagreement between layers of pretrained language encoders.
Discussion (0). Continue with ORCID to comment.