Pre-training BERT on automatically generated multiple-choice questions from ConceptNet and Wikipedia improves commonsense benchmarks and leaves GLUE performance essentially unchanged.
Exploring Unsupervised Pretraining and Sentence Structure Modelling for Winograd Schema Challenge
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
Winograd Schema Challenge (WSC) was proposed as an AI-hard problem in testing computers' intelligence on common sense representation and reasoning. This paper presents the new state-of-theart on WSC, achieving an accuracy of 71.1%. We demonstrate that the leading performance benefits from jointly modelling sentence structures, utilizing knowledge learned from cutting-edge pretraining models, and performing fine-tuning. We conduct detailed analyses, showing that fine-tuning is critical for achieving the performance, but it helps more on the simpler associative problems. Modelling sentence dependency structures, however, consistently helps on the harder non-associative subset of WSC. Analysis also shows that larger fine-tuning datasets yield better performances, suggesting the potential benefit of future work on annotating more Winograd schema sentences.
fields
cs.CL 1years
2019 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Align, Mask and Select: A Simple Method for Incorporating Commonsense Knowledge into Language Representation Models
Pre-training BERT on automatically generated multiple-choice questions from ConceptNet and Wikipedia improves commonsense benchmarks and leaves GLUE performance essentially unchanged.