Pith. sign in

REVIEW 1 cited by

Does BERT Make Any Sense? Interpretable Word Sense Disambiguation with Contextualized Embeddings

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1909.10430 v2 pith:G77G75YM submitted 2019-09-23 cs.CL

classification cs.CL
keywords wordsensebertembeddingsclassificationcontextcontextualizedcwes
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Contextualized word embeddings (CWE) such as provided by ELMo (Peters et al., 2018), Flair NLP (Akbik et al., 2018), or BERT (Devlin et al., 2019) are a major recent innovation in NLP. CWEs provide semantic vector representations of words depending on their respective context. Their advantage over static word embeddings has been shown for a number of tasks, such as text classification, sequence tagging, or machine translation. Since vectors of the same word type can vary depending on the respective context, they implicitly provide a model for word sense disambiguation (WSD). We introduce a simple but effective approach to WSD using a nearest neighbor classification on CWEs. We compare the performance of different CWE models for the task and can report improvements above the current state of the art for two standard WSD benchmark datasets. We further show that the pre-trained BERT model is able to place polysemic words into distinct 'sense' regions of the embedding space, while ELMo and Flair NLP do not seem to possess this ability.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. "Dialogue" vs "Dialog" in NLP and AI research: Statistics from a Confused Discourse

    cs.CL 2024-12 conditional novelty 6.0 of 10

    Analysis of tens of thousands of papers shows NLP/AI research mixes 'dialogue' and 'dialog' with no clear trend, author, or context explanation.

Pith tools