Pith. sign in

REVIEW 1 cited by

SBERT-WK: A Sentence Embedding Method by Dissecting BERT-based Word Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2002.06652 v2 pith:J6VB7UQ4 submitted 2020-02-16 cs.CL cs.LGcs.MM

classification cs.CLcs.LGcs.MM
keywords wordrepresentationsbert-wksentencemodelstasksbert-basedembedding
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Sentence embedding is an important research topic in natural language processing (NLP) since it can transfer knowledge to downstream tasks. Meanwhile, a contextualized word representation, called BERT, achieves the state-of-the-art performance in quite a few NLP tasks. Yet, it is an open problem to generate a high quality sentence representation from BERT-based word models. It was shown in previous study that different layers of BERT capture different linguistic properties. This allows us to fusion information across layers to find better sentence representation. In this work, we study the layer-wise pattern of the word representation of deep contextualized models. Then, we propose a new sentence embedding method by dissecting BERT-based word models through geometric analysis of the space spanned by the word representation. It is called the SBERT-WK method. No further training is required in SBERT-WK. We evaluate SBERT-WK on semantic textual similarity and downstream supervised tasks. Furthermore, ten sentence-level probing tasks are presented for detailed linguistic analysis. Experiments show that SBERT-WK achieves the state-of-the-art performance. Our codes are publicly available.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Semantic Compression for Word and Sentence Embeddings using Discrete Wavelet Transform

    cs.CL 2025-07 conditional novelty 5.0 of 10

    Keeping only the low-frequency DWT coefficients of word and sentence embeddings preserves most of their semantic quality at 50 to 93 percent fewer dimensions.

Pith tools