Pith. sign in

REVIEW

Self-Adaptive Reconstruction with Contrastive Learning for Unsupervised Sentence Embeddings

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.15153 v1 pith:RCUNTUW3 submitted 2024-02-23 cs.CL cs.LG

classification cs.CLcs.LG
keywords sentenceembeddingsmodelsreconstructionself-adaptivesentencesbiascontrastive
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Unsupervised sentence embeddings task aims to convert sentences to semantic vector representations. Most previous works directly use the sentence representations derived from pretrained language models. However, due to the token bias in pretrained language models, the models can not capture the fine-grained semantics in sentences, which leads to poor predictions. To address this issue, we propose a novel Self-Adaptive Reconstruction Contrastive Sentence Embeddings (SARCSE) framework, which reconstructs all tokens in sentences with an AutoEncoder to help the model to preserve more fine-grained semantics during tokens aggregating. In addition, we proposed a self-adaptive reconstruction loss to alleviate the token bias towards frequency. Experimental results show that SARCSE gains significant improvements compared with the strong baseline SimCSE on the 7 STS tasks.

Discussion (0). Sign in to comment.

Pith tools