BanglaEmbed-MSE, a 66M-parameter model trained via cross-lingual distillation from English sentence embeddings, outperforms existing Bangla sentence transformers on paraphrase detection and semantic textual similarity benchmarks.
Progress in machine translation,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
BanglaEmbed: Efficient Sentence Embedding Models for a Low-Resource Language Using Cross-Lingual Distillation Techniques
BanglaEmbed-MSE, a 66M-parameter model trained via cross-lingual distillation from English sentence embeddings, outperforms existing Bangla sentence transformers on paraphrase detection and semantic textual similarity benchmarks.