CamemBERTv2 and CamemBERTav2, trained on 275B tokens with an improved tokenizer, outperform prior French encoder models on several downstream tasks.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
CamemBERT 2.0: A Smarter French Language Model Aged to Perfection
CamemBERTv2 and CamemBERTav2, trained on 275B tokens with an improved tokenizer, outperform prior French encoder models on several downstream tasks.