Pith. sign in

REVIEW

JuriBERT: A Masked-Language Model Adaptation for French Legal Text

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2110.01485 v2 pith:MIGRC5EJ submitted 2021-10-04 cs.CL

classification cs.CL
keywords frenchlegalmodelslanguageadapteddomain-specifictextadaptation
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Language models have proven to be very useful when adapted to specific domains. Nonetheless, little research has been done on the adaptation of domain-specific BERT models in the French language. In this paper, we focus on creating a language model adapted to French legal text with the goal of helping law professionals. We conclude that some specific tasks do not benefit from generic language models pre-trained on large amounts of data. We explore the use of smaller architectures in domain-specific sub-languages and their benefits for French legal text. We prove that domain-specific pre-trained models can perform better than their equivalent generalised ones in the legal domain. Finally, we release JuriBERT, a new set of BERT models adapted to the French legal domain.

Discussion (0). Sign in to comment.

Pith tools