REVIEW 3 cited by
Yes, BM25 is a Strong Baseline for Legal Case Retrieval
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We describe our single submission to task 1 of COLIEE 2021. Our vanilla BM25 got second place, well above the median of submissions. Code is available at https://github.com/neuralmind-ai/coliee.
Forward citations
Cited by 3 Pith papers
-
BriefMe: A Legal NLP Benchmark for Assisting with Legal Briefs
BriefMe introduces a legal brief benchmark with argument summarization, argument completion, and case retrieval, and shows LLMs beat human headings on the first two but struggle on the latter two.
-
Assessing the Performance Gap Between Lexical and Semantic Models for Information Retrieval With Formulaic Legal Language
On CJEU paragraph retrieval, BM25 beats off-the-shelf dense models on most metrics, fine-tuned dense models beat BM25, and BM25 wins mainly when queries have less verbatim overlap with the target.
-
ASP2LJ : An Adversarial Self-Play Laywer Augmented Legal Judgment Framework
ASP2LJ combines synthetic case generation with adversarial self-play for lawyer agents, improving legal judgment prediction on a Chinese benchmark and on a new rare-case dataset.
Discussion (0). Continue with ORCID to comment.