Pith. sign in

The judge, the ai, and the crown: a collusive network,

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.CL 1

years

2024 1

verdicts

REJECT 1

representative citing papers

Legal Evalutions and Challenges of Large Language Models

cs.CL · 2024-11-15 · reject · novelty 3.0

In a small human-scored evaluation of 10 LLMs on 26 legal cases, o1-preview received the highest overall human score (3.96/5), while ROUGE and BLEU scores did not track human preference.

citing papers explorer

Showing 1 of 1 citing paper.

  • Legal Evalutions and Challenges of Large Language Models cs.CL · 2024-11-15 · reject · none · ref 7

    In a small human-scored evaluation of 10 LLMs on 26 legal cases, o1-preview received the highest overall human score (3.96/5), while ROUGE and BLEU scores did not track human preference.