Title resolution pending

Raoyuan Zhao, Beiduo Chen, Barbara Plank, Michael A · 2025 · DOI 10.18653/v1/2025.findings-emnlp.1256

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

open at publisher browse 1 citing papers

Title metadata for this work has not finished resolving. The hub is built from the citation graph; the title resolver retries DOI and OpenAlex on its next pass.

representative citing papers

JuICE: A Benchmark for Evaluating LLM-Judge in Identifying Cultural Errors

cs.CL · 2026-05-26 · unverdicted · novelty 6.0

JuICE is a new multilingual benchmark dataset showing top LLM judges reach only F1 0.52 on span-level cultural error detection and miss errors locals readily spot.

citing papers explorer

Showing 1 of 1 citing paper after filters.

JuICE: A Benchmark for Evaluating LLM-Judge in Identifying Cultural Errors cs.CL · 2026-05-26 · unverdicted · none · ref 41
JuICE is a new multilingual benchmark dataset showing top LLM judges reach only F1 0.52 on span-level cultural error detection and miss errors locals readily spot.

Title resolution pending

fields

years

verdicts

representative citing papers

citing papers explorer