Title resolution pending

doi: 10 · 2025 · DOI 10.1162/coli.a.14

2 Pith papers cite this work. Polarity classification is still indexing.

2 Pith papers citing it

open at publisher browse 2 citing papers

Title metadata for this work has not finished resolving. The hub is built from the citation graph; the title resolver retries DOI and OpenAlex on its next pass.

representative citing papers

JuICE: A Benchmark for Evaluating LLM-Judge in Identifying Cultural Errors

cs.CL · 2026-05-26 · unverdicted · novelty 6.0

JuICE is a new multilingual benchmark dataset showing top LLM judges reach only F1 0.52 on span-level cultural error detection and miss errors locals readily spot.

Mitigating Cross-Lingual Cultural Inconsistencies in LLMs via Consensus-Driven Preference Optimisation

cs.CL · 2026-04-02

citing papers explorer

Showing 1 of 1 citing paper after filters.

JuICE: A Benchmark for Evaluating LLM-Judge in Identifying Cultural Errors cs.CL · 2026-05-26 · unverdicted · none · ref 31
JuICE is a new multilingual benchmark dataset showing top LLM judges reach only F1 0.52 on span-level cultural error detection and miss errors locals readily spot.

Title resolution pending

fields

years

verdicts

representative citing papers

citing papers explorer