LLMs exhibit data referencing errors across model sizes; a critic model detects them at 78.2% F1 and boosts accuracy up to 12% via filtering and rejection sampling.
all teams with more than 5 wins
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2026 1verdicts
UNVERDICTED 1representative citing papers
citing papers explorer
-
When LLMs Read Tables Carelessly: Measuring and Reducing Data Referencing Errors
LLMs exhibit data referencing errors across model sizes; a critic model detects them at 78.2% F1 and boosts accuracy up to 12% via filtering and rejection sampling.