Vision-language models vary widely in how trustworthy their confidence scores are on document extraction, with stronger models and OCR-plus-image input helping most, as measured on the new ConfBench benchmark.
Title resolution pending
1 Pith paper cite this work, alongside 36 external citations. Polarity classification is still indexing.
1
Pith paper citing it
36
external citations · OpenAlex
fields
cs.AI 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Can You Trust the Confidence? ConfBench for Vision-Language Models on Document Extraction
Vision-language models vary widely in how trustworthy their confidence scores are on document extraction, with stronger models and OCR-plus-image input helping most, as measured on the new ConfBench benchmark.