A new French financial document benchmark shows vision-language models are strong at text/table extraction but brittle on charts and multi-turn dialogue, with accuracy converging near 50% in conversational settings.
Title resolution pending
1 Pith paper cite this work, alongside 92 external citations. Polarity classification is still indexing.
1
Pith paper citing it
92
external citations · OpenAlex
fields
cs.CL 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
When Tables Go Crazy: Evaluating Multimodal Models on French Financial Documents
A new French financial document benchmark shows vision-language models are strong at text/table extraction but brittle on charts and multi-turn dialogue, with accuracy converging near 50% in conversational settings.