DashArena evaluates LLM-generated interactive dashboards by replaying each model's own interaction walkthrough in a browser and judging the evidence with a distilled VLM judge, showing that current frontier models frequently generate non-functional or analytically weak dashboards.
IEEE Transactions on Visualization and Computer Graphics , volume=
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.AI 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
DashArena: Benchmarking LLMs on Interactive Analytic Dashboard Generation
DashArena evaluates LLM-generated interactive dashboards by replaying each model's own interaction walkthrough in a browser and judging the evidence with a distilled VLM judge, showing that current frontier models frequently generate non-functional or analytically weak dashboards.