The authors propose the Flourishing AI Benchmark, which uses 1,229 objective and subjective questions plus LLM judges to score 28 chatbots across seven dimensions of human flourishing, and find none reach the 90-point threshold.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.AI 1years
2025 1verdicts
UNVERDICTED 1representative citing papers
citing papers explorer
-
Measuring AI Alignment with Human Flourishing
The authors propose the Flourishing AI Benchmark, which uses 1,229 objective and subjective questions plus LLM judges to score 28 chatbots across seven dimensions of human flourishing, and find none reach the 90-point threshold.