An ensemble of expert medical LLMs with triage and weighted consensus reports accuracy gains over single frontier models on medical QA benchmarks.
Evaluating large language mod- els for use in healthcare: A framework for translational value assessment
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.AI 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Second Opinion Matters: Towards Adaptive Clinical AI via the Consensus of Expert Model Ensemble
An ensemble of expert medical LLMs with triage and weighted consensus reports accuracy gains over single frontier models on medical QA benchmarks.