Four LLMs scored OSCE transcripts with low exact agreement (27-44%) but moderate to high broad-band agreement (75-88%) with expert consensus.
Journal of Graduate Medical Education 9(5), 645–649 (2017)
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Benchmarking Generative AI for Scoring Medical Student Interviews in Objective Structured Clinical Examinations (OSCEs)
Four LLMs scored OSCE transcripts with low exact agreement (27-44%) but moderate to high broad-band agreement (75-88%) with expert consensus.