On 6,990 ophthalmology MCQs, OpenAI o1 beat five other LLMs on answer accuracy yet lagged GPT-4o and GPT-4 on text-similarity metrics used to gauge reasoning.
Large language models and their impact in ophthalmology
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Can OpenAI o1 Reason Well in Ophthalmology? A 6,990-Question Head-to-Head Evaluation Study
On 6,990 ophthalmology MCQs, OpenAI o1 beat five other LLMs on answer accuracy yet lagged GPT-4o and GPT-4 on text-similarity metrics used to gauge reasoning.