An ensemble of prompted GPT-4o outputs scores F1=0.95 on medical entity classification in EHR text, but only after excluding the 63 percent of gold entities that the extraction step missed.
The evolving use of electronic health records (ehr) for research,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.AI 1years
2025 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
LLM-based Prompt Ensemble for Reliable Medical Entity Recognition from EHRs
An ensemble of prompted GPT-4o outputs scores F1=0.95 on medical entity classification in EHR text, but only after excluding the 63 percent of gold entities that the extraction step missed.