The paper releases the first public, rule-complete benchmark for LLM reasoning over NordDRG hospital payment logic, with top models scoring 13/13 on logic tasks and 7/13 on full grouper emulation.
Large Language Model in Medical Informatics: Direct Classification and Enhanced Text Representations for Automatic ICD Coding
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
Addressing the complexity of accurately classifying International Classification of Diseases (ICD) codes from medical discharge summaries is challenging due to the intricate nature of medical documentation. This paper explores the use of Large Language Models (LLM), specifically the LLAMA architecture, to enhance ICD code classification through two methodologies: direct application as a classifier and as a generator of enriched text representations within a Multi-Filter Residual Convolutional Neural Network (MultiResCNN) framework. We evaluate these methods by comparing them against state-of-the-art approaches, revealing LLAMA's potential to significantly improve classification outcomes by providing deep contextual insights into medical texts.
fields
cs.AI 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
The NordDRG AI Benchmark for Large Language Models
The paper releases the first public, rule-complete benchmark for LLM reasoning over NordDRG hospital payment logic, with top models scoring 13/13 on logic tasks and 7/13 on full grouper emulation.