Pith. sign in

REVIEW 1 cited by

A Medical Information Extraction Workbench to Process German Clinical Text

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2207.03885 v2 pith:HI7ODZWU submitted 2022-07-08 cs.CL

classification cs.CL
keywords germanmodelstextclinicalprocessingworkbenchaccessibleavailable
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Background: In the information extraction and natural language processing domain, accessible datasets are crucial to reproduce and compare results. Publicly available implementations and tools can serve as benchmark and facilitate the development of more complex applications. However, in the context of clinical text processing the number of accessible datasets is scarce -- and so is the number of existing tools. One of the main reasons is the sensitivity of the data. This problem is even more evident for non-English languages. Approach: In order to address this situation, we introduce a workbench: a collection of German clinical text processing models. The models are trained on a de-identified corpus of German nephrology reports. Result: The presented models provide promising results on in-domain data. Moreover, we show that our models can be also successfully applied to other biomedical text in German. Our workbench is made publicly available so it can be used out of the box, as a benchmark or transferred to related problems.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. ELMTEX: Fine-Tuning Large Language Models for Structured Clinical Information Extraction. A Case Study on Clinical Reports

    cs.CL 2025-02 conditional novelty 5.0 of 10

    Fine-tuned small Llama models (1B to 8B) outperform Llama 405B with prompting on a new 60k English and 24k German clinical-summary extraction benchmark.

Pith tools