REVIEW 6 cited by
Improving Clinical Note Generation from Complex Doctor-Patient Conversation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Writing clinical notes and documenting medical exams is a critical task for healthcare professionals, serving as a vital component of patient care documentation. However, manually writing these notes is time-consuming and can impact the amount of time clinicians can spend on direct patient interaction and other tasks. Consequently, the development of automated clinical note generation systems has emerged as a clinically meaningful area of research within AI for health. In this paper, we present three key contributions to the field of clinical note generation using large language models (LLMs). First, we introduce CliniKnote, a comprehensive dataset consisting of 1,200 complex doctor-patient conversations paired with their full clinical notes. This dataset, created and curated by medical experts with the help of modern neural networks, provides a valuable resource for training and evaluating models in clinical note generation tasks. Second, we propose the K-SOAP (Keyword, Subjective, Objective, Assessment, and Plan) note format, which enhances traditional SOAP~\cite{podder2023soap} (Subjective, Objective, Assessment, and Plan) notes by adding a keyword section at the top, allowing for quick identification of essential information. Third, we develop an automatic pipeline to generate K-SOAP notes from doctor-patient conversations and benchmark various modern LLMs using various metrics. Our results demonstrate significant improvements in efficiency and performance compared to standard LLM finetuning methods.
Forward citations
Cited by 6 Pith papers
-
Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting
In a retrospective sample of 500 urgent-care telehealth visits, a proprietary AI doctor matched clinicians' top diagnosis 81% of the time and treatment plans 99.2% of the time, but the design cannot support claims of ...
-
Towards Scalable SOAP Note Generation: A Weakly Supervised Multimodal Framework
A weakly supervised, retrieval-augmented vision-language framework generates structured SOAP notes from lesion images and sparse clinical text, with evaluation against GPT-4o, Claude, and Janus Pro on a small set of cases.
-
Process-Supervised Reward Models for Verifying Clinical Note Generation: A Scalable Approach Guided by Domain Expertise
A step-level reward model trained on expert-designed synthetic clinical errors detects injected note errors with 98.8% accuracy and selects physician-preferred notes with 56.2% accuracy.
-
Embeddings to Diagnosis: Latent Fragility under Agentic Perturbations in Clinical LLMs
LDFR is a proposed diagnostic signal that flags LLM diagnostic instability when embeddings cross a PCA-based boundary under small clinical edits, tested on synthetic and 90 real notes.
-
Assessment of AI-Generated Pediatric Rehabilitation SOAP-Note Quality
AI-drafted and human-written pediatric SOAP notes received statistically indistinguishable clinician quality scores, but the study's ANOVA used only four observations per group and cannot support the claim of comparab...
-
Skin-SOAP: A Weakly Supervised Framework for Generating Structured SOAP Notes
Skin-SOAP is a weakly supervised multimodal system that turns a skin lesion image and sparse clinical text into structured SOAP notes, evaluated with two new metrics.
Discussion (0). Continue with ORCID to comment.