Pith. sign in

REVIEW 2 cited by

Exploring Robustness in Doctor-Patient Conversation Summarization: An Analysis of Out-of-Domain SOAP Notes

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2406.02826 v1 pith:RU27YCTB submitted 2024-06-05 cs.CL cs.LG

classification cs.CLcs.LG
keywords notessoapanalysisconversationdatadoctor-patientmodelout-of-domain
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Summarizing medical conversations poses unique challenges due to the specialized domain and the difficulty of collecting in-domain training data. In this study, we investigate the performance of state-of-the-art doctor-patient conversation generative summarization models on the out-of-domain data. We divide the summarization model of doctor-patient conversation into two configurations: (1) a general model, without specifying subjective (S), objective (O), and assessment (A) and plan (P) notes; (2) a SOAP-oriented model that generates a summary with SOAP sections. We analyzed the limitations and strengths of the fine-tuning language model-based methods and GPTs on both configurations. We also conducted a Linguistic Inquiry and Word Count analysis to compare the SOAP notes from different datasets. The results exhibit a strong correlation for reference notes across different datasets, indicating that format mismatch (i.e., discrepancies in word distribution) is not the main cause of performance decline on out-of-domain data. Lastly, a detailed analysis of SOAP notes is included to provide insights into missing information and hallucinations introduced by the models.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Towards Scalable SOAP Note Generation: A Weakly Supervised Multimodal Framework

    cs.CV 2025-06 conditional novelty 6.0 of 10

    A weakly supervised, retrieval-augmented vision-language framework generates structured SOAP notes from lesion images and sparse clinical text, with evaluation against GPT-4o, Claude, and Janus Pro on a small set of cases.

  2. Skin-SOAP: A Weakly Supervised Framework for Generating Structured SOAP Notes

    cs.CV 2025-08 unverdicted novelty 4.0 of 10

    Skin-SOAP is a weakly supervised multimodal system that turns a skin lesion image and sparse clinical text into structured SOAP notes, evaluated with two new metrics.

Pith tools