REVIEW 4 cited by
Publicly Shareable Clinical Large Language Model Built on Synthetic Clinical Notes
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
The development of large language models tailored for handling patients' clinical notes is often hindered by the limited accessibility and usability of these notes due to strict privacy regulations. To address these challenges, we first create synthetic large-scale clinical notes using publicly available case reports extracted from biomedical literature. We then use these synthetic notes to train our specialized clinical large language model, Asclepius. While Asclepius is trained on synthetic data, we assess its potential performance in real-world applications by evaluating it using real clinical notes. We benchmark Asclepius against several other large language models, including GPT-3.5-turbo and other open-source alternatives. To further validate our approach using synthetic notes, we also compare Asclepius with its variants trained on real clinical notes. Our findings convincingly demonstrate that synthetic clinical notes can serve as viable substitutes for real ones when constructing high-performing clinical language models. This conclusion is supported by detailed evaluations conducted by both GPT-4 and medical professionals. All resources including weights, codes, and data used in the development of Asclepius are made publicly accessible for future research. (https://github.com/starmpcc/Asclepius)
Forward citations
Cited by 4 Pith papers
-
DENSE: Longitudinal Progress Note Generation with Temporal Modeling of Heterogeneous Clinical Notes Across Hospital Visits
DENSE synthesizes progress notes across hospital visits using retrieval over heterogeneous clinical notes, claiming temporal continuity that even exceeds gold-standard notes.
-
Lossless Compression of Large Language Model-Generated Text via Next-Token Prediction
Using an LLM as a next-token predictor with arithmetic coding compresses LLM-generated text about 20x, roughly 4 to 7 times better than Gzip, LZMA, or neural compressors in the paper's benchmarks.
-
Mitigating hallucinations in healthcare LLMs with granular fact-checking and domain-specific adaptation
A deterministic, proposition-level fact-checker that compares clinical summaries against electronic health records via (entity, attribute, value, time) claims and hard-coded logical checks reports 0.8904 precision and...
-
Synthetic Data Generation with LLM for Improved Depression Prediction
LLM-generated synthetic synopses conditioned on target PHQ-8 scores, added to real DAIC-WOZ synopses, reduce PHQ-8 regression error (RMSE 4.64, MAE 3.66) in a single-run BERT evaluation.
Discussion (0). Continue with ORCID to comment.