REVIEW 3 cited by
Zero-Shot Clinical Trial Patient Matching with LLMs
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Matching patients to clinical trials is a key unsolved challenge in bringing new drugs to market. Today, identifying patients who meet a trial's eligibility criteria is highly manual, taking up to 1 hour per patient. Automated screening is challenging, however, as it requires understanding unstructured clinical text. Large language models (LLMs) offer a promising solution. In this work, we explore their application to trial matching. First, we design an LLM-based system which, given a patient's medical history as unstructured clinical text, evaluates whether that patient meets a set of inclusion criteria (also specified as free text). Our zero-shot system achieves state-of-the-art scores on the n2c2 2018 cohort selection benchmark. Second, we improve the data and cost efficiency of our method by identifying a prompting strategy which matches patients an order of magnitude faster and more cheaply than the status quo, and develop a two-stage retrieval pipeline that reduces the number of tokens processed by up to a third while retaining high performance. Third, we evaluate the interpretability of our system by having clinicians evaluate the natural language justifications generated by the LLM for each eligibility decision, and show that it can output coherent explanations for 97% of its correct decisions and 75% of its incorrect ones. Our results establish the feasibility of using LLMs to accelerate clinical trial operations.
Forward citations
Cited by 3 Pith papers
-
Embedding-Driven Diversity Sampling to Improve Few-Shot Synthetic Data Generation
Embedding-driven diversity sampling of few-shot examples improves downstream classification with synthetic clinical text over random and zero-shot baselines on CheXpert radiology reports.
-
A Contrastive Pretrain Model with Prompt Tuning for Multi-center Medication Recommendation
TEMPT, a contrastive pretraining model with per-hospital prompt tuning, outperforms existing baselines on multi-center medication recommendation in the eICU dataset.
-
Distilling Large Language Models for Efficient Clinical Information Extraction
Small BERT models distilled from LLM and ontology labels match the teachers on medication and disease extraction at a fraction of cost, but trail on symptoms.
Discussion (0). Continue with ORCID to comment.