REVIEW 3 cited by
Prompt engineering paradigms for medical applications: scoping review and recommendations for better practices
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Prompt engineering is crucial for harnessing the potential of large language models (LLMs), especially in the medical domain where specialized terminology and phrasing is used. However, the efficacy of prompt engineering in the medical domain remains to be explored. In this work, 114 recent studies (2022-2024) applying prompt engineering in medicine, covering prompt learning (PL), prompt tuning (PT), and prompt design (PD) are reviewed. PD is the most prevalent (78 articles). In 12 papers, PD, PL, and PT terms were used interchangeably. ChatGPT is the most commonly used LLM, with seven papers using it for processing sensitive clinical data. Chain-of-Thought emerges as the most common prompt engineering technique. While PL and PT articles typically provide a baseline for evaluating prompt-based approaches, 64% of PD studies lack non-prompt-related baselines. We provide tables and figures summarizing existing work, and reporting recommendations to guide future research contributions.
Forward citations
Cited by 3 Pith papers
-
BiomedCoOp: Learning to Prompt for Biomedical Vision-Language Models
BiomedCoOp improves few-shot biomedical image classification by aligning learnable prompts with selectively pruned LLM-generated prompt ensembles and distilling their knowledge into BiomedCLIP.
-
A Hybrid Artificial Intelligence System for Automated EEG Background Analysis and Report Generation
A hybrid AI system combining deep learning, artifact removal, and expert heuristics interprets EEG background activity and uses Gemini to write reports, but key validation claims are weakened by circular LLM verificat...
-
Clinical trial cohort selection using Large Language Models on n2c2 Challenges
Open-source LLMs achieve moderate F1 on straightforward clinical trial criteria but underperform challenge-winning systems on criteria requiring fine-grained reasoning across three n2c2 datasets.
Discussion (0). Continue with ORCID to comment.