REVIEW 6 cited by
Agentic LLM Workflows for Generating Patient-Friendly Medical Reports
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
The application of Large Language Models (LLMs) in healthcare is expanding rapidly, with one potential use case being the translation of formal medical reports into patient-legible equivalents. Currently, LLM outputs often need to be edited and evaluated by a human to ensure both factual accuracy and comprehensibility, and this is true for the above use case. We aim to minimize this step by proposing an agentic workflow with the Reflexion framework, which uses iterative self-reflection to correct outputs from an LLM. This pipeline was tested and compared to zero-shot prompting on 16 randomized radiology reports. In our multi-agent approach, reports had an accuracy rate of 94.94% when looking at verification of ICD-10 codes, compared to zero-shot prompted reports, which had an accuracy rate of 68.23%. Additionally, 81.25% of the final reflected reports required no corrections for accuracy or readability, while only 25% of zero-shot prompted reports met these criteria without needing modifications. These results indicate that our approach presents a feasible method for communicating clinical findings to patients in a quick, efficient and coherent manner whilst also retaining medical accuracy. The codebase is available for viewing at http://github.com/malavikhasudarshan/Multi-Agent-Patient-Letter-Generation.
Forward citations
Cited by 6 Pith papers
-
MedSentry: Understanding and Mitigating Safety Risks in Medical LLM Multi-Agent Systems
A 5,000-prompt medical safety benchmark reveals that decentralized LLM multi-agent teams resist a malicious insider agent better than shared-pool teams, and a personality-screening defense partially restores safety.
-
FALCON: Transforming Cyber Threat Intelligence into Deployable IDS Rules with Self-Reflection
FALCON automates the generation of Snort and YARA intrusion detection rules from cyber threat intelligence using an LLM agent pipeline with a contrastively trained CTI-rule semantic scorer as a ground-truth-free validator.
-
MAARTA:Multi-Agentic Adaptive Radiology Teaching Assistant
MAARTA, a multi-agent LLM framework comparing expert and student gaze graphs, reports higher accuracy than single-agent baselines on simulated perceptual errors in chest X-ray interpretation.
-
Generative to Agentic AI: Survey, Conceptualization, and Challenges
Agentic AI is characterized over Generative AI by iterative reasoning, environment interaction, memory, and tool use, with autonomy as the defining difference.
-
Towards a HIPAA Compliant Agentic AI System in Healthcare
The paper claims a HIPAA-compliant agentic AI framework using ABAC, hybrid regex+BERT PHI redaction, and immutable audit trails, supported only by synthetic-data experiments.
-
Agentic AI Systems Applied to tasks in Financial Services: Modeling and model risk management crews
A CrewAI-based multi-agent system with human oversight built financial models and carried out model risk management checks on three public credit datasets, with results comparable to AutoML and Kaggle baselines.
Discussion (0). Continue with ORCID to comment.