REVIEW 3 cited by
RadFusion: Benchmarking Performance and Fairness for Multimodal Pulmonary Embolism Detection from CT and EHR
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Despite the routine use of electronic health record (EHR) data by radiologists to contextualize clinical history and inform image interpretation, the majority of deep learning architectures for medical imaging are unimodal, i.e., they only learn features from pixel-level information. Recent research revealing how race can be recovered from pixel data alone highlights the potential for serious biases in models which fail to account for demographics and other key patient attributes. Yet the lack of imaging datasets which capture clinical context, inclusive of demographics and longitudinal medical history, has left multimodal medical imaging underexplored. To better assess these challenges, we present RadFusion, a multimodal, benchmark dataset of 1794 patients with corresponding EHR data and high-resolution computed tomography (CT) scans labeled for pulmonary embolism. We evaluate several representative multimodal fusion models and benchmark their fairness properties across protected subgroups, e.g., gender, race/ethnicity, age. Our results suggest that integrating imaging and EHR data can improve classification performance and robustness without introducing large disparities in the true positive rate between population groups.
Forward citations
Cited by 3 Pith papers
-
Multimodal Routing for Interpretable, Robust, and Auditable Clinical Prediction
Explicit unimodal, directional-bimodal and trimodal routes plus inference-time masking yield higher AUROC/F1 than fusion baselines on MIMIC-IV mortality and 25-phenotype tasks while exposing modality reliance.
-
Benchmarking Foundation Models with Multimodal Public Electronic Health Records
A standardized MIMIC-IV benchmark comparing eight unimodal and multimodal foundation models shows multimodal inputs improve predictive performance without adding bias, while medical LVLMs underperform on length-of-sta...
-
Harnessing EHRs for Diffusion-based Anomaly Detection on Chest X-rays
Diff3M conditions a diffusion-based chest X-ray anomaly detector on EHR tokens and a checkerboard mask, yielding small AUROC gains over prior medical UAD methods.
Discussion (0). Continue with ORCID to comment.