REVIEW 3 cited by
RadGazeGen: Radiomics and Gaze-guided Medical Image Generation using Diffusion Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
In this work, we present RadGazeGen, a novel framework for integrating experts' eye gaze patterns and radiomic feature maps as controls to text-to-image diffusion models for high fidelity medical image generation. Despite the recent success of text-to-image diffusion models, text descriptions are often found to be inadequate and fail to convey detailed disease-specific information to these models to generate clinically accurate images. The anatomy, disease texture patterns, and location of the disease are extremely important to generate realistic images; moreover the fidelity of image generation can have significant implications in downstream tasks involving disease diagnosis or treatment repose assessment. Hence, there is a growing need to carefully define the controls used in diffusion models for medical image generation. Eye gaze patterns of radiologists are important visuo-cognitive information, indicative of subtle disease patterns and spatial location. Radiomic features further provide important subvisual cues regarding disease phenotype. In this work, we propose to use these gaze patterns in combination with standard radiomics descriptors, as controls, to generate anatomically correct and disease-aware medical images. RadGazeGen is evaluated for image generation quality and diversity on the REFLACX dataset. To demonstrate clinical applicability, we also show classification performance on the generated images from the CheXpert test set (n=500) and long-tailed learning performance on the MIMIC-CXR-LT test set (n=23550).
Forward citations
Cited by 3 Pith papers
-
Pathologist Attention-Aligned Report Generation for Prostate Histopathology
Using pathologists' eye movements as a training signal improves prostate pathology report generation and makes model attention more human-like.
-
ImmunoDiff: A Diffusion Model for Immunotherapy Response Prediction in Lung Cancer
An anatomy- and clinical-conditioned diffusion model that synthesizes post-treatment CT and uses its features to improve immunotherapy response prediction in NSCLC.
-
GazeLT: Visual attention-guided long-tailed disease classification in chest radiographs
GazeLT classifies long-tailed chest X-ray diseases by training a student model with a teacher that learns time-windowed radiologist gaze attention, improving tail-class accuracy on the NIH-CXR-LT and MIMIC-CXR-LT benchmarks.
Discussion (0). Continue with ORCID to comment.