Pith. sign in

REVIEW 1 cited by

A study on the adequacy of common IQA measures for medical images

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2405.19224 v4 pith:NKSHY5UC submitted 2024-05-29 eess.IV cs.CV

classification eess.IVcs.CV
keywords imagesmeasuresmedicalnaturaldataimagealgorithmsassessment
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Image quality assessment (IQA) is standard practice in the development stage of novel machine learning algorithms that operate on images. The most commonly used IQA measures have been developed and tested for natural images, but not in the medical setting. Reported inconsistencies arising in medical images are not surprising, as they have different properties than natural images. In this study, we test the applicability of common IQA measures for medical image data by comparing their assessment to manually rated chest X-ray (5 experts) and photoacoustic image data (2 experts). Moreover, we include supplementary studies on grayscale natural images and accelerated brain MRI data. The results of all experiments show a similar outcome in line with previous findings for medical images: PSNR and SSIM in the default setting are in the lower range of the result list and HaarPSI outperforms the other tested measures in the overall performance. Also among the top performers in our experiments are the full reference measures FSIM, LPIPS and MS-SSIM. Generally, the results on natural images yield considerably higher correlations, suggesting that additional employment of tailored IQA measures for medical imaging algorithms is needed.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Pathology-Guided Virtual Staining Metric for Evaluation and Training

    eess.IV 2025-07 reject novelty 6.0 of 10

    PaPIS is a pathology-aware full-reference similarity metric for virtual staining, built from cell-morphology segmentation features and Retinex decomposition, demonstrated as both an evaluation score and a training loss.

Pith tools