Pith. sign in

REVIEW 2 cited by

Weakly Supervised Medical Diagnosis and Localization from Multiple Resolutions

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1803.07703 v1 pith:EICZYG6Z submitted 2018-03-21 cs.CV

classification cs.CV
keywords abnormalitiesapplyingapproachfindingsgeneratingglobalimage-levellabels
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Diagnostic imaging often requires the simultaneous identification of a multitude of findings of varied size and appearance. Beyond global indication of said findings, the prediction and display of localization information improves trust in and understanding of results when augmenting clinical workflow. Medical training data rarely includes more than global image-level labels as segmentations are time-consuming and expensive to collect. We introduce an approach to managing these practical constraints by applying a novel architecture which learns at multiple resolutions while generating saliency maps with weak supervision. Further, we parameterize the Log-Sum-Exp pooling function with a learnable lower-bounded adaptation (LSE-LBA) to build in a sharpness prior and better handle localizing abnormalities of different sizes using only image-level labels. Applying this approach to interpreting chest x-rays, we set the state of the art on 9 abnormalities in the NIH's CXR14 dataset while generating saliency maps with the highest resolution to date.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Interpreting Radiologist's Intention from Eye Movements in Chest X-ray Diagnosis

    cs.CV 2025-07 reject novelty 6.0 of 10

    RadGazeIntent, a transformer model, predicts per-fixation diagnostic intention from radiologist gaze on chest X-rays, evaluated on three newly constructed intention-labeled datasets.

  2. DeepChest: Dynamic Gradient-Free Task Weighting for Effective Multi-Task Learning in Chest X-ray Classification

    cs.CV 2025-05 reject novelty 5.0 of 10

    DeepChest weights each chest X-ray pathology task by comparing its current training accuracy to the average, boosting weak tasks and shrinking strong ones.

Pith tools