Pith. sign in

REVIEW 1 cited by

StudyFormer : Attention-Based and Dynamic Multi View Classifier for X-ray images

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2302.11840 v1 pith:SFJPMFHE submitted 2023-02-23 cs.CV cs.AIcs.LG

classification cs.CVcs.AIcs.LG
keywords x-rayimagesapproachclassificationmodelsviewinformationmultiple
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Chest X-ray images are commonly used in medical diagnosis, and AI models have been developed to assist with the interpretation of these images. However, many of these models rely on information from a single view of the X-ray, while multiple views may be available. In this work, we propose a novel approach for combining information from multiple views to improve the performance of X-ray image classification. Our approach is based on the use of a convolutional neural network to extract feature maps from each view, followed by an attention mechanism implemented using a Vision Transformer. The resulting model is able to perform multi-label classification on 41 labels and outperforms both single-view models and traditional multi-view classification architectures. We demonstrate the effectiveness of our approach through experiments on a dataset of 363,000 X-ray images.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. VET-DINO: Learning Anatomical Understanding Through Multi-View Distillation in Veterinary Imaging

    cs.CV 2025-05 conditional novelty 6.0 of 10

    Using real multi-view radiographs from the same study as self-supervised training pairs yields better anatomical representations and downstream veterinary task performance than synthetic single-image augmentations.

Pith tools