Pith. sign in

REVIEW 2 cited by

COVID-VIT: Classification of COVID-19 from CT chest images based on vision transformer models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2107.01682 v1 pith:ZDMT3G34 submitted 2021-07-04 eess.IV cs.CV

classification eess.IVcs.CV
keywords covid-19dataimagestransformervisionaxialchestclassification
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

This paper is responding to the MIA-COV19 challenge to classify COVID from non-COVID based on CT lung images. The COVID-19 virus has devastated the world in the last eighteen months by infecting more than 182 million people and causing over 3.9 million deaths. The overarching aim is to predict the diagnosis of the COVID-19 virus from chest radiographs, through the development of explainable vision transformer deep learning techniques, leading to population screening in a more rapid, accurate and transparent way. In this competition, there are 5381 three-dimensional (3D) datasets in total, including 1552 for training, 374 for evaluation and 3455 for testing. While most of the data volumes are in axial view, there are a number of subjects' data are in coronal or sagittal views with 1 or 2 slices are in axial view. Hence, while 3D data based classification is investigated, in this competition, 2D images remains the main focus. Two deep learning methods are studied, which are vision transformer (ViT) based on attention models and DenseNet that is built upon conventional convolutional neural network (CNN). Initial evaluation results based on validation datasets whereby the ground truth is known indicate that ViT performs better than DenseNet with F1 scores being 0.76 and 0.72 respectively. Codes are available at GitHub at <https://github/xiaohong1/COVID-ViT>.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Capturing star formation activity from compressed photometric images of galaxies

    astro-ph.GA 2025-07 conditional novelty 5.0 of 10

    A Vision Transformer trained on SDSS JPEG composites distinguishes star-forming galaxies from passive ones with about 86% accuracy and predicts BPT line ratios.

  2. Unmasking Interstitial Lung Diseases: Leveraging Masked Autoencoders for Diagnosis

    eess.IV 2025-08 conditional novelty 4.0 of 10

    A masked autoencoder pretrained on mixed COVID-19 and ILD CT scans improves multiclass interstitial lung disease classification over supervised baselines on a small cohort.

Pith tools