Pith. sign in

REVIEW

VT-ADL: A Vision Transformer Network for Image Anomaly Detection and Localization

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2104.10036 v1 pith:ZBSGA2U5 submitted 2021-04-20 cs.CV cs.AIcs.LG

classification cs.CVcs.AIcs.LG
keywords anomalynetworkdetectionimagelocalizationtransformeradditionalgorithms
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We present a transformer-based image anomaly detection and localization network. Our proposed model is a combination of a reconstruction-based approach and patch embedding. The use of transformer networks helps to preserve the spatial information of the embedded patches, which are later processed by a Gaussian mixture density network to localize the anomalous areas. In addition, we also publish BTAD, a real-world industrial anomaly dataset. Our results are compared with other state-of-the-art algorithms using publicly available datasets like MNIST and MVTec.

Discussion (0). Continue with ORCID to comment.

Pith tools