Pith. sign in

REVIEW 1 cited by

Understanding Transformer-based Vision Models through Inversion

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2412.06534 v4 pith:4R5P4PZ2 submitted 2024-12-09 cs.CV cs.AIcs.LGcs.NE

classification cs.CVcs.AIcs.LGcs.NE
keywords visionmodelsinversiontransformertransformer-basedunderstandingfeatureimage
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Understanding the mechanisms underlying deep neural networks remains a fundamental challenge in machine learning and computer vision. One promising, yet only preliminarily explored approach, is feature inversion, which attempts to reconstruct images from intermediate representations using trained inverse neural networks. In this study, we revisit feature inversion, introducing a novel, modular variation that enables significantly more efficient application of the technique. We demonstrate how our method can be systematically applied to the large-scale transformer-based vision models, Detection Transformer and Vision Transformer, and how reconstructed images can be qualitatively interpreted in a meaningful way. We further quantitatively evaluate our method, thereby uncovering underlying mechanisms of representing image features that emerge in the two transformer architectures. Our analysis reveals key insights into how these models encode contextual shape and image details, how their layers correlate, and their robustness against color perturbations. These findings contribute to a deeper understanding of transformer-based vision models and their internal representations. The code for reproducing our experiments is available at github.com/wiskott-lab/inverse-tvm.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Inverting the Hidden: Unveiling Multimodal Privacy Leakage in Collaborative LVLM Inference

    cs.CR 2026-08 conditional novelty 6.0 of 10

    Intermediate hidden states transmitted during collaborative LVLM inference leak enough information to reconstruct input images and text with near-perfect token accuracy.

Pith tools