pith. sign in

arxiv: 1411.5908 · v2 · pith:INZLDBQPnew · submitted 2014-11-21 · 💻 cs.CV · cs.LG· cs.NE

Understanding image representations by measuring their equivariance and equivalence

classification 💻 cs.CV cs.LGcs.NE
keywords representationsequivalenceequivarianceimageincludinginvariancelayersmethods
0
0 comments X
read the original abstract

Despite the importance of image representations such as histograms of oriented gradients and deep Convolutional Neural Networks (CNN), our theoretical understanding of them remains limited. Aiming at filling this gap, we investigate three key mathematical properties of representations: equivariance, invariance, and equivalence. Equivariance studies how transformations of the input image are encoded by the representation, invariance being a special case where a transformation has no effect. Equivalence studies whether two representations, for example two different parametrisations of a CNN, capture the same visual information or not. A number of methods to establish these properties empirically are proposed, including introducing transformation and stitching layers in CNNs. These methods are then applied to popular representations to reveal insightful aspects of their structure, including clarifying at which layers in a CNN certain geometric invariances are achieved. While the focus of the paper is theoretical, direct applications to structured-output regression are demonstrated too.

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Transformer Field Theory: A Response-Theoretic Approach to Mechanistic Interpretability

    cs.LG 2026-05 unverdicted novelty 7.0

    Transformer Field Theory frames the residual stream as a field, models patching as source insertion, and uses first-order sensitivities plus Green functions to predict and describe responses, with empirical tests on G...