Pith. sign in

REVIEW 1 cited by

Understanding Cross-Model Perceptual Invariances Through Ensemble Metamers

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2504.01739 v2 pith:YCI5IDK5 submitted 2025-04-02 cs.CV

Understanding Cross-Model Perceptual Invariances Through Ensemble Metamers

classification cs.CV
keywords metamersneuralinvariancesnetworksvisionartificialconvolutionalperceptual
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
Share X Bluesky LinkedIn Reddit HN
read the original abstract

Understanding the perceptual invariances of artificial neural networks is essential for improving explainability and aligning models with human vision. Metamers - stimuli that are physically distinct yet produce identical neural activations - serve as a valuable tool for investigating these invariances. We introduce a novel approach to metamer generation by leveraging ensembles of artificial neural networks, capturing shared representational subspaces across diverse architectures, including convolutional neural networks and vision transformers. To characterize the properties of the generated metamers, we employ a suite of image-based metrics that assess factors such as semantic fidelity and naturalness. Our findings show that convolutional neural networks generate more recognizable and human-like metamers, while vision transformers produce realistic but less transferable metamers, highlighting the impact of architectural biases on representational invariances.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. MRD: Using Physically Based Differentiable Rendering to Probe Vision Models for 3D Scene Understanding

    cs.CV 2025-12 conditional novelty 6.0

    MRD finds physically different 3D scenes that reproduce a target model activation, revealing which shape and material properties vision models are sensitive to.