REVIEW 2 cited by
LipSim: A Provably Robust Perceptual Similarity Metric
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
abstract
Recent years have seen growing interest in developing and applying perceptual similarity metrics. Research has shown the superiority of perceptual metrics over pixel-wise metrics in aligning with human perception and serving as a proxy for the human visual system. On the other hand, as perceptual metrics rely on neural networks, there is a growing concern regarding their resilience, given the established vulnerability of neural networks to adversarial attacks. It is indeed logical to infer that perceptual metrics may inherit both the strengths and shortcomings of neural networks. In this work, we demonstrate the vulnerability of state-of-the-art perceptual similarity metrics based on an ensemble of ViT-based feature extractors to adversarial attacks. We then propose a framework to train a robust perceptual similarity metric called LipSim (Lipschitz Similarity Metric) with provable guarantees. By leveraging 1-Lipschitz neural networks as the backbone, LipSim provides guarded areas around each data point and certificates for all perturbations within an $\ell_2$ ball. Finally, a comprehensive set of experiments shows the performance of LipSim in terms of natural and certified scores and on the image retrieval application. The code is available at https://github.com/SaraGhazanfari/LipSim.
Forward citations
Cited by 2 Pith papers
-
Disentangling Safe and Unsafe Corruptions via Anisotropy and Locality
Projected Displacement assigns low threat to safe corruptions like blur and noise even at large norms, and high threat to perturbations aligned with label-changing directions.
-
Stochastic BIQA: Median Randomized Smoothing for Certified Blind Image Quality Assessment
Median smoothing plus a trained denoiser with ranking loss yields certified l2 robustness for no-reference image quality metrics while preserving correlation with subjective scores better than prior smoothing baselines.
Discussion (0). Continue with ORCID to comment.