Pith. sign in

REVIEW 1 cited by

CAMixerSR: Only Details Need More "Attention"

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.19289 v2 pith:KCN2PDZI submitted 2024-02-29 eess.IV cs.CV

classification eess.IVcs.CV
keywords convolutionattentioncamixercamixersrcontent-awarefurthermixernetworks
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

To satisfy the rapidly increasing demands on the large image (2K-8K) super-resolution (SR), prevailing methods follow two independent tracks: 1) accelerate existing networks by content-aware routing, and 2) design better super-resolution networks via token mixer refining. Despite directness, they encounter unavoidable defects (e.g., inflexible route or non-discriminative processing) limiting further improvements of quality-complexity trade-off. To erase the drawbacks, we integrate these schemes by proposing a content-aware mixer (CAMixer), which assigns convolution for simple contexts and additional deformable window-attention for sparse textures. Specifically, the CAMixer uses a learnable predictor to generate multiple bootstraps, including offsets for windows warping, a mask for classifying windows, and convolutional attentions for endowing convolution with the dynamic property, which modulates attention to include more useful textures self-adaptively and improves the representation capability of convolution. We further introduce a global classification loss to improve the accuracy of predictors. By simply stacking CAMixers, we obtain CAMixerSR which achieves superior performance on large-image SR, lightweight SR, and omnidirectional-image SR.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. PixelSR: Efficient Screen Content Super-Resolution via Pixel Classification

    cs.CV 2026-08 conditional novelty 5.5 of 10

    PixelSR achieves faster screen-content super-resolution by classifying pixels into unique/repeated/background and using an on-the-fly lookup table plus nearest-neighbor prediction.

Pith tools