Pith. sign in

REVIEW 1 cited by

Defects of Convolutional Decoder Networks in Frequency Representation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2210.09020 v2 pith:N65LHX4S submitted 2022-10-17 cs.LG cs.AIcs.CV

classification cs.LGcs.AIcs.CV
keywords decoderfrequencyprovecomponentsconvolutionaldefectsfeaturenetwork
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

In this paper, we prove the representation defects of a cascaded convolutional decoder network, considering the capacity of representing different frequency components of an input sample. We conduct the discrete Fourier transform on each channel of the feature map in an intermediate layer of the decoder network. Then, we extend the 2D circular convolution theorem to represent the forward and backward propagations through convolutional layers in the frequency domain. Based on this, we prove three defects in representing feature spectrums. First, we prove that the convolution operation, the zero-padding operation, and a set of other settings all make a convolutional decoder network more likely to weaken high-frequency components. Second, we prove that the upsampling operation generates a feature spectrum, in which strong signals repetitively appear at certain frequencies. Third, we prove that if the frequency components in the input sample and frequency components in the target output for regression have a small shift, then the decoder usually cannot be effectively learned.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Adaptive High-Pass Kernel Prediction for Efficient Video Deblurring

    cs.CV 2024-12 conditional novelty 5.0 of 10

    AHFNet dynamically weights four handpicked high-pass kernels (Sobel and temporal gradients) to extract sharpening features, reaching 33.25 dB PSNR on GOPRO with roughly one-sixth the training memory of heavier models.

Pith tools