Pith. sign in

REVIEW 1 cited by

Weighted Residuals for Very Deep Networks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1605.08831 v1 pith:EWBWOAL5 submitted 2016-05-28 cs.CV

classification cs.CV
keywords networksresidualdeepweightedlayersnetworkoriginalaccuracy
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Deep residual networks have recently shown appealing performance on many challenging computer vision tasks. However, the original residual structure still has some defects making it difficult to converge on very deep networks. In this paper, we introduce a weighted residual network to address the incompatibility between \texttt{ReLU} and element-wise addition and the deep network initialization problem. The weighted residual network is able to learn to combine residuals from different layers effectively and efficiently. The proposed models enjoy a consistent improvement over accuracy and convergence with increasing depths from 100+ layers to 1000+ layers. Besides, the weighted residual networks have little more computation and GPU memory burden than the original residual networks. The networks are optimized by projected stochastic gradient descent. Experiments on CIFAR-10 have shown that our algorithm has a \emph{faster convergence speed} than the original residual networks and reaches a \emph{high accuracy} at 95.3\% with a 1192-layer model.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. CNNtention: Can CNNs do better with Attention?

    cs.CV 2024-12 conditional novelty 3.0 of 10

    Adding attention blocks between ResNet-20 feature extractors gives small accuracy improvements on CIFAR-10 and MNIST, but the gains are within a range that needs error bars to trust.

Pith tools