arXiv:1512.03385 · 2015

Deep Residual Learning for Image Recognition

Kaiming He, Xiangyu Zhang, Shaoqing Ren, Jian Sun

Deeper neural networks are more difficult to train. We present a residual learning framework to ease the training of networks that are substantially deeper than those used previously. We explicitly reformulate the layers as learning residual functions with reference to the layer inputs, instead of learning unreferenced functions. We provide comprehensive empirical evidence showing that these residual networks are easier to optimize, and can gain accuracy from considerably increased depth. On the ImageNet dataset we evaluate residual nets with a depth of up to 152 layers—8×\times deeper than VGG nets [41] but still having lower complexity. An ensemble of these residual nets achieves 3.57% error on the ImageNet test set. This result won the 1st place on the ILSVRC 2015 classification task. We also present analysis on CIFAR-10 with 100 and 1000 layers.

Residual LearningDeep LearningComputer VisionResNetOptimization

Current guide

Version 5

Difficulty
Intermediate
Study time
120 minutes
Coverage
7 learning objectives
Quality
Equation fidelity: high

Immutable history

Guide versions

  1. Version 5Current
    Read guide version 5
Show 4 earlier guide versions
  1. Version 4Archived
    Read guide version 4
  2. Version 3Archived
    Read guide version 3
  3. Version 2Archived
    Read guide version 2
  4. Version 1Archived
    Read guide version 1