Deep Residual Networks with 1K Layers

By Kaiming He, Xiangyu Zhang, Shaoqing Ren, Jian Sun.

Microsoft Research Asia (MSRA).

Introduction
Notes
Usage

Introduction

This repository contains re-implemented code for the paper "Identity Mappings in Deep Residual Networks" (http://arxiv.org/abs/1603.05027). This work enables training quality 1k-layer neural networks in a super simple way.

Acknowledgement: This code is re-implemented by Xiang Ming from Xi'an Jiaotong Univeristy for the ease of release.

Seel Also: Re-implementations of ResNet-200 [a] on ImageNet from Facebook AI Research (FAIR): https://github.com/facebook/fb.resnet.torch/tree/master/pretrained

Notes

This code is based on the implementation of Torch ResNets (https://github.com/facebook/fb.resnet.torch).
The experiments in the paper were conducted in Caffe, whereas this code is re-implemented in Torch. We observed similar results within reasonable statistical variations.
To fit the 1k-layer models into memory without modifying much code, we simply reduced the mini-batch size to 64, noting that results in the paper were obtained with a mini-batch size of 128. Less expectedly, the results with the mini-batch size of 64 are slightly better:

mini-batch CIFAR-10 test error (%): (median (mean+/-std))

128 (as in [a]) 4.92 (4.89+/-0.14)

64 (as in this code) 4.62 (4.69+/-0.20)
Curves obtained by running this code with a mini-batch size of 64 (training loss: y-axis on the left; test error: y-axis on the right):

mini-batch	CIFAR-10 test error (%): (median (mean+/-std))
128 (as in [a])	4.92 (4.89+/-0.14)
64 (as in this code)	4.62 (4.69+/-0.20)

Usage

Install Torch ResNets (https://github.com/facebook/fb.resnet.torch) following instructions therein.
Add the file resnet-pre-act.lua from this repository to ./models.
To train ResNet-1001 as of the form in [a]:

th main.lua -netType resnet-pre-act -depth 1001 -batchSize 64 -nGPU 2 -nThreads 4 -dataset cifar10 -nEpochs 200 -shareGradInput false

Note: ``shareGradInput=true'' is not valid for this model yet.

Deep Residual Networks with 1K Layers

Related tags

Overview

Deep Residual Networks with 1K Layers

Table of Contents

Introduction

Notes

Usage

Owner

Kaiming He

Copy Paste positive polyp using poisson image blending for medical image segmentation

ADSPM: Attribute-Driven Spontaneous Motion in Unpaired Image Translation

Python scripts for performing road segemtnation and car detection using the HybridNets multitask model in ONNX.

Self-Adaptable Point Processes with Nonparametric Time Decays

Moon-patrol - A faithful recreation of the 1983 hit classic Moon Patrol for the Atari 2600 created using the Pygame library for Python

Torchyolo - Yolov3 ve Yolov4 modellerin Pytorch uygulamasıdır

Time Dependent DFT in Tamm-Dancoff Approximation

Algorithmic Trading using RNN

A simple baseline for the 2022 IEEE GRSS Data Fusion Contest (DFC2022)

State-Relabeling Adversarial Active Learning

GraphGT: Machine Learning Datasets for Graph Generation and Transformation

Leveraging Two Types of Global Graph for Sequential Fashion Recommendation, ICMR 2021

Object Depth via Motion and Detection Dataset

Rethinking Transformer-based Set Prediction for Object Detection

A PaddlePaddle implementation of Time Interval Aware Self-Attentive Sequential Recommendation.

A collection of implementations of deep domain adaptation algorithms

[Open Source]. The improved version of AnimeGAN. Landscape photos/videos to anime

HarDNeXt: Official HarDNeXt repository

Code for our paper "Graph Pre-training for AMR Parsing and Generation" in ACL2022

MPI-IS Mesh Processing Library