ResNeSt: Split-Attention Networks

2020-04-19Code Available3· sign in to hype

Hang Zhang, Chongruo wu, Zhongyue Zhang, Yi Zhu, Haibin Lin, Zhi Zhang, Yue Sun, Tong He, Jonas Mueller, R. Manmatha, Mu Li, Alexander Smola

arXiv PDF

Code Available — Be the first to reproduce this paper.

Reproduce

Code

github.com/zhanghang1989/ResNeSt
OfficialIn paperpytorch★ 3,264
github.com/dmlc/gluon-cv
tf★ 5,919
github.com/zhanghang1989/PyTorch-Encoding
pytorch★ 2,048
github.com/chongruo/detectron2-resnest
pytorch★ 387
github.com/zhanghang1989/detectron2-ResNeSt
pytorch★ 387
github.com/ZJCV/ZCls
pytorch★ 143
github.com/YeongHyeon/ResNeSt-TF2
tf★ 67
github.com/Burf/ResNeSt-Tensorflow2
tf★ 25
github.com/STomoya/ResNeSt
pytorch★ 18
github.com/AnudeepKonda/MIMII_anamoly_detection
pytorch★ 3

Abstract

It is well known that featuremap attention and multi-path representation are important for visual recognition. In this paper, we present a modularized architecture, which applies the channel-wise attention on different network branches to leverage their success in capturing cross-feature interactions and learning diverse representations. Our design results in a simple and unified computation block, which can be parameterized using only a few variables. Our model, named ResNeSt, outperforms EfficientNet in accuracy and latency trade-off on image classification. In addition, ResNeSt has achieved superior transfer learning results on several public benchmarks serving as the backbone, and has been adopted by the winning entries of COCO-LVIS challenge. The source code for complete system and pretrained models are publicly available.

Tasks

image-classification Image Classification Instance Segmentation Object Detection Panoptic Segmentation Semantic Segmentation Transfer Learning

Benchmark Results

Dataset	Model	Metric	Claimed	Verified	Status
ImageNet	ResNeSt-269	Top 1 Accuracy	84.5	—	Unverified
ImageNet	ResNeSt-50-fast	Top 1 Accuracy	80.64	—	Unverified
ImageNet	ResNeSt-50	Top 1 Accuracy	81.13	—	Unverified
ImageNet	ResNeSt-101	Top 1 Accuracy	83	—	Unverified
ImageNet	ResNeSt-200	Top 1 Accuracy	83.9	—	Unverified

ResNeSt: Split-Attention Networks

Code

Abstract

Tasks

Benchmark Results

Reproductions