Hang Zhang, Chongruo wu, Zhongyue Zhang, Yi Zhu, Haibin Lin, Zhi Zhang, Yue Sun, Tong He, Jonas Mueller, R. Manmatha, Mu Li, Alexander Smola
It is well known that featuremap attention and multi-path representation are important for visual recognition. In this paper, we present a modularized architecture, which applies the channel-wise attention on different network branches to leverage their success in capturing cross-feature interactions and learning diverse representations. Our design results in a simple and unified computation block, which can be parameterized using only a few variables. Our model, named ResNeSt, outperforms EfficientNet in accuracy and latency trade-off on image classification. In addition, ResNeSt has achieved superior transfer learning results on several public benchmarks serving as the backbone, and has been adopted by the winning entries of COCO-LVIS challenge. The source code for complete system and pretrained models are publicly available.
| Task | Dataset | Metric | Value | Model |
|---|---|---|---|---|
| Semantic Segmentation | Cityscapes val | mIoU | 82.7 | ResNeSt-200 |
| Semantic Segmentation | ADE20K val | mIoU | 48.36 | ResNeSt-200 |
| Semantic Segmentation | ADE20K val | mIoU | 47.6 | ResNeSt-269 |
| Semantic Segmentation | ADE20K val | mIoU | 46.91 | ResNeSt-101 |
| Semantic Segmentation | PASCAL Context | mIoU | 58.9 | ResNeSt-269 |
| Semantic Segmentation | PASCAL Context | mIoU | 58.4 | ResNeSt-200 |
| Semantic Segmentation | PASCAL Context | mIoU | 56.5 | ResNeSt-101 |
| Semantic Segmentation | DADA-seg | mIoU | 19.99 | ResNeSt (ResNeSt-101) |
| Semantic Segmentation | ADE20K | Validation mIoU | 48.36 | ResNeSt-200 |
| Semantic Segmentation | ADE20K | Validation mIoU | 47.6 | ResNeSt-269 |
| Semantic Segmentation | ADE20K | Validation mIoU | 46.91 | ResNeSt-101 |
| Semantic Segmentation | COCO minival | PQ | 47.9 | PanopticFPN+ResNeSt(single-scale) |
| Semantic Segmentation | COCO minival | PQst | 37 | PanopticFPN+ResNeSt(single-scale) |
| Semantic Segmentation | COCO minival | PQth | 55.1 | PanopticFPN+ResNeSt(single-scale) |
| Object Detection | COCO test-dev | AP50 | 72 | ResNeSt-200 (multi-scale) |
| Object Detection | COCO test-dev | AP75 | 58 | ResNeSt-200 (multi-scale) |
| Object Detection | COCO test-dev | APL | 66.8 | ResNeSt-200 (multi-scale) |
| Object Detection | COCO test-dev | APM | 56.2 | ResNeSt-200 (multi-scale) |
| Object Detection | COCO test-dev | APS | 35.1 | ResNeSt-200 (multi-scale) |
| Object Detection | COCO test-dev | box mAP | 53.3 | ResNeSt-200 (multi-scale) |
| Object Detection | COCO minival | AP50 | 71 | ResNeSt-200 (multi-scale) |
| Object Detection | COCO minival | AP75 | 57.07 | ResNeSt-200 (multi-scale) |
| Object Detection | COCO minival | APL | 66.29 | ResNeSt-200 (multi-scale) |
| Object Detection | COCO minival | APM | 56.36 | ResNeSt-200 (multi-scale) |
| Object Detection | COCO minival | APS | 36.8 | ResNeSt-200 (multi-scale) |
| Object Detection | COCO minival | box AP | 52.47 | ResNeSt-200 (multi-scale) |
| Object Detection | COCO minival | AP50 | 69.53 | ResNeSt-200-DCN (single-scale) |
| Object Detection | COCO minival | AP75 | 55.4 | ResNeSt-200-DCN (single-scale) |
| Object Detection | COCO minival | APL | 65.83 | ResNeSt-200-DCN (single-scale) |
| Object Detection | COCO minival | APM | 54.66 | ResNeSt-200-DCN (single-scale) |
| Object Detection | COCO minival | APS | 32.67 | ResNeSt-200-DCN (single-scale) |
| Object Detection | COCO minival | box AP | 50.91 | ResNeSt-200-DCN (single-scale) |
| Object Detection | COCO minival | AP50 | 68.78 | ResNeSt-200 (single-scale) |
| Object Detection | COCO minival | AP75 | 55.17 | ResNeSt-200 (single-scale) |
| Object Detection | COCO minival | APL | 63.9 | ResNeSt-200 (single-scale) |
| Object Detection | COCO minival | APM | 54.2 | ResNeSt-200 (single-scale) |
| Object Detection | COCO minival | box AP | 50.54 | ResNeSt-200 (single-scale) |
| Image Classification | ImageNet | GFLOPs | 5.39 | ResNeSt-50 |
| Image Classification | ImageNet | GFLOPs | 4.34 | ResNeSt-50-fast |
| 3D | COCO test-dev | AP50 | 72 | ResNeSt-200 (multi-scale) |
| 3D | COCO test-dev | AP75 | 58 | ResNeSt-200 (multi-scale) |
| 3D | COCO test-dev | APL | 66.8 | ResNeSt-200 (multi-scale) |
| 3D | COCO test-dev | APM | 56.2 | ResNeSt-200 (multi-scale) |
| 3D | COCO test-dev | APS | 35.1 | ResNeSt-200 (multi-scale) |
| 3D | COCO test-dev | box mAP | 53.3 | ResNeSt-200 (multi-scale) |
| 3D | COCO minival | AP50 | 71 | ResNeSt-200 (multi-scale) |
| 3D | COCO minival | AP75 | 57.07 | ResNeSt-200 (multi-scale) |
| 3D | COCO minival | APL | 66.29 | ResNeSt-200 (multi-scale) |
| 3D | COCO minival | APM | 56.36 | ResNeSt-200 (multi-scale) |
| 3D | COCO minival | APS | 36.8 | ResNeSt-200 (multi-scale) |
| 3D | COCO minival | box AP | 52.47 | ResNeSt-200 (multi-scale) |
| 3D | COCO minival | AP50 | 69.53 | ResNeSt-200-DCN (single-scale) |
| 3D | COCO minival | AP75 | 55.4 | ResNeSt-200-DCN (single-scale) |
| 3D | COCO minival | APL | 65.83 | ResNeSt-200-DCN (single-scale) |
| 3D | COCO minival | APM | 54.66 | ResNeSt-200-DCN (single-scale) |
| 3D | COCO minival | APS | 32.67 | ResNeSt-200-DCN (single-scale) |
| 3D | COCO minival | box AP | 50.91 | ResNeSt-200-DCN (single-scale) |
| 3D | COCO minival | AP50 | 68.78 | ResNeSt-200 (single-scale) |
| 3D | COCO minival | AP75 | 55.17 | ResNeSt-200 (single-scale) |
| 3D | COCO minival | APL | 63.9 | ResNeSt-200 (single-scale) |
| 3D | COCO minival | APM | 54.2 | ResNeSt-200 (single-scale) |
| 3D | COCO minival | box AP | 50.54 | ResNeSt-200 (single-scale) |
| Instance Segmentation | COCO minival | mask AP | 46.25 | ResNeSt-200 (multi-scale) |
| Instance Segmentation | COCO minival | mask AP | 44.5 | ResNeSt-200-DCN (single-scale) |
| Instance Segmentation | COCO minival | mask AP | 44.21 | ResNeSt-200 (single-scale) |
| Instance Segmentation | COCO minival | mask AP | 41.56 | ResNeSt-101 (single-scale) |
| Instance Segmentation | COCO test-dev | AP50 | 70.2 | ResNeSt-200 (multi-scale) |
| Instance Segmentation | COCO test-dev | AP75 | 51.5 | ResNeSt-200 (multi-scale) |
| Instance Segmentation | COCO test-dev | APL | 60.6 | ResNeSt-200 (multi-scale) |
| Instance Segmentation | COCO test-dev | APM | 49.6 | ResNeSt-200 (multi-scale) |
| Instance Segmentation | COCO test-dev | APS | 30 | ResNeSt-200 (multi-scale) |
| 2D Classification | COCO test-dev | AP50 | 72 | ResNeSt-200 (multi-scale) |
| 2D Classification | COCO test-dev | AP75 | 58 | ResNeSt-200 (multi-scale) |
| 2D Classification | COCO test-dev | APL | 66.8 | ResNeSt-200 (multi-scale) |
| 2D Classification | COCO test-dev | APM | 56.2 | ResNeSt-200 (multi-scale) |
| 2D Classification | COCO test-dev | APS | 35.1 | ResNeSt-200 (multi-scale) |
| 2D Classification | COCO test-dev | box mAP | 53.3 | ResNeSt-200 (multi-scale) |
| 2D Classification | COCO minival | AP50 | 71 | ResNeSt-200 (multi-scale) |
| 2D Classification | COCO minival | AP75 | 57.07 | ResNeSt-200 (multi-scale) |
| 2D Classification | COCO minival | APL | 66.29 | ResNeSt-200 (multi-scale) |
| 2D Classification | COCO minival | APM | 56.36 | ResNeSt-200 (multi-scale) |
| 2D Classification | COCO minival | APS | 36.8 | ResNeSt-200 (multi-scale) |
| 2D Classification | COCO minival | box AP | 52.47 | ResNeSt-200 (multi-scale) |
| 2D Classification | COCO minival | AP50 | 69.53 | ResNeSt-200-DCN (single-scale) |
| 2D Classification | COCO minival | AP75 | 55.4 | ResNeSt-200-DCN (single-scale) |
| 2D Classification | COCO minival | APL | 65.83 | ResNeSt-200-DCN (single-scale) |
| 2D Classification | COCO minival | APM | 54.66 | ResNeSt-200-DCN (single-scale) |
| 2D Classification | COCO minival | APS | 32.67 | ResNeSt-200-DCN (single-scale) |
| 2D Classification | COCO minival | box AP | 50.91 | ResNeSt-200-DCN (single-scale) |
| 2D Classification | COCO minival | AP50 | 68.78 | ResNeSt-200 (single-scale) |
| 2D Classification | COCO minival | AP75 | 55.17 | ResNeSt-200 (single-scale) |
| 2D Classification | COCO minival | APL | 63.9 | ResNeSt-200 (single-scale) |
| 2D Classification | COCO minival | APM | 54.2 | ResNeSt-200 (single-scale) |
| 2D Classification | COCO minival | box AP | 50.54 | ResNeSt-200 (single-scale) |
| 2D Object Detection | COCO test-dev | AP50 | 72 | ResNeSt-200 (multi-scale) |
| 2D Object Detection | COCO test-dev | AP75 | 58 | ResNeSt-200 (multi-scale) |
| 2D Object Detection | COCO test-dev | APL | 66.8 | ResNeSt-200 (multi-scale) |
| 2D Object Detection | COCO test-dev | APM | 56.2 | ResNeSt-200 (multi-scale) |
| 2D Object Detection | COCO test-dev | APS | 35.1 | ResNeSt-200 (multi-scale) |
| 2D Object Detection | COCO test-dev | box mAP | 53.3 | ResNeSt-200 (multi-scale) |
| 2D Object Detection | COCO minival | AP50 | 71 | ResNeSt-200 (multi-scale) |
| 2D Object Detection | COCO minival | AP75 | 57.07 | ResNeSt-200 (multi-scale) |
| 2D Object Detection | COCO minival | APL | 66.29 | ResNeSt-200 (multi-scale) |
| 2D Object Detection | COCO minival | APM | 56.36 | ResNeSt-200 (multi-scale) |
| 2D Object Detection | COCO minival | APS | 36.8 | ResNeSt-200 (multi-scale) |
| 2D Object Detection | COCO minival | box AP | 52.47 | ResNeSt-200 (multi-scale) |
| 2D Object Detection | COCO minival | AP50 | 69.53 | ResNeSt-200-DCN (single-scale) |
| 2D Object Detection | COCO minival | AP75 | 55.4 | ResNeSt-200-DCN (single-scale) |
| 2D Object Detection | COCO minival | APL | 65.83 | ResNeSt-200-DCN (single-scale) |
| 2D Object Detection | COCO minival | APM | 54.66 | ResNeSt-200-DCN (single-scale) |
| 2D Object Detection | COCO minival | APS | 32.67 | ResNeSt-200-DCN (single-scale) |
| 2D Object Detection | COCO minival | box AP | 50.91 | ResNeSt-200-DCN (single-scale) |
| 2D Object Detection | COCO minival | AP50 | 68.78 | ResNeSt-200 (single-scale) |
| 2D Object Detection | COCO minival | AP75 | 55.17 | ResNeSt-200 (single-scale) |
| 2D Object Detection | COCO minival | APL | 63.9 | ResNeSt-200 (single-scale) |
| 2D Object Detection | COCO minival | APM | 54.2 | ResNeSt-200 (single-scale) |
| 2D Object Detection | COCO minival | box AP | 50.54 | ResNeSt-200 (single-scale) |
| 10-shot image generation | Cityscapes val | mIoU | 82.7 | ResNeSt-200 |
| 10-shot image generation | ADE20K val | mIoU | 48.36 | ResNeSt-200 |
| 10-shot image generation | ADE20K val | mIoU | 47.6 | ResNeSt-269 |
| 10-shot image generation | ADE20K val | mIoU | 46.91 | ResNeSt-101 |
| 10-shot image generation | PASCAL Context | mIoU | 58.9 | ResNeSt-269 |
| 10-shot image generation | PASCAL Context | mIoU | 58.4 | ResNeSt-200 |
| 10-shot image generation | PASCAL Context | mIoU | 56.5 | ResNeSt-101 |
| 10-shot image generation | DADA-seg | mIoU | 19.99 | ResNeSt (ResNeSt-101) |
| 10-shot image generation | ADE20K | Validation mIoU | 48.36 | ResNeSt-200 |
| 10-shot image generation | ADE20K | Validation mIoU | 47.6 | ResNeSt-269 |
| 10-shot image generation | ADE20K | Validation mIoU | 46.91 | ResNeSt-101 |
| 10-shot image generation | COCO minival | PQ | 47.9 | PanopticFPN+ResNeSt(single-scale) |
| 10-shot image generation | COCO minival | PQst | 37 | PanopticFPN+ResNeSt(single-scale) |
| 10-shot image generation | COCO minival | PQth | 55.1 | PanopticFPN+ResNeSt(single-scale) |
| Panoptic Segmentation | COCO minival | PQ | 47.9 | PanopticFPN+ResNeSt(single-scale) |
| Panoptic Segmentation | COCO minival | PQst | 37 | PanopticFPN+ResNeSt(single-scale) |
| Panoptic Segmentation | COCO minival | PQth | 55.1 | PanopticFPN+ResNeSt(single-scale) |
| 16k | COCO test-dev | AP50 | 72 | ResNeSt-200 (multi-scale) |
| 16k | COCO test-dev | AP75 | 58 | ResNeSt-200 (multi-scale) |
| 16k | COCO test-dev | APL | 66.8 | ResNeSt-200 (multi-scale) |
| 16k | COCO test-dev | APM | 56.2 | ResNeSt-200 (multi-scale) |
| 16k | COCO test-dev | APS | 35.1 | ResNeSt-200 (multi-scale) |
| 16k | COCO test-dev | box mAP | 53.3 | ResNeSt-200 (multi-scale) |
| 16k | COCO minival | AP50 | 71 | ResNeSt-200 (multi-scale) |
| 16k | COCO minival | AP75 | 57.07 | ResNeSt-200 (multi-scale) |
| 16k | COCO minival | APL | 66.29 | ResNeSt-200 (multi-scale) |
| 16k | COCO minival | APM | 56.36 | ResNeSt-200 (multi-scale) |
| 16k | COCO minival | APS | 36.8 | ResNeSt-200 (multi-scale) |
| 16k | COCO minival | box AP | 52.47 | ResNeSt-200 (multi-scale) |
| 16k | COCO minival | AP50 | 69.53 | ResNeSt-200-DCN (single-scale) |
| 16k | COCO minival | AP75 | 55.4 | ResNeSt-200-DCN (single-scale) |
| 16k | COCO minival | APL | 65.83 | ResNeSt-200-DCN (single-scale) |
| 16k | COCO minival | APM | 54.66 | ResNeSt-200-DCN (single-scale) |
| 16k | COCO minival | APS | 32.67 | ResNeSt-200-DCN (single-scale) |
| 16k | COCO minival | box AP | 50.91 | ResNeSt-200-DCN (single-scale) |
| 16k | COCO minival | AP50 | 68.78 | ResNeSt-200 (single-scale) |
| 16k | COCO minival | AP75 | 55.17 | ResNeSt-200 (single-scale) |
| 16k | COCO minival | APL | 63.9 | ResNeSt-200 (single-scale) |
| 16k | COCO minival | APM | 54.2 | ResNeSt-200 (single-scale) |
| 16k | COCO minival | box AP | 50.54 | ResNeSt-200 (single-scale) |