RF-Next: Efficient Receptive Field Search for Convolutional Neural Networks

ShangHua Gao, Zhong-Yu Li, Qi Han, Ming-Ming Cheng, Liang Wang

2022-06-14Action Segmentation Temporal Action Segmentation Segmentation Semantic Segmentation Speech Synthesis Instance Segmentation object-detection Object Detection

Paper PDF Code Code(official)

Abstract

Temporal/spatial receptive fields of models play an important role in sequential/spatial tasks. Large receptive fields facilitate long-term relations, while small receptive fields help to capture the local details. Existing methods construct models with hand-designed receptive fields in layers. Can we effectively search for receptive field combinations to replace hand-designed patterns? To answer this question, we propose to find better receptive field combinations through a global-to-local search scheme. Our search scheme exploits both global search to find the coarse combinations and local search to get the refined receptive field combinations further. The global search finds possible coarse combinations other than human-designed patterns. On top of the global search, we propose an expectation-guided iterative local search scheme to refine combinations effectively. Our RF-Next models, plugging receptive field search to various models, boost the performance on many tasks, e.g., temporal action segmentation, object detection, instance segmentation, and speech synthesis. The source code is publicly available on http://mmcheng.net/rfnext.

Results

Task	Dataset	Metric	Value	Model
Semantic Segmentation	ImageNet-S	mIoU (test)	51.1	RF-ConvNext-Tiny (rfmerge, P4, 224x224, SUP)
Semantic Segmentation	ImageNet-S	mIoU (val)	51.3	RF-ConvNext-Tiny (rfmerge, P4, 224x224, SUP)
Semantic Segmentation	ImageNet-S	mIoU (test)	50.5	RF-ConvNext-Tiny (rfmultiple, P4, 224x224, SUP)
Semantic Segmentation	ImageNet-S	mIoU (val)	50.8	RF-ConvNext-Tiny (rfmultiple, P4, 224x224, SUP)
Semantic Segmentation	ImageNet-S	mIoU (test)	50.5	RF-ConvNext-Tiny (rfsingle, P4, 224x224, SUP)
Semantic Segmentation	ImageNet-S	mIoU (val)	50.7	RF-ConvNext-Tiny (rfsingle, P4, 224x224, SUP)
Action Localization	Breakfast	Acc	70.8	RF++-SSTDA
Object Detection	COCO 2017 val	AP	50.9	RF-ConvNeXt-T Cascade R-CNN
3D	COCO 2017 val	AP	50.9	RF-ConvNeXt-T Cascade R-CNN
Instance Segmentation	COCO 2017 val	AP	44.3	RF-ConvNeXt-T Cascade R-CNN
Action Segmentation	Breakfast	Acc	70.8	RF++-SSTDA
2D Classification	COCO 2017 val	AP	50.9	RF-ConvNeXt-T Cascade R-CNN
2D Object Detection	COCO 2017 val	AP	50.9	RF-ConvNeXt-T Cascade R-CNN
10-shot image generation	ImageNet-S	mIoU (test)	51.1	RF-ConvNext-Tiny (rfmerge, P4, 224x224, SUP)
10-shot image generation	ImageNet-S	mIoU (val)	51.3	RF-ConvNext-Tiny (rfmerge, P4, 224x224, SUP)
10-shot image generation	ImageNet-S	mIoU (test)	50.5	RF-ConvNext-Tiny (rfmultiple, P4, 224x224, SUP)
10-shot image generation	ImageNet-S	mIoU (val)	50.8	RF-ConvNext-Tiny (rfmultiple, P4, 224x224, SUP)
10-shot image generation	ImageNet-S	mIoU (test)	50.5	RF-ConvNext-Tiny (rfsingle, P4, 224x224, SUP)
10-shot image generation	ImageNet-S	mIoU (val)	50.7	RF-ConvNext-Tiny (rfsingle, P4, 224x224, SUP)
16k	COCO 2017 val	AP	50.9	RF-ConvNeXt-T Cascade R-CNN

RF-Next: Efficient Receptive Field Search for Convolutional Neural Networks

Abstract

Results

Related Papers

RF-Next: Efficient Receptive Field Search for Convolutional Neural Networks

Abstract

Results

Related Papers