CenterFormer: Center-based Transformer for 3D Object Detection

Zixiang Zhou, Xiangchen Zhao, Yu Wang, Panqu Wang, Hassan Foroosh

2022-09-12object-detection 3D Object Detection Object Detection

Abstract

Query-based transformer has shown great potential in constructing long-range attention in many image-domain tasks, but has rarely been considered in LiDAR-based 3D object detection due to the overwhelming size of the point cloud data. In this paper, we propose CenterFormer, a center-based transformer network for 3D object detection. CenterFormer first uses a center heatmap to select center candidates on top of a standard voxel-based point cloud encoder. It then uses the feature of the center candidate as the query embedding in the transformer. To further aggregate features from multiple frames, we design an approach to fuse features through cross-attention. Lastly, regression heads are added to predict the bounding box on the output center feature representation. Our design reduces the convergence difficulty and computational complexity of the transformer structure. The results show significant improvements over the strong baseline of anchor-free object detection networks. CenterFormer achieves state-of-the-art performance for a single model on the Waymo Open Dataset, with 73.7% mAPH on the validation set and 75.6% mAPH on the test set, significantly outperforming all previously published CNN and transformer-based methods. Our code is publicly available at https://github.com/TuSimple/centerformer

Results

Task	Dataset	Metric	Value	Model
Object Detection	Waymo Open Dataset	mAPH/L2	68.9	CenterFormer
Object Detection	waymo cyclist	APH/L2	73.3	CenterFormer
Object Detection	waymo vehicle	APH/L2	73.8	CenterFormer
Object Detection	waymo pedestrian	APH/L2	75	CenterFormer
3D	Waymo Open Dataset	mAPH/L2	68.9	CenterFormer
3D	waymo cyclist	APH/L2	73.3	CenterFormer
3D	waymo vehicle	APH/L2	73.8	CenterFormer
3D	waymo pedestrian	APH/L2	75	CenterFormer
3D Object Detection	Waymo Open Dataset	mAPH/L2	68.9	CenterFormer
3D Object Detection	waymo cyclist	APH/L2	73.3	CenterFormer
3D Object Detection	waymo vehicle	APH/L2	73.8	CenterFormer
3D Object Detection	waymo pedestrian	APH/L2	75	CenterFormer
2D Classification	Waymo Open Dataset	mAPH/L2	68.9	CenterFormer
2D Classification	waymo cyclist	APH/L2	73.3	CenterFormer
2D Classification	waymo vehicle	APH/L2	73.8	CenterFormer
2D Classification	waymo pedestrian	APH/L2	75	CenterFormer
2D Object Detection	Waymo Open Dataset	mAPH/L2	68.9	CenterFormer
2D Object Detection	waymo cyclist	APH/L2	73.3	CenterFormer
2D Object Detection	waymo vehicle	APH/L2	73.8	CenterFormer
2D Object Detection	waymo pedestrian	APH/L2	75	CenterFormer
16k	Waymo Open Dataset	mAPH/L2	68.9	CenterFormer
16k	waymo cyclist	APH/L2	73.3	CenterFormer
16k	waymo vehicle	APH/L2	73.8	CenterFormer
16k	waymo pedestrian	APH/L2	75	CenterFormer

CenterFormer: Center-based Transformer for 3D Object Detection

Abstract

Results

Related Papers

CenterFormer: Center-based Transformer for 3D Object Detection

Abstract

Results

Related Papers