SVGA-Net: Sparse Voxel-Graph Attention Network for 3D Object Detection from Point Clouds

Qingdong He, Zhengning Wang, Hao Zeng, Yi Zeng, Yijun Liu

2020-06-07regression Autonomous Driving object-detection 3D Object Detection Object Detection Graph Attention

Abstract

Accurate 3D object detection from point clouds has become a crucial component in autonomous driving. However, the volumetric representations and the projection methods in previous works fail to establish the relationships between the local point sets. In this paper, we propose Sparse Voxel-Graph Attention Network (SVGA-Net), a novel end-to-end trainable network which mainly contains voxel-graph module and sparse-to-dense regression module to achieve comparable 3D detection tasks from raw LIDAR data. Specifically, SVGA-Net constructs the local complete graph within each divided 3D spherical voxel and global KNN graph through all voxels. The local and global graphs serve as the attention mechanism to enhance the extracted features. In addition, the novel sparse-to-dense regression module enhances the 3D box estimation accuracy through feature maps aggregation at different levels. Experiments on KITTI detection benchmark demonstrate the efficiency of extending the graph representation to 3D object detection and the proposed SVGA-Net can achieve decent detection accuracy.

Results

Task	Dataset	Metric	Value	Model
Object Detection	KITTI Cars Hard val	AP	79.15	SVGA-Net
Object Detection	KITTI Cars Moderate val	AP	80.23	SVGA-Net
Object Detection	KITTI Cars Easy val	AP	90.59	SVGA-Net
3D	KITTI Cars Hard val	AP	79.15	SVGA-Net
3D	KITTI Cars Moderate val	AP	80.23	SVGA-Net
3D	KITTI Cars Easy val	AP	90.59	SVGA-Net
3D Object Detection	KITTI Cars Hard val	AP	79.15	SVGA-Net
3D Object Detection	KITTI Cars Moderate val	AP	80.23	SVGA-Net
3D Object Detection	KITTI Cars Easy val	AP	90.59	SVGA-Net
2D Classification	KITTI Cars Hard val	AP	79.15	SVGA-Net
2D Classification	KITTI Cars Moderate val	AP	80.23	SVGA-Net
2D Classification	KITTI Cars Easy val	AP	90.59	SVGA-Net
2D Object Detection	KITTI Cars Hard val	AP	79.15	SVGA-Net
2D Object Detection	KITTI Cars Moderate val	AP	80.23	SVGA-Net
2D Object Detection	KITTI Cars Easy val	AP	90.59	SVGA-Net
16k	KITTI Cars Hard val	AP	79.15	SVGA-Net
16k	KITTI Cars Moderate val	AP	80.23	SVGA-Net
16k	KITTI Cars Easy val	AP	90.59	SVGA-Net

SVGA-Net: Sparse Voxel-Graph Attention Network for 3D Object Detection from Point Clouds

Abstract

Results

Related Papers

SVGA-Net: Sparse Voxel-Graph Attention Network for 3D Object Detection from Point Clouds

Abstract

Results

Related Papers