SVGA-Net: Sparse Voxel-Graph Attention Network for 3D Object Detection from Point Clouds

He, Qingdong; Wang, Zhengning; Zeng, Hao; Zeng, Yi; Liu, Yijun

Full-text links:

Download:

Current browse context:

cs.CV

< prev | next >

new | recent | 2006

Change to browse by:

Computer Science > Computer Vision and Pattern Recognition

Title: SVGA-Net: Sparse Voxel-Graph Attention Network for 3D Object Detection from Point Clouds

Authors: Qingdong He, Zhengning Wang, Hao Zeng, Yi Zeng, Yijun Liu

(Submitted on 7 Jun 2020 (v1), last revised 23 Dec 2021 (this version, v2))

Abstract: Accurate 3D object detection from point clouds has become a crucial component in autonomous driving. However, the volumetric representations and the projection methods in previous works fail to establish the relationships between the local point sets. In this paper, we propose Sparse Voxel-Graph Attention Network (SVGA-Net), a novel end-to-end trainable network which mainly contains voxel-graph module and sparse-to-dense regression module to achieve comparable 3D detection tasks from raw LIDAR data. Specifically, SVGA-Net constructs the local complete graph within each divided 3D spherical voxel and global KNN graph through all voxels. The local and global graphs serve as the attention mechanism to enhance the extracted features. In addition, the novel sparse-to-dense regression module enhances the 3D box estimation accuracy through feature maps aggregation at different levels. Experiments on KITTI detection benchmark demonstrate the efficiency of extending the graph representation to 3D object detection and the proposed SVGA-Net can achieve decent detection accuracy.

Comments:	9 pages, 4 figures
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2006.04043 [cs.CV]
	(or arXiv:2006.04043v2 [cs.CV] for this version)

Submission history

From: Qingdong He [view email]
[v1] Sun, 7 Jun 2020 05:01:06 GMT (462kb,D)
[v2] Thu, 23 Dec 2021 13:17:48 GMT (547kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2006.04043

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Computer Vision and Pattern Recognition

Title: SVGA-Net: Sparse Voxel-Graph Attention Network for 3D Object Detection from Point Clouds

Submission history