Expandable YOLO: 3D Object Detection from RGB-D Images

Takahashi, Masahiro; Moro, Alessandro; Ji, Yonghoon; Umeda, Kazunori

Full-text links:

Download:

PDF only

Current browse context:

cs.CV

< prev | next >

new | recent | 2006

Change to browse by:

Computer Science > Computer Vision and Pattern Recognition

Title: Expandable YOLO: 3D Object Detection from RGB-D Images

Authors: Masahiro Takahashi, Alessandro Moro, Yonghoon Ji, Kazunori Umeda

(Submitted on 26 Jun 2020)

Abstract: This paper aims at constructing a light-weight object detector that inputs a depth and a color image from a stereo camera. Specifically, by extending the network architecture of YOLOv3 to 3D in the middle, it is possible to output in the depth direction. In addition, Intersection over Uninon (IoU) in 3D space is introduced to confirm the accuracy of region extraction results. In the field of deep learning, object detectors that use distance information as input are actively studied for utilizing automated driving. However, the conventional detector has a large network structure, and the real-time property is impaired. The effectiveness of the detector constructed as described above is verified using datasets. As a result of this experiment, the proposed model is able to output 3D bounding boxes and detect people whose part of the body is hidden. Further, the processing speed of the model is 44.35 fps.

Comments:	5 pages, 8 figures
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2006.14837 [cs.CV]
	(or arXiv:2006.14837v1 [cs.CV] for this version)

Submission history

From: Masahiro Takahashi [view email]
[v1] Fri, 26 Jun 2020 07:32:30 GMT (843kb)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2006.14837

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Computer Vision and Pattern Recognition

Title: Expandable YOLO: 3D Object Detection from RGB-D Images

Submission history