We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.CV

Change to browse by:

cs

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Computer Science > Computer Vision and Pattern Recognition

Title: Box2Mask: Weakly Supervised 3D Semantic Instance Segmentation Using Bounding Boxes

Abstract: Current 3D segmentation methods heavily rely on large-scale point-cloud datasets, which are notoriously laborious to annotate. Few attempts have been made to circumvent the need for dense per-point annotations. In this work, we look at weakly-supervised 3D semantic instance segmentation. The key idea is to leverage 3D bounding box labels which are easier and faster to annotate. Indeed, we show that it is possible to train dense segmentation models using only bounding box labels. At the core of our method, \name{}, lies a deep model, inspired by classical Hough voting, that directly votes for bounding box parameters, and a clustering method specifically tailored to bounding box votes. This goes beyond commonly used center votes, which would not fully exploit the bounding box annotations. On ScanNet test, our weakly supervised model attains leading performance among other weakly supervised approaches (+18 mAP@50). Remarkably, it also achieves 97% of the mAP@50 score of current fully supervised models. To further illustrate the practicality of our work, we train Box2Mask on the recently released ARKitScenes dataset which is annotated with 3D bounding boxes only, and show, for the first time, compelling 3D instance segmentation masks.
Comments: Project page: this https URL
Subjects: Computer Vision and Pattern Recognition (cs.CV)
Journal reference: European Conference on Computer Vision (ECCV), 2022, Oral Presentation
Cite as: arXiv:2206.01203 [cs.CV]
  (or arXiv:2206.01203v3 [cs.CV] for this version)

Submission history

From: Julian Chibane [view email]
[v1] Thu, 2 Jun 2022 17:59:57 GMT (8616kb,D)
[v2] Tue, 26 Jul 2022 10:35:09 GMT (8674kb,D)
[v3] Tue, 31 Oct 2023 16:50:00 GMT (8673kb,D)

Link back to: arXiv, form interface, contact.