We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.CV

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Computer Science > Computer Vision and Pattern Recognition

Title: Classifier-Guided Visual Correction of Noisy Labels for Image Classification Tasks

Abstract: Training data plays an essential role in modern applications of machine learning. However, gathering labeled training data is time-consuming. Therefore, labeling is often outsourced to less experienced users, or completely automated. This can introduce errors, which compromise valuable training data, and lead to suboptimal training results. We thus propose a novel approach that uses the power of pretrained classifiers to visually guide users to noisy labels, and let them interactively check error candidates, to iteratively improve the training data set. To systematically investigate training data, we propose a categorization of labeling errors into three different types, based on an analysis of potential pitfalls in label acquisition processes. For each of these types, we present approaches to detect, reason about, and resolve error candidates, as we propose measures and visual guidance techniques to support machine learning users. Our approach has been used to spot errors in well-known machine learning benchmark data sets, and we tested its usability during a user evaluation. While initially developed for images, the techniques presented in this paper are independent of the classification algorithm, and can also be extended to many other types of training data.
Subjects: Computer Vision and Pattern Recognition (cs.CV); Human-Computer Interaction (cs.HC); Machine Learning (cs.LG)
DOI: 10.1111/cgf.13973
Cite as: arXiv:1808.03114 [cs.CV]
  (or arXiv:1808.03114v4 [cs.CV] for this version)

Submission history

From: Alex Bäuerle [view email]
[v1] Thu, 9 Aug 2018 12:34:33 GMT (3980kb,D)
[v2] Fri, 6 Dec 2019 12:07:13 GMT (1335kb,D)
[v3] Thu, 2 Apr 2020 12:20:21 GMT (1338kb,D)
[v4] Mon, 6 Apr 2020 13:55:12 GMT (1338kb,D)

Link back to: arXiv, form interface, contact.