We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.CV

Change to browse by:

cs

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Computer Science > Computer Vision and Pattern Recognition

Title: A Survey of Automated Data Augmentation Algorithms for Deep Learning-based Image Classification Tasks

Abstract: In recent years, one of the most popular techniques in the computer vision community has been the deep learning technique. As a data-driven technique, deep model requires enormous amounts of accurately labelled training data, which is often inaccessible in many real-world applications. A data-space solution is Data Augmentation (DA), that can artificially generate new images out of original samples. Image augmentation strategies can vary by dataset, as different data types might require different augmentations to facilitate model training. However, the design of DA policies has been largely decided by the human experts with domain knowledge, which is considered to be highly subjective and error-prone. To mitigate such problem, a novel direction is to automatically learn the image augmentation policies from the given dataset using Automated Data Augmentation (AutoDA) techniques. The goal of AutoDA models is to find the optimal DA policies that can maximize the model performance gains. This survey discusses the underlying reasons of the emergence of AutoDA technology from the perspective of image classification. We identify three key components of a standard AutoDA model: a search space, a search algorithm and an evaluation function. Based on their architecture, we provide a systematic taxonomy of existing image AutoDA approaches. This paper presents the major works in AutoDA field, discussing their pros and cons, and proposing several potential directions for future improvements.
Comments: 68 pages, 9 figures. Submitted to Knowledge and Information Systems (KAIS)
Subjects: Computer Vision and Pattern Recognition (cs.CV)
MSC classes: A.1, I.4.3, I.5.2
Cite as: arXiv:2206.06544 [cs.CV]
  (or arXiv:2206.06544v2 [cs.CV] for this version)

Submission history

From: Zihan Yang [view email]
[v1] Tue, 14 Jun 2022 01:40:09 GMT (1330kb)
[v2] Thu, 6 Oct 2022 23:49:31 GMT (1330kb)

Link back to: arXiv, form interface, contact.