A survey of automated data augmentation algorithms for deep learning-based image classification tasks

被引:0
作者
Zihan Yang
Richard O. Sinnott
James Bailey
Qiuhong Ke
机构
[1] The University of Melbourne,Faculty of Engineering and Information Technology
[2] Monash University,Faculty of Information Technology
来源
Knowledge and Information Systems | 2023年 / 65卷
关键词
Automated data augmentation; Deep learning; Image classification; Big data;
D O I
暂无
中图分类号
学科分类号
摘要
In recent years, one of the most popular techniques in the computer vision community has been the deep learning technique. As a data-driven technique, deep model requires enormous amounts of accurately labelled training data, which is often inaccessible in many real-world applications. A data-space solution is Data Augmentation (DA), that can artificially generate new images out of original samples. Image augmentation strategies can vary by dataset, as different data types might require different augmentations to facilitate model training. However, the design of DA policies has been largely decided by the human experts with domain knowledge, which is considered to be highly subjective and error-prone. To mitigate such problem, a novel direction is to automatically learn the image augmentation policies from the given dataset using Automated Data Augmentation (AutoDA) techniques. The goal of AutoDA models is to find the optimal DA policies that can maximize the model performance gains. This survey discusses the underlying reasons of the emergence of AutoDA technology from the perspective of image classification. We identify three key components of a standard AutoDA model: a search space, a search algorithm and an evaluation function. Based on their architecture, we provide a systematic taxonomy of existing image AutoDA approaches. This paper presents the major works in AutoDA field, discussing their pros and cons, and proposing several potential directions for future improvements.
引用
收藏
页码:2805 / 2861
页数:56
相关论文
共 111 条
  • [21] Kong J-L(2019)Regularized evolution for image classifier architecture search Proc AAAI Conf Artif Intell 33 4780-21333
  • [22] Jin X-B(1996)Planning, learning and coordination in multiagent decision processes TARK 96 195-62
  • [23] Wang X-Y(2020)A group-theoretic framework for data augmentation Adv Neural Inf Process Syst 33 21321-115
  • [24] Su T-L(2020)Monte carlo gradient estimation in machine learning J Mach Learn Res 21 1-186136
  • [25] Zuo M(2021)Understanding deep learning (still) requires rethinking generalization Commun ACM 64 107-undefined
  • [26] Kamilaris A(2019)Shakedrop regularization for deep residual learning IEEE Access 7 186126-undefined
  • [27] Prenafeta-Boldú FX(undefined)undefined undefined undefined undefined-undefined
  • [28] Lemley J(undefined)undefined undefined undefined undefined-undefined
  • [29] Bazrafkan S(undefined)undefined undefined undefined undefined-undefined
  • [30] Corcoran P(undefined)undefined undefined undefined undefined-undefined