Survey: Image Mixing and Deleting for Data Augmentation

Naveed, Humza; Anwar, Saeed; Hayat, Munawar; Javed, Kashif; Mian, Ajmal

Computer Science > Computer Vision and Pattern Recognition

arXiv:2106.07085 (cs)

[Submitted on 13 Jun 2021 (v1), last revised 6 Feb 2023 (this version, v4)]

Title:Survey: Image Mixing and Deleting for Data Augmentation

Authors:Humza Naveed, Saeed Anwar, Munawar Hayat, Kashif Javed, Ajmal Mian

View PDF

Abstract:Neural networks are prone to overfitting and memorizing data patterns. To avoid over-fitting and enhance their generalization and performance, various methods have been suggested in the literature, including dropout, regularization, label smoothing, etc. One such method is augmentation which introduces different types of corruption in the data to prevent the model from overfitting and to memorize patterns present in the data. A sub-area of data augmentation is image mixing and deleting. This specific type of augmentation either deletes image regions or mixes two images to hide or make particular characteristics of images confusing for the network, forcing it to emphasize the overall structure of the object in an image. Models trained with this approach have proven to perform and generalize well compared to those trained without image mixing or deleting. An added benefit that comes with this method of training is robustness against image corruption. Due to its low computational cost and recent success, researchers have proposed many image mixing and deleting techniques. We furnish an in-depth survey of image mixing and deleting techniques and provide categorization via their most distinguishing features. We initiate our discussion with some fundamental relevant concepts. Next, we present essentials, such as each category's strengths and limitations, describing their working mechanism, basic formulations, and applications. We also discuss the general challenges and recommend possible future research directions for image mixing and deleting data augmentation techniques. Datasets and codes for evaluation are publicly available here.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2106.07085 [cs.CV]
	(or arXiv:2106.07085v4 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2106.07085

Submission history

From: Humza Naveed [view email]
[v1] Sun, 13 Jun 2021 20:32:24 UTC (2,810 KB)
[v2] Mon, 1 Nov 2021 18:53:37 UTC (3,146 KB)
[v3] Wed, 25 Jan 2023 19:37:57 UTC (28,105 KB)
[v4] Mon, 6 Feb 2023 19:21:18 UTC (28,276 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Survey: Image Mixing and Deleting for Data Augmentation

Submission history

Access Paper:

References & Citations

1 blog link

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Survey: Image Mixing and Deleting for Data Augmentation

Submission history

Access Paper:

References & Citations

1 blog link

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators