CISA: Context Substitution for Image Semantics Augmentation-Reference-Cited by-同舟云学术

CISA: Context Substitution for Image Semantics Augmentation

Published:2023-04-11 Issue:8 Volume:11 Page:1818
ISSN:2227-7390
Container-title:Mathematics
language:en
Short-container-title:Mathematics

Author:

Nesteruk Sergey¹^ORCID,Zherebtsov Ilya²,Illarionova Svetlana¹^ORCID,Shadrin Dmitrii¹³^ORCID,Somov Andrey¹^ORCID,Bezzateev Sergey V.⁴^ORCID,Yelina Tatiana⁴,Denisenko Vladimir²^ORCID,Oseledets Ivan¹^ORCID

Affiliation:

1. Skolkovo Institute of Science and Technology (Skoltech), 121205 Moscow, Russia

2. Voronezh State University of Engineering Technology (VSUET), 394036 Voronezh, Russia

3. Irkutsk National Research Technical University (INRTU), 664074 Irkutsk, Russia

4. Saint-Petrsburg State University of Aerospace Instrumentation (SUAI), 190000 Saint Petersburg, Russia

Abstract

Large datasets catalyze the rapid expansion of deep learning and computer vision. At the same time, in many domains, there is a lack of training data, which may become an obstacle for the practical application of deep computer vision models. To overcome this problem, it is popular to apply image augmentation. When a dataset contains instance segmentation masks, it is possible to apply instance-level augmentation. It operates by cutting an instance from the original image and pasting to new backgrounds. This article challenges a dataset with the same objects present in various domains. We introduce the Context Substitution for Image Semantics Augmentation framework (CISA), which is focused on choosing good background images. We compare several ways to find backgrounds that match the context of the test set, including Contrastive Language–Image Pre-Training (CLIP) image retrieval and diffusion image generation. We prove that our augmentation method is effective for classification, segmentation, and object detection with different dataset complexity and different model types. The average percentage increase in accuracy across all the tasks on a fruits and vegetables recognition dataset is 4.95%. Moreover, we show that the Fréchet Inception Distance (FID) metrics has a strong correlation with model accuracy, and it can help to choose better backgrounds without model training. The average negative correlation between model accuracy and the FID between the augmented and test datasets is 0.55 in our experiments.

Publisher

MDPI AG

Subject

General Mathematics,Engineering (miscellaneous),Computer Science (miscellaneous)

Link

https://www.mdpi.com/2227-7390/11/8/1818/pdf

Reference61 articles.

1. Effects of data set features on the performances of classification algorithms;Kwon;Expert Syst. Appl.,2013

2. Sbai, O., Couprie, C., and Aubry, M. (2020). Proceedings of the European Conference on Computer Vision, Springer.

3. How good is my test data? Introducing safety analysis for computer vision;Zendel;Int. J. Comput. Vis.,2017

4. Impact of dataset size and variety on the effectiveness of deep learning and transfer learning for plant disease classification;Barbedo;Comput. Electron. Agric.,2018

5. Zheng, S., Song, Y., Leung, T., and Goodfellow, I. (2016, January 27–30). Improving the robustness of deep neural networks via stability training. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.

Cited by 7 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. A Novel Applicable Shadow Resistant Neural Network Model for High-Efficiency Grid-Level Pavement Crack Detection;IEEE Transactions on Artificial Intelligence;2024-09

2. Machine Vision for Solid Waste Detection;Scanning Technologies for Autonomous Systems;2024

3. Apple Tree Health Recognition Through the Application of Transfer Learning for UAV Imagery;2023 IEEE 28th International Conference on Emerging Technologies and Factory Automation (ETFA);2023-09-12

4. MineralImage5k: A benchmark for zero-shot raw mineral visual recognition and description;Computers & Geosciences;2023-09

5. Enabling Multi-Part Plant Segmentation with Instance-Level Augmentation Using Weak Annotations;Information;2023-07-03