PLACE Dropout: A Progressive Layer-wise and Channel-wise Dropout for Domain Generalization-Reference-Cited by-同舟云学术

PLACE Dropout: A Progressive Layer-wise and Channel-wise Dropout for Domain Generalization

Published:2023-10-23 Issue:3 Volume:20 Page:1-23
ISSN:1551-6857
Container-title:ACM Transactions on Multimedia Computing, Communications, and Applications
language:en
Short-container-title:ACM Trans. Multimedia Comput. Commun. Appl.

Author:

Guo Jintao¹^ORCID,Qi Lei²^ORCID,Shi Yinghuan¹^ORCID,Gao Yang¹^ORCID

Affiliation:

1. The State Key Laboratory for Novel Software Technology, Nanjing University, China

2. The School of Computer Science and Engineering, Southeast University and Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications (Southeast University), Ministry of Education, China

Abstract

Domain generalization (DG) aims to learn a generic model from multiple observed source domains that generalizes well to arbitrary unseen target domains without further training. The major challenge in DG is that the model inevitably faces a severe overfitting issue due to the domain gap between source and target domains. To mitigate this problem, some dropout-based methods have been proposed to resist overfitting by discarding part of the representation of the intermediate layers. However, we observe that most of these methods only conduct the dropout operation in some specific layers, leading to an insufficient regularization effect on the model. We argue that applying dropout at multiple layers can produce stronger regularization effects, which could alleviate the overfitting problem on source domains more adequately than previous layer-specific dropout methods. In this article, we develop a novel layer-wise and channel-wise dropout for DG, which randomly selects one layer and then randomly selects its channels to conduct dropout. Particularly, the proposed method can generate a variety of data variants to better deal with the overfitting issue. We also provide theoretical analysis for our dropout method and prove that it can effectively reduce the generalization error bound. Besides, we leverage the progressive scheme to increase the dropout ratio with the training progress, which can gradually boost the difficulty of training the model to enhance its robustness. Extensive experiments on three standard benchmark datasets have demonstrated that our method outperforms several state-of-the-art DG methods. Our code is available at https://github.com/lingeringlight/PLACEdropout .

Funder

NSFC

Jiangsu Natural Science Foundation Project

Publisher

Association for Computing Machinery (ACM)

Subject

Computer Networks and Communications,Hardware and Architecture

Link

https://dl.acm.org/doi/pdf/10.1145/3624015

Reference68 articles.

1. Yogesh Balaji, Swami Sankaranarayanan, and Rama Chellappa. 2018. Metareg: Towards domain generalization using meta-regularization. In Proceedings of the NeurIPS.

2. Curriculum learning

3. Francesco Cappio Borlino, Antonio D’Innocente, and Tatiana Tommasi. 2021. Rethinking domain generalization baselines. In Proceedings of the ICPR.

4. Fabio M Carlucci, Antonio D’Innocente, Silvia Bucci, Barbara Caputo, and Tatiana Tommasi. 2019. Domain generalization by solving jigsaw puzzles. In Proceedings of the CVPR.

5. Chaoqi Chen, Jiongcheng Li, Xiaoguang Han, Xiaoqing Liu, and Yizhou Yu. 2022. Compound domain generalization via meta-knowledge encoding. In Proceedings of the CVPR.

Cited by 5 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Mask-Shift-Inference: A novel paradigm for domain generalization;Neural Networks;2024-11

2. A Random Focusing Method with Jensen–Shannon Divergence for Improving Deep Neural Network Performance Ensuring Architecture Consistency;Neural Processing Letters;2024-06-17

3. Reducing Overfitting Risk in Small-Sample Learning with ANN: A Case of Predicting Graduate Admission Probability;Communications in Computer and Information Science;2024

4. Gradient-Guided Channel Masking for Cross-Domain Few-Shot Learning;2024

5. Meta Channel Masking for Cross-Domain Few-Shot Learning;2024