MEAL: Multi-Model Ensemble via Adversarial Learning-Reference-Cited by-同舟云学术

MEAL: Multi-Model Ensemble via Adversarial Learning

Published:2019-07-17 Issue: Volume:33 Page:4886-4893
ISSN:2374-3468
Container-title:Proceedings of the AAAI Conference on Artificial Intelligence
language:
Short-container-title:AAAI

Author:

Shen Zhiqiang,He Zhankui,Xue Xiangyang

Abstract

Often the best performing deep neural models are ensembles of multiple base-level networks. Unfortunately, the space required to store these many networks, and the time required to execute them at test-time, prohibits their use in applications where test sets are large (e.g., ImageNet). In this paper, we present a method for compressing large, complex trained ensembles into a single network, where knowledge from a variety of trained deep neural networks (DNNs) is distilled and transferred to a single DNN. In order to distill diverse knowledge from different trained (teacher) models, we propose to use adversarial-based learning strategy where we define a block-wise training loss to guide and optimize the predefined student network to recover the knowledge in teacher models, and to promote the discriminator network to distinguish teacher vs. student features simultaneously. The proposed ensemble method (MEAL) of transferring distilled knowledge with adversarial learning exhibits three important advantages: (1) the student network that learns the distilled knowledge with discriminators is optimized better than the original model; (2) fast inference is realized by a single forward pass, while the performance is even better than traditional ensembles from multi-original models; (3) the student network can learn the distilled knowledge from a teacher model that has arbitrary structures. Extensive experiments on CIFAR-10/100, SVHN and ImageNet datasets demonstrate the effectiveness of our MEAL method. On ImageNet, our ResNet-50 based MEAL achieves top-1/5 21.79%/5.99% val error, which outperforms the original model by 2.06%/1.14%.

Publisher

Association for the Advancement of Artificial Intelligence (AAAI)

Subject

General Medicine

Cited by 90 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Multi-label category enhancement fusion distillation based on variational estimation;Knowledge-Based Systems;2024-09

2. PURF: Improving teacher representations by imposing smoothness constraints for knowledge distillation;Applied Soft Computing;2024-07

3. MulKD: Multi-layer Knowledge Distillation via collaborative learning;Engineering Applications of Artificial Intelligence;2024-07

4. Efficient Multi-Model Fusion with Adversarial Complementary Representation Learning;2024 International Joint Conference on Neural Networks (IJCNN);2024-06-30

5. Multi-Stage Adversarial Learning for Cloth-Changing Person Re-Identification;2024 5th International Conference on Computer Engineering and Application (ICCEA);2024-04-12