Geometric Regularization of Local Activations for Knowledge Transfer in Convolutional Neural Networks-Reference-Cited by-同舟云学术

Geometric Regularization of Local Activations for Knowledge Transfer in Convolutional Neural Networks

Published:2021-08-19 Issue:8 Volume:12 Page:333
ISSN:2078-2489
Container-title:Information
language:en
Short-container-title:Information

Author:

Theodorakopoulos Ilias,Fotopoulou Foteini,Economou George

Abstract

In this work, we propose a mechanism for knowledge transfer between Convolutional Neural Networks via the geometric regularization of local features produced by the activations of convolutional layers. We formulate appropriate loss functions, driving a “student” model to adapt such that its local features exhibit similar geometrical characteristics to those of an “instructor” model, at corresponding layers. The investigated functions, inspired by manifold-to-manifold distance measures, are designed to compare the neighboring information inside the feature space of the involved activations without any restrictions in the features’ dimensionality, thus enabling knowledge transfer between different architectures. Experimental evidence demonstrates that the proposed technique is effective in different settings, including knowledge-transfer to smaller models, transfer between different deep architectures and harnessing knowledge from external data, producing models with increased accuracy compared to a typical training. Furthermore, results indicate that the presented method can work synergistically with methods such as knowledge distillation, further increasing the accuracy of the trained models. Finally, experiments on training with limited data show that a combined regularization scheme can achieve the same generalization as a non-regularized training with 50% of the data in the CIFAR-10 classification task.

Publisher

MDPI AG

Subject

Information Systems

Link

https://www.mdpi.com/2078-2489/12/8/333/pdf

Reference52 articles.

1. Deep Learning in Computer Vision: Principles and Applications;Hassaballah,2020

2. Manifold-based synthetic oversampling with manifold conformance estimation

3. Distilling the Knowledge in a Neural Network;Hinton;arXiv,2015

4. A Survey of Biometric Recognition Using Deep Learning

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Parameter-free surrounding neighborhood based regression methods;Expert Systems with Applications;2022-08