Information Flows of Diverse Autoencoders-Reference-Cited by-同舟云学术

Information Flows of Diverse Autoencoders

Published:2021-07-05 Issue:7 Volume:23 Page:862
ISSN:1099-4300
Container-title:Entropy
language:en
Short-container-title:Entropy

Author:

Lee Sungyeop^ORCID,Jo Junghyo

Abstract

Deep learning methods have had outstanding performances in various fields. A fundamental query is why they are so effective. Information theory provides a potential answer by interpreting the learning process as the information transmission and compression of data. The information flows can be visualized on the information plane of the mutual information among the input, hidden, and output layers. In this study, we examine how the information flows are shaped by the network parameters, such as depth, sparsity, weight constraints, and hidden representations. Here, we adopt autoencoders as models of deep learning, because (i) they have clear guidelines for their information flows, and (ii) they have various species, such as vanilla, sparse, tied, variational, and label autoencoders. We measured their information flows using Rényi’s matrix-based α-order entropy functional. As learning progresses, they show a typical fitting phase where the amounts of input-to-hidden and hidden-to-output mutual information both increase. In the last stage of learning, however, some autoencoders show a simplifying phase, previously called the “compression phase”, where input-to-hidden mutual information diminishes. In particular, the sparsity regularization of hidden activities amplifies the simplifying phase. However, tied, variational, and label autoencoders do not have a simplifying phase. Nevertheless, all autoencoders have similar reconstruction errors for training and test data. Thus, the simplifying phase does not seem to be necessary for the generalization of learning.

Funder

Ministry of Science and ICT, South Korea

New Faculty Startup Fund from Seoul National University,

Publisher

MDPI AG

Subject

General Physics and Astronomy

Link

https://www.mdpi.com/1099-4300/23/7/862/pdf

Reference37 articles.

1. A Mathematical Theory of Communication

2. Elements of Information Theory;Cover,1999

3. Information Theory and Statistical Mechanics

4. Information Theory, Evolution, and the Origin of Life;Yockey,2005

5. Information Theory, Inference and Learning Algorithms;MacKay,2003

Cited by 8 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Comparative Analysis of Dimensionality Reduction Techniques Applied to Disease Classification Tasks;Studies in Computational Intelligence;2024

2. An AI-Guided Data Centric Strategy to Detect and Mitigate Biases in Healthcare Datasets;2023-11-07

3. Autoencoder with Orthogonal Variant;2023 20th International Conference on Electrical Engineering, Computing Science and Automatic Control (CCE);2023-10-25

4. Multivariate Time Series Information Bottleneck;Entropy;2023-05-22

5. On the Size and Width of the Decoder of a Boolean Threshold Autoencoder;IEEE Transactions on Neural Networks and Learning Systems;2023