The Conditional Entropy Bottleneck-Reference-Cited by-同舟云学术

The Conditional Entropy Bottleneck

Published:2020-09-08 Issue:9 Volume:22 Page:999
ISSN:1099-4300
Container-title:Entropy
language:en
Short-container-title:Entropy

Author:

Fischer Ian^ORCID

Abstract

Much of the field of Machine Learning exhibits a prominent set of failure modes, including vulnerability to adversarial examples, poor out-of-distribution (OoD) detection, miscalibration, and willingness to memorize random labelings of datasets. We characterize these as failures of robust generalization, which extends the traditional measure of generalization as accuracy or related metrics on a held-out set. We hypothesize that these failures to robustly generalize are due to the learning systems retaining too much information about the training data. To test this hypothesis, we propose the Minimum Necessary Information (MNI) criterion for evaluating the quality of a model. In order to train models that perform well with respect to the MNI criterion, we present a new objective function, the Conditional Entropy Bottleneck (CEB), which is closely related to the Information Bottleneck (IB). We experimentally test our hypothesis by comparing the performance of CEB models with deterministic models and Variational Information Bottleneck (VIB) models on a variety of different datasets and robustness challenges. We find strong empirical evidence supporting our hypothesis that MNI models improve on these problems of robust generalization.

Publisher

MDPI AG

Subject

General Physics and Astronomy

Link

https://www.mdpi.com/1099-4300/22/9/999/pdf

Reference55 articles.

1. Simple and Scalable Predictive Uncertainty Estimation using Deep Ensembles;Lakshminarayanan,2017

Cited by 37 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Adversarial Machine Learning in Industry: A Systematic Literature Review;Computers & Security;2024-10

2. Interpretability of deep neural networks: A review of methods, classification and hardware;Neurocomputing;2024-10

3. Multimodal information bottleneck for deep reinforcement learning with multiple sensors;Neural Networks;2024-08

4. Contrast Information Dynamics: A Novel Information Measure for Cognitive Modelling;Entropy;2024-07-27

5. Partial Information Decomposition: Redundancy as Information Bottleneck;Entropy;2024-06-26