Supposed Maximum Mutual Information for Improving Generalization and Interpretation of Multi-Layered Neural Networks

15Citations
Citations of this article
8Readers
Mendeley users who have this article in their library.

Abstract

The present paper1 aims to propose a new type of information-theoretic method to maximize mutual information between inputs and outputs. The importance of mutual information in neural networks is well known, but the actual implementation of mutual information maximization has been quite difficult to undertake. In addition, mutual information has not extensively been used in neural networks, meaning that its applicability is very limited. To overcome the shortcoming of mutual information maximization, we present it here in a very simplified manner by supposing that mutual information is already maximized before learning, or at least at the beginning of learning. The method was applied to three data sets (crab data set, wholesale data set, and human resources data set) and examined in terms of generalization performance and connection weights. The results showed that by disentangling connection weights, maximizing mutual information made it possible to explicitly interpret the relations between inputs and outputs.

Cite

CITATION STYLE

APA

Kamimura, R. (2019). Supposed Maximum Mutual Information for Improving Generalization and Interpretation of Multi-Layered Neural Networks. Journal of Artificial Intelligence and Soft Computing Research, 9(2), 123–147. https://doi.org/10.2478/jaiscr-2018-0029

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free