Data Mining of Inputs: Analysing Magnitude and Functional Measures-Reference-Cited by-同舟云学术

Data Mining of Inputs: Analysing Magnitude and Functional Measures

Published:1997-04 Issue:02 Volume:08 Page:209-218
ISSN:0129-0657
Container-title:International Journal of Neural Systems
language:en
Short-container-title:Int. J. Neur. Syst.

Author:

Gedeon Tamás D.¹

Affiliation:

1. Department of Information Engineering, School of Computer Science and Engineering, The University of New South Wales, Sydney 2052, Australia

Abstract

The problem of data encoding and feature selection for training back-propagation neural networks is well known. The basic principles are to avoid encrypting the underlying structure of the data, and to avoid using irrelevant inputs. This is not easy in the real world, where we often receive data which has been processed by at least one previous user. The data may contain too many instances of some class, and too few instances of other classes. Real data sets often include many irrelevant or redundant input fields. This paper examines the use of weight matrix analysis techniques and functional measures using two real (and hence noisy) data sets. The first part of this paper examines the use of the weight matrix of the trained neural network itself to determine which inputs are significant. A new technique is introduced and compared with two other techniques from the literature. We present our experience and results on some satellite data augmented by a terrain model. The task was to predict the forest supra-type based on the available information. A brute force technique eliminating randomly selected inputs was used to validate our approach. The second part of this paper examines the use of measures to determine the functional contribution of inputs to outputs. Inputs which include minor but unique information to the network are more significant than inputs with higher magnitude contribution but providing redundant information, which is also provided by another input. A comparison is made to sensitivity analysis, where the sensitivity of outputs to input perturbation is used as a measure of the significance of inputs. This paper presents a novel functional analysis of the weight matrix based on a technique developed for determining the behavioral significance of hidden neurons. This is compared with the application of the same technique to the training and test data. Finally, a novel aggregation technique is introduced.

Publisher

World Scientific Pub Co Pte Lt

Subject

Computer Networks and Communications,General Medicine

Link

https://www.worldscientific.com/doi/pdf/10.1142/S0129065797000227

Reference3 articles.

Cited by 127 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Leveraging network topology for credit risk assessment in P2P lending: A comparative study under the lens of machine learning;Expert Systems with Applications;2024-10

2. Geospatial Data and Deep Learning Expose ESG Risks to Critical Raw Materials Supply: The Case of Lithium;Earth Science, Systems and Society;2024-07-04

3. Using comparative extinction risk analysis to prioritize the IUCN Red List reassessments of amphibians;Conservation Biology;2024-07

4. Deep neural network (DNN) modelling for prediction of the mode of delivery;European Journal of Obstetrics & Gynecology and Reproductive Biology;2024-06

5. Risk transmission, systemic fragility of banks’ interacting customers and credit worthiness assessment;Finance Research Letters;2024-04