Isolation-based feature Selection for Unsupervised Outlier Detection-Reference-Cited by-同舟云学术

Isolation-based feature Selection for Unsupervised Outlier Detection

Published:2019-09-22 Issue:1 Volume:11 Page:
ISSN:2325-0178
Container-title:Annual Conference of the PHM Society
language:
Short-container-title:PHM_CONF

Author:

Yang Qibo,Singh Jaskaran,Lee Jay

Abstract

For high-dimensional datasets, bad features and complex interactions between features can cause high computational costs and make outlier detection algorithms inefficient. Most feature selection methods are designed for supervised classification and regression, and limited works are specifically for unsupervised outlier detection. This paper proposes a novel isolation-based feature selection (IBFS) method for unsupervised outlier detection. It is based on the training process of isolation forest. When a point of a feature is used to split the data, the imbalanced distribution of split data is measured and used to quantify how strong this feature can detect outliers. We also compare the proposed method with variance, Laplacian score and kurtosis. These methods are benchmarked on simulated data to show their characteristics. Then we evaluate the performance using one-class support vector machine, isolation forest and local outlier factor on several real-word datasets. The evaluation results show that the proposed method can improve the performance of isolation forest, and its results are similar to and sometimes better than another useful outlier indicator: kurtosis, which demonstrate the effectiveness of the proposed method. We also notice that sometimes variance and Laplacian score has similar performance on the datasets.

Publisher

PHM Society

Cited by 10 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. AcME-AD: Accelerated Model Explanations for Anomaly Detection;Communications in Computer and Information Science;2024

2. A Survey on Explainable Anomaly Detection;ACM Transactions on Knowledge Discovery from Data;2023-09-06

3. Feature selection considering interaction, redundancy and complementarity for outlier detection in categorical data;Knowledge-Based Systems;2023-09

4. An Ensemble Learning Based Intrusion Detection Model for Industrial IoT Security;Big Data Mining and Analytics;2023-09

5. An effective intrusion detection approach based on ensemble learning for IIoT edge computing;Journal of Computer Virology and Hacking Techniques;2022-12-14