Cluster Analysis in Practice: Dealing with Outliers in Managerial Research-Reference-Cited by-同舟云学术

Cluster Analysis in Practice: Dealing with Outliers in Managerial Research

Published:2021 Issue:1 Volume:25 Page:
ISSN:1982-7849
Container-title:Revista de Administração Contemporânea
language:
Short-container-title:Rev. adm. contemp.

Author:

Lopes Humberto Elias Garcia¹^ORCID,Gosling Marlusa de Sevilha²^ORCID

Affiliation:

1. Pontifícia Universidade Católica de Minas Gerais, Brazil

2. Universidade Federal de Minas Gerais, Brazil

Abstract

ABSTRACT Context: in recent years, cluster analysis has stimulated researchers to explore new ways to understand data behavior. The computational ease of this method and its ability to generate consistent outputs, even in small datasets, explain that to some extent. However, researchers are often mistaken in holding that clustering is a terrain in which anything goes. The literature shows the opposite: they must be careful, especially regarding the effect of outliers on cluster formation. Objective: in this tutorial paper, we contribute to this discussion by presenting four clustering techniques and their respective advantages and disadvantages in the treatment of outliers. Methods: for that, we worked from a managerial dataset and analyzed it using k-means, PAM, DBSCAN, and FCM techniques. Results: our analyzes indicate that researchers have distinct clustering techniques for dealing with outliers accordingly. Conclusion: we concluded that researchers need to have a more diversified repertoire of clustering techniques. After all, this would give them two relevant empirical alternatives: choose the most appropriate technique for their research objectives or adopt a multi-method approach.

Publisher

FapUNIFESP (SciELO)

Link

http://www.scielo.br/pdf/rac/v25n1/1982-7849-rac-25-01-e200081.pdf

Reference44 articles.

1. A gentle introduction to Stata;Acock A. C.,2014

2. Identifying and treating outliers in finance;Adams J.;Financial Management,2019

3. Data clustering: Algorithms and applications;Aggarwal C.,2014

4. Economics of strategy;Besanko D.,2016

5. Introduction to deep learning using R: A step-by-step guide to learning and implementing deep learning models using R;Beysolow T.,2017

Cited by 8 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. In situ single-droplet analysis of emulsified fat using confocal Raman microscopy: insights into crystal network formation within spatial resolution;Soft Matter;2024

2. Mapping Homogeneous Response Areas for Forest Fuel Management Using Geospatial Data, K-Means, and Random Forest Classification;Forests;2022-11-22

3. Data-driven versus a domain-led approach to k-means clustering on an open heart failure dataset;International Journal of Data Science and Analytics;2022-07-25

4. A Case-control Study to Determine Metalloproteinase-12 and Lysyl Oxidase Levels in Iraqi women with Osteoporosis;Research Journal of Pharmacy and Technology;2022-06-28

5. Benefit Segmentation of Tourists to Geosites and Its Implications for Sustainable Development of Geotourism in the Southern Lake Tana Region, Ethiopia;Sustainability;2022-03-14