Affiliation:
1. Hebei University of Water Resources and Electric Engineering, Cangzhou, Hebei, China
2. Cangzhou Municipal Human Resources and Social Security Bureau, Cangzhou, Hebei, China
Abstract
Clustering algorithm is the foundation and important technology in data mining. In fact, in the real world, the data itself often has a hierarchical structure. Hierarchical clustering aims at constructing a cluster tree, which reveals the underlying modal structure of a complex density. Due to its inherent complexity, most existing hierarchical clustering algorithms are usually designed heuristically without an explicit objective function, which limits its utilization and analysis. K-means clustering, the well-known simple yet effective algorithm which can be expressed from the view of probability distribution, has inherent connection to Mixture of Gaussians (MoG). At this point, we consider combining Bayesian theory analysis with K-means algorithm. This motivates us to develop a hierarchical clustering based on K-means under the probability distribution framework, which is different from existing hierarchical K-means algorithms processing data in a single-pass manner along with heuristic strategies. For this goal, we propose an explicit objective function for hierarchical clustering, termed as Bayesian hierarchical K-means (BHK-means). In our method, a cascaded clustering tree is constructed, in which all layers interact with each other in the network-like manner. In this cluster tree, the clustering results of each layer are influenced by the parent and child nodes. Therefore, the clustering result of each layer is dynamically improved in accordance with the global hierarchical clustering objective function. The objective function is solved using the same algorithm as K-means, the Expectation-maximization algorithm. The experimental results on both synthetic data and benchmark datasets demonstrate the effectiveness of our algorithm over the existing related ones.
Subject
Artificial Intelligence,Computer Vision and Pattern Recognition,Theoretical Computer Science
Reference29 articles.
1. L. Fei-Fei and P. Perona, A Bayesian hierarchical model for learning natural scene categories, in: 2005 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR’05), IEEE, 2005.
2. X. Wang, X. Ma and E. Grimson, Unsupervised Activity Perception by Hierarchical Bayesian Models, in: 2007 IEEE Conference on Computer Vision and Pattern Recognition, IEEE Computer Society, 2007.
3. A Bayesian hierarchical model for monitoring harbor seal changes in Prince William Sound, Alaska;Hoef;Environmental and Ecological Statistics,2003
4. A Bayesian sampling approach to decision fusion using hierarchical models;Chen;IEEE Transactions on Signal Processing,2002
5. K.A. Heller and Z. Ghahramani, Bayesian hierarchical clustering, in: International Conference on Machine Learning, 2005.
Cited by
18 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献