Abstract
Purpose
Data mining has been a popular research area in the past decades. Many researchers study data-mining theories, methods, applications and trends; however, there are very few studies on data-mining-related topics in social media. This paper aims to explore the topics related to data mining based on the data collected from Wikipedia.
Design/methodology/approach
In total, 402 data-mining-related articles were obtained from Wikipedia. These articles were manually classified into several categories by the coding method. Each category formed an article-term matrix. These matrices were analysed and visualized by the self-organizing map approach. Several clusters were observed in each category. Finally, the topics of these clusters were extracted by content analysis.
Findings
The articles obtained were classified into six categories: applications, foundation and concepts, methodologies, organizations, related fields and topics and technology support. Business, biology and security were the three prominent topics of the applications category. The technologies supporting data mining were software, systems, databases, programming languages and so forth. The general public was more interested in data-mining organizations than the researchers. They also focused on the applications of data mining in business more than in other fields.
Originality/value
This study will help researchers gain insight into the general public’s perceptions of data mining and discover the gap between the general public and themselves. It will assist researchers in finding new techniques and methods which will potentially provide them with new data-mining methods and research topics.
Subject
Library and Information Sciences,Computer Science Applications
Reference61 articles.
1. Social media road maps exploring the futures triggered by social media;VTT Tiedotteita-Valtion Teknillinen Tutkimuskeskus,2008
2. Application of data mining: diabetes health care in young and old patients;Journal of King Saud University-Computer and Information Sciences,2013
3. The visual subject analysis of library and information science journals with self-organizing map;Knowledge Organization,2011
4. Motivating and discouraging factors for Wikipedians: the case study of Persian Wikipedia;Library Review,2013
5. Commons-based peer production and virtue;Journal of Political Philosophy,2006
Cited by
3 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献