ONLINE NEWS CLASSIFICATION USING MACHINE LEARNING TECHNIQUES-Reference-Cited by-同舟云学术

ONLINE NEWS CLASSIFICATION USING MACHINE LEARNING TECHNIQUES

Published:2021-07-04 Issue:2 Volume:22 Page:210-225
ISSN:2289-7860
Container-title:IIUM Engineering Journal
language:
Short-container-title:IIUMEJ

Author:

Ahmed Jeelani,Ahmed Muqeem

Abstract

A massive rise in web-based online content today pushes businesses to implement new approaches and resources that might support better navigation, processing, and handling of high-dimensional data. Over the Internet, 90% of the data is unstructured, and there are several approaches through which this data can translate into useful, structured data—classification is one such approach. Classification of knowledge into a good collection of groups is significant and necessary. As the number of machine-readable documents proliferates, automatic text classification is badly needed to classify these documents. Unlabeled documents are categorized into predefined classes of labeled documents using text labeling, a supervised learning technique. This paper reviewed some existing approaches for classifying online news articles and discusses a framework for the automatic classification of online news articles. For achieving high accuracy, different classifiers were tried. Our experimental method achieved 93% accuracy using a Bayesian classifier and present in terms of confusion metrics. ABSTRAK: Peningkatan tinggi pada masa kini pada maklumat dalam talian berasaskan web menyebabkan kaedah baru dalam bisnes telah diguna pakai dan sumber sokongan seperti navigasi, proses, dan pengurusan data berdimensi-tinggi adalah perlu. 90% data di internet adalah data tidak berstruktur, dan terdapat pelbagai kaedah data ini dapat diterjemahkan kepada data berguna, lebih berstruktur — iaitu melalui kaedah klasifikasi. Klasifikasi ilmu kepada koleksi kumpulan baik adalah penting dan perlu. Seperti mana mesin-boleh baca dokumen berkembang pesat, teks klasifikasi automatik juga sangat diperlukan bagi mengklasifikasi dokumen-dokumen ini. Dokumen yang tidak dilabel dikategori sebagai pengelasan pratakrif dokumen berlabel melalui teks label, iaitu teknik pembelajaran berpenyelia. Kajian ini mengkaji semula pendekatan sedia ada bagi artikel berita dalam talian dan membincangkan rangka kerja bagi pengelasan automatik artikel berita dalam talian. Bagi menghasilkan ketepatan yang tinggi, kami menggunakan pelbagai alat klasifikasi. Kaedah eksperimen ini mempunyai ketepatan 93% menggunakan pengelas Bayesian dan data dibentangkan berdasarkan matriks kekeliruan.

Publisher

IIUM Press

Subject

Applied Mathematics,General Engineering,General Chemical Engineering,General Computer Science

Reference54 articles.

1. Jindal R, Malhotra R, Jain A. (2015) Techniques for text classification: Literature review

2. and current trends. Webology, 12(2): Article 139.

3. https://www.webology.org/2015/v12n2/a139.pdf

4. Turney P. (2002) Thumbs Up or Thumbs Down? Semantic Orientation Applied to Unsupervised Classification of Reviews. Computing Research Repository, 417-424. doi:10.3115/1073083.1073153.

5. Wilson T, Wiebe J, Hoffmann P. (2009) Recognizing Contextual Polarity: An Exploration of Features for Phrase-Level Sentiment Analysis. Computational Linguistics, 35(3): 399-433. doi:10.1162/coli.08-012-r1-06-90

Cited by 6 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. K-means and Support Vector Machines for Thai Language IT News Retrieval System;2023 7th International Conference on Computer, Software and Modeling (ICCSM);2023-07-21

2. Arabic News Classification Based on the Country of Origin Using Machine Learning and Deep Learning Techniques;Applied Sciences;2023-06-13

3. Natural Language Contents Evaluation System for Multi-class News Categorization Using Machine Learning and Transformers;Communications in Computer and Information Science;2023

4. Machine Learning Based Text Classification Technology;2022 IEEE 2nd International Conference on Mobile Networks and Wireless Communications (ICMNWC);2022-12-02

5. Classification of Multi-Labeled Text Articles with Reuters Dataset using SVM;2022 International Conference on Science and Technology (ICOSTECH);2022-02-03