Affiliation:
1. Laboratory for Big Data and Decision, National University of Defense Technology, Changsha 410073, China
Abstract
(1) Background: Chinese news text is a popular form of media communication, which can be seen everywhere in China. Chinese news text classification is an important direction in natural language processing (NLP). How to use high-quality text classification technology to help humans to efficiently organize and manage the massive amount of web news is an urgent problem to be solved. It is noted that the existing deep learning methods rely on a large-scale tagged corpus for news text classification tasks and this model is poorly interpretable because the size is large. (2) Methods: To solve the above problems, this paper proposes a Chinese news text classification method based on key feature enhancement named KFE-CNN. It can effectively expand the semantic information of key features to enhance sample data and then combine the zero–one binary vector representation to transform text features into binary vectors and input them into CNN model for training and implementation, thus improving the interpretability of the model and effectively compressing the size of the model. (3) Results: The experimental results show that our method can significantly improve the overall performance of the model and the average accuracy and F1-score of the THUCNews subset of the public dataset reached 97.84% and 98%. (4) Conclusions: this fully proved the effectiveness of the KFE-CNN method for the Chinese news text classification task and it also fully demonstrates that key feature enhancement can improve classification performance.
Funder
National Natural Science Foundation of China
Subject
Fluid Flow and Transfer Processes,Computer Science Applications,Process Chemistry and Technology,General Engineering,Instrumentation,General Materials Science
Reference26 articles.
1. Survey on supervised machine learning techniques for automatic text classification;Kadhim;Artif. Intell. Rev.,2019
2. Deep learning-based text classification: A comprehensive review;Minaee;ACM Comput. Surv. (CSUR),2021
3. A survey of the usages of deep learning for natural language processing;Otter;IEEE Trans. Neural Netw. Learn. Syst.,2020
4. Feature selection for text classification: A review;Deng;Multimed. Tools Appl.,2019
5. VM, N., and Kumar, R.D. (2019, January 17–18). Implementation on Text Classification Using Bag of Words Model. Proceedings of the Second International Conference on Emerging Trends in Science & Technologies for Engineering Systems, Sarigam, India.
Cited by
2 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献