Affiliation:
1. Department of Computer Science and Engineering, Koneru Lakshmaiah Education Foundation, Veddeswaram, AP, India
Abstract
The main intention of this paper is to develop a new intelligent framework for web page classification and re-ranking. The two main phases of the proposed model are (a) classification, and (b) re-ranking-based retrieval. In the classification phase, pre-processing is initially performed, which follows the steps like HTML (Hyper Text Markup Language) tag removal, punctuation marks removal, stop words removal, and stemming. After pre-processing, word to vector formation is done and then, feature extraction is performed by Principle Component Analysis (PCA). From this, optimal feature selection is accomplished, which is the important process for the accurate classification of web pages. Web pages contain several features, which reduces the classification accuracy. Here, the adoption of a new meta-heuristic algorithm termed Opposition based-Tunicate Swarm Algorithm (O-TSA) is employed to perform the optimal feature selection. Finally, the selected features are subjected to the Enhanced Convolutional-Recurrent Neural Network (E-CRNN) for accurate web page classification with enhancement based on O-TSA. The outcome of this phase is the categorization of different web page classes. In the second phase, the re-ranking is involved utilizing the O-TSA, which derives the objective function based on similarity function (correlation) for URL matching, which results in optimal re-ranking of web pages for retrieval. Thus, the proposed method yields better classification and re-ranking performance and reduce space requirements and search time in the web documents compared with the existing methods.
Publisher
World Scientific Pub Co Pte Ltd
Subject
Artificial Intelligence,Information Systems,Control and Systems Engineering,Software
Cited by
2 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献
1. Missing Data Filling of Model Based on Neural Network;International Journal of Uncertainty, Fuzziness and Knowledge-Based Systems;2024-06
2. Error Modeling Analysis of Parallel Robot Based on Improved Fish Swarm Algorithm;2023 2nd International Conference for Innovation in Technology (INOCON);2023-03-03