A generic framework for ontology-based information retrieval and image retrieval in web data-Reference-Cited by-同舟云学术

A generic framework for ontology-based information retrieval and image retrieval in web data

Published:2016-11-05 Issue:1 Volume:6 Page:
ISSN:2192-1962
Container-title:Human-centric Computing and Information Sciences
language:en
Short-container-title:Hum. Cent. Comput. Inf. Sci.

Author:

Vijayarajan V.,Dinakaran M.,Tejaswin Priyam,Lohani Mayank

Abstract

AbstractIn the internet era, search engines play a vital role in information retrieval from web pages. Search engines arrange the retrieved results using various ranking algorithms. Additionally, retrieval is based on statistical searching techniques or content-based information extraction methods. It is still difficult for the user to understand the abstract details of every web page unless the user opens it separately to view the web content. This key point provided the motivation to propose and display an ontology-based object-attribute-value (O-A-V) information extraction system as a web model that acts as a user dictionary to refine the search keywords in the query for subsequent attempts. This first model is evaluated using various natural language processing (NLP) queries given as English sentences. Additionally, image search engines, such as Google Images, use content-based image information extraction and retrieval of web pages against the user query. To minimize the semantic gap between the image retrieval results and the expected user results, the domain ontology is built using image descriptions. The second proposed model initially examines natural language user queries using an NLP parser algorithm that will identify the subject-predicate-object (S-P-O) for the query. S-P-O extraction is an extended idea from the ontology-based O-A-V web model. Using this S-P-O extraction and considering the complex nature of writing SPARQL protocol and RDF query language (SPARQL) from the user point of view, the SPARQL auto query generation module is proposed, and it will auto generate the SPARQL query. Then, the query is deployed on the ontology, and images are retrieved based on the auto-generated SPARQL query. With the proposed methodology above, this paper seeks answers to following two questions. First, how to combine the use of domain ontology and semantics to improve information retrieval and user experience? Second, does this new unified framework improve the standard information retrieval systems? To answer these questions, a document retrieval system and an image retrieval system were built to test our proposed framework. The web document retrieval was tested against three key-words/bag-of-words models and a semantic ontology model. Image retrieval was tested on IAPR TC-12 benchmark dataset. The precision, recall and accuracy results were then compared against standard information retrieval systems using TREC_EVAL. The results indicated improvements over the standard systems. A controlled experiment was performed by test subjects querying the retrieval system in the absence and presence of our proposed framework. The queries were measured using two metrics, time and click-count. Comparisons were made on the retrieval performed with and without our proposed framework. The results were encouraging.

Publisher

Springer Science and Business Media LLC

Subject

General Computer Science

Link

https://link.springer.com/content/pdf/10.1186/s13673-016-0074-1.pdf

Reference39 articles.

1. Berners-Lee T, Hendler J, Lassila O et al (2001) The Semantic Web. Sci Am 284(5):28–37

2. Bizer C, Heath T, Berners-Lee T (2009) Linked data-the story so far. Semantic services, interoperability and web applications: emerging concepts, p 205–227

3. Meehan A, Brennan R, O’Sullivan D (2015) Sparql based mapping management. In: IEEE International Conference on Semantic Computing (ICSC), 2015. IEEE, New York, p 456–459

4. Heath T, Bizer C (2011) Linked data: Evolving the web into a global data space. Synth Lect Semant Web Theory Technol 1(1):1–136

5. Kompridis N (2000) So we need something else for reason to mean. Int J Philos Stud 8(3):271–295

Cited by 31 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Automatic Query Generation Based on Adaptive Naked Mole-Rate Algorithm;Multimedia Tools and Applications;2024-06-27

2. Secure CPS Content-Based Image Retrieval Using Tripartite Delayed Homomorphic Secret Sharing & CNN;HUM-CENT COMPUT INFO;2024

3. BERT-Based Natural Language Processing System for Online Social Information Retrieval on Illness Tracking;2023 International Conference on Advances in Computation, Communication and Information Technology (ICAICCIT);2023-11-23

4. Semantic Context and Attention-driven Framework for Predicting Visual Description Utilizing a Deep Neural Network and Natural Language Processing;International Journal of Case Studies in Business, IT, and Education;2023-07-28

5. Image Retrieval Through Free-Form Query using Intelligent Text Processing;International Journal of Innovative Technology and Exploring Engineering;2023-06-30