A Flexible Supervised Term-Weighting Technique and its Application to Variable Extraction and Information Retrieval-Reference-Cited by-同舟云学术

A Flexible Supervised Term-Weighting Technique and its Application to Variable Extraction and Information Retrieval

Published:2019-02-27 Issue:63 Volume:22 Page:61-80
ISSN:1988-3064
Container-title:Inteligencia Artificial
language:
Short-container-title:ia

Author:

Maisonnave Mariano,Delbianco Fernando,Tohmé Fernando Abel,Maguitman Ana Gabriela

Abstract

Successful modeling and prediction depend on effective methods for the extraction of domain-relevant variables. This paper proposes a methodology for identifying domain-specific terms. The proposed methodology relies on a collection of documents labeled as relevant or irrelevant to the domain under analysis. Based on the labeled document collection, we propose a supervised technique that weights terms based on their descriptive and discriminating power. Finally, the descriptive and discriminating values are combined into a general measure that, through the use of an adjustable parameter, allows to independently favor different aspects of retrieval such as maximizing precision or recall, or achieving a balance between both of them. The proposed technique is applied to the economic domain and is empirically evaluated through a human-subject experiment involving experts and non-experts in Economy. It is also evaluated as a term-weighting technique for query-term selection showing promising results. We finally illustrate the applicability of the proposed technique to address diverse problems such as building prediction models, supporting knowledge modeling, and achieving total recall.

Publisher

IBERAMIA: Sociedad Iberoamericana de Inteligencia Artificial

Subject

Artificial Intelligence,Software

Cited by 4 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Causal graph extraction from news: a comparative study of time-series causality learning techniques;PeerJ Computer Science;2022-08-03

2. Probabilistic Term Weighting Based on Three-Way Decisions for Class Based Feature Selection;Lecture Notes in Computer Science;2022

3. Assessing the behavior and performance of a supervised term-weighting technique for topic-based retrieval;Information Processing & Management;2021-05

4. An improved term weighting scheme for text classification;Concurrency and Computation: Practice and Experience;2020-05-10