Lost in Transduction: Transductive Transfer Learning in Text Classification-Reference-Cited by-同舟云学术

Lost in Transduction: Transductive Transfer Learning in Text Classification

Published:2021-07-03 Issue:1 Volume:16 Page:1-21
ISSN:1556-4681
Container-title:ACM Transactions on Knowledge Discovery from Data
language:en
Short-container-title:ACM Trans. Knowl. Discov. Data

Author:

Moreo Alejandro¹,Esuli Andrea¹,Sebastiani Fabrizio¹

Affiliation:

1. Istituto di Scienza e Tecnologie dell’Informazione, Consiglio Nazionale delle Ricerche, Pisa, Italy

Abstract

Obtaining high-quality labelled data for training a classifier in a new application domain is often costly. Transfer Learning (a.k.a. “Inductive Transfer”) tries to alleviate these costs by transferring, to the “target” domain of interest, knowledge available from a different “source” domain. In transfer learning the lack of labelled information from the target domain is compensated by the availability at training time of a set of unlabelled examples from the target distribution. Transductive Transfer Learning denotes the transfer learning setting in which the only set of target documents that we are interested in classifying is known and available at training time. Although this definition is indeed in line with Vapnik’s original definition of “transduction”, current terminology in the field is confused. In this article, we discuss how the term “transduction” has been misused in the transfer learning literature, and propose a clarification consistent with the original characterization of this term given by Vapnik. We go on to observe that the above terminology misuse has brought about misleading experimental comparisons, with inductive transfer learning methods that have been incorrectly compared with transductive transfer learning methods. We then, give empirical evidence that the difference in performance between the inductive version and the transductive version of a transfer learning method can indeed be statistically significant (i.e., that knowing at training time the only data one needs to classify indeed gives an advantage). Our clarification allows a reassessment of the field, and of the relative merits of the major, state-of-the-art algorithms for transfer learning in text classification.

Funder

ARIADNEplus

European Commission

Publisher

Association for Computing Machinery (ACM)

Subject

General Computer Science

Link

https://dl.acm.org/doi/pdf/10.1145/3453146

Reference61 articles.

1. Amar P. Azad Dinesh Garg Priyanka Agrawal and Arun Kumar. 2018. Deep domain adaptation under deep label scarcity. arXiv:1809.08097. Retrieved from https://arxiv.org/abs/1809.08097. Amar P. Azad Dinesh Garg Priyanka Agrawal and Arun Kumar. 2018. Deep domain adaptation under deep label scarcity. arXiv:1809.08097. Retrieved from https://arxiv.org/abs/1809.08097.

2. Learning with Minimum Supervision: A General Framework for Transductive Transfer Learning

3. A partially supervised cross-collection topic model for cross-domain text classification

4. Long term bank failure prediction using Fuzzy Refinement-based Transductive Transfer learning

Cited by 13 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Predicting semantic category of answers for question answering systems using transformers: a transfer learning approach;Multimedia Tools and Applications;2024-02-24

2. Genetic Programming for Document Classification: A Transductive Transfer Learning System;IEEE Transactions on Cybernetics;2024-02

3. Transfer Learning based Location-Aided Modulation Classification in Indoor Environments for Cognitive Radio Applications;Radioengineering;2023-12

4. Classification of Unstructured Power Grid Data Based on Self-organizing Map Neural Network;2023 3rd International Conference on New Energy and Power Engineering (ICNEPE);2023-11-24

5. Why Did You Go There? Semantic Knowledge Extraction from Trajectory Traces;IGARSS 2023 - 2023 IEEE International Geoscience and Remote Sensing Symposium;2023-07-16