A Survey of Deep Active Learning-Reference-Cited by-同舟云学术

A Survey of Deep Active Learning

Published:2022-12-31 Issue:9 Volume:54 Page:1-40
ISSN:0360-0300
Container-title:ACM Computing Surveys
language:en
Short-container-title:ACM Comput. Surv.

Author:

Ren Pengzhen¹,Xiao Yun¹,Chang Xiaojun²,Huang Po-Yao³,Li Zhihui⁴,Gupta Brij B.⁵,Chen Xiaojiang¹,Wang Xin¹

Affiliation:

1. Northwest University, Shaanxi Province, China

2. RMIT University

3. Carnegie Mellon University, Pittsburgh

4. Qilu University of Technology (Shandong Academy of Sciences), Shandong, China

5. National Institute of Technology Kurukshetra, Haryana, India

Abstract

Active learning (AL) attempts to maximize a model’s performance gain while annotating the fewest samples possible. Deep learning (DL) is greedy for data and requires a large amount of data supply to optimize a massive number of parameters if the model is to learn how to extract high-quality features. In recent years, due to the rapid development of internet technology, we have entered an era of information abundance characterized by massive amounts of available data. As a result, DL has attracted significant attention from researchers and has been rapidly developed. Compared with DL, however, researchers have a relatively low interest in AL. This is mainly because before the rise of DL, traditional machine learning requires relatively few labeled samples, meaning that early AL is rarely according the value it deserves. Although DL has made breakthroughs in various fields, most of this success is due to a large number of publicly available annotated datasets. However, the acquisition of a large number of high-quality annotated datasets consumes a lot of manpower, making it unfeasible in fields that require high levels of expertise (such as speech recognition, information extraction, medical images, etc.). Therefore, AL is gradually coming to receive the attention it is due. It is therefore natural to investigate whether AL can be used to reduce the cost of sample annotation while retaining the powerful learning capabilities of DL. As a result of such investigations, deep active learning (DeepAL) has emerged. Although research on this topic is quite abundant, there has not yet been a comprehensive survey of DeepAL-related works; accordingly, this article aims to fill this gap. We provide a formal classification method for the existing work, along with a comprehensive and systematic overview. In addition, we also analyze and summarize the development of DeepAL from an application perspective. Finally, we discuss the confusion and problems associated with DeepAL and provide some possible development directions.

Funder

NSFC

Shaanxi Science and Technology Innovation Team Support

Australian Research Council Discovery Early Career Researcher Award

Publisher

Association for Computing Machinery (ACM)

Subject

General Computer Science,Theoretical Computer Science

Link

https://dl.acm.org/doi/pdf/10.1145/3472291

Reference259 articles.

1. Toward the next generation of recommender systems: a survey of the state-of-the-art and possible extensions

2. Massively Multilingual Neural Machine Translation

Cited by 544 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Active learning concerning sampling cost for enhancing AI-enabled building energy system modeling;Advances in Applied Energy;2024-12

2. Improving the Identification of Diabetic Retinopathy and Related Conditions in the Electronic Health Record Using Natural Language Processing Methods;Ophthalmology Science;2024-11

3. Curvature index of image samples used to evaluate the interpretability informativeness;Engineering Applications of Artificial Intelligence;2024-11

4. Human-in-the-loop: Using classifier decision boundary maps to improve pseudo labels;Computers & Graphics;2024-11

5. Mapping high entropy state spaces for novel material discovery;Acta Materialia;2024-10