Concept Drift Adaptation by Exploiting Drift Type

Author:

Li Jinpeng1,Yu Hang1,zhenyuzhang 1,Luo Xiangfeng1,Xie Shaorong1

Affiliation:

1. School of Computer Engineering and Science, China

Abstract

Concept drift is a phenomenon where the distribution of data streams changes over time. When this happens, model predictions become less accurate. Hence, models built in the past need to be re-learned for the current data. Two design questions need to be addressed in designing a strategy to re-learn models: which type of concept drift has occurred, and how to utilize the drift type to improve re-learning performance. Existing drift detection methods are often good at determining when drift has occurred. However, few retrieve information about how the drift came to be present in the stream. Hence, determining the impact of the type of drift on adaptation is difficult. Filling this gap, we designed a framework based on a lazy strategy called Type-Driven Lazy Drift Adaptor (Type-LDA). Type-LDA first retrieves information about both how and when a drift has occurred, then it uses this information to re-learn the new model. To identify the type of drift, a drift type identifier is pre-trained on synthetic data of known drift types. Further, a drift point locator locates the optimal point of drift via a sharing loss. Hence, Type-LDA can select the optimal point, according to the drift type, to re-learn the new model. Experiments validate Type-LDA on both synthetic data and real-world data, and the results show that accurately identifying drift type can improve adaptation accuracy.

Publisher

Association for Computing Machinery (ACM)

Subject

General Computer Science

Reference53 articles.

1. ElStream: An Ensemble Learning Approach for Concept Drift Detection in Dynamic Social Big Data Stream Learning

2. Supriya Agrahari and Anil Kumar Singh . 2021. Concept Drift Detection in Data Stream Mining: A literature review . Journal of King Saud University-Computer and Information Sciences ( 2021 ). Supriya Agrahari and Anil Kumar Singh. 2021. Concept Drift Detection in Data Stream Mining: A literature review. Journal of King Saud University-Computer and Information Sciences (2021).

3. Rakesh Agrawal , Tomasz Imielinski , and Arun Swami . 1993. Database mining: A performance perspective . IEEE transactions on knowledge and data engineering 5, 6( 1993 ), 914–925. Rakesh Agrawal, Tomasz Imielinski, and Arun Swami. 1993. Database mining: A performance perspective. IEEE transactions on knowledge and data engineering 5, 6(1993), 914–925.

4. M. Baena-Garc , JD Campo-Ávila , R. Fidalgo , A. Bifet , and R. Morales-Bueno . 2006 . Early Drift Detection Method. International Workshop on Knowledge Discovery from Data Streams ( 2006 ). M. Baena-Garc, JD Campo-Ávila, R. Fidalgo, A. Bifet, and R. Morales-Bueno. 2006. Early Drift Detection Method. International Workshop on Knowledge Discovery from Data Streams (2006).

5. Manuel Baena-Garcıa , José del Campo-Ávila , Raúl Fidalgo , Albert Bifet , R Gavalda , and Rafael Morales-Bueno . 2006 . Early drift detection method . In Fourth international workshop on knowledge discovery from data streams, Vol.  6. 77–86 . Manuel Baena-Garcıa, José del Campo-Ávila, Raúl Fidalgo, Albert Bifet, R Gavalda, and Rafael Morales-Bueno. 2006. Early drift detection method. In Fourth international workshop on knowledge discovery from data streams, Vol.  6. 77–86.

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3