Incremental Ant-Miner Classifier for Online Big Data Analytics

被引:0
|
作者
Al-Dawsari, Amal [1 ]
Al-Turaiki, Isra [2 ]
Kurdi, Heba [1 ,3 ]
机构
[1] King Saud Univ, Coll Comp & Informat Sci, Comp Sci Dept, Riyadh 11451, Saudi Arabia
[2] King Saud Univ, Coll Comp & Informat Sci, Informat Technol Dept, Riyadh 11451, Saudi Arabia
[3] MIT, Mech Engn Dept, Cambridge, MA 02142 USA
关键词
machine learning; association rule mining; ant colony optimization; incremental classifier; big data analytics; IoT; COLONY; ALGORITHM; FLOOD;
D O I
10.3390/s22062223
中图分类号
O65 [分析化学];
学科分类号
070302 ; 081704 ;
摘要
Internet of Things (IoT) environments produce large amounts of data that are challenging to analyze. The most challenging aspect is reducing the quantity of consumed resources and time required to retrain a machine learning model as new data records arrive. Therefore, for big data analytics in IoT environments where datasets are highly dynamic, evolving over time, it is highly advised to adopt an online (also called incremental) machine learning model that can analyze incoming data instantaneously, rather than an offline model (also called static), that should be retrained on the entire dataset as new records arrive. The main contribution of this paper is to introduce the Incremental Ant-Miner (IAM), a machine learning algorithm for online prediction based on one of the most well-established machine learning algorithms, Ant-Miner. IAM classifier tackles the challenge of reducing the time and space overheads associated with the classic offline classifiers, when used for online prediction. IAM can be exploited in managing dynamic environments to ensure timely and space-efficient prediction, achieving high accuracy, precision, recall, and F-measure scores. To show its effectiveness, the proposed IAM was run on six different datasets from different domains, namely horse colic, credit cards, flags, ionosphere, and two breast cancer datasets. The performance of the proposed model was compared to ten state-of-the-art classifiers: naive Bayes, logistic regression, multilayer perceptron, support vector machine, K*, adaptive boosting (AdaBoost), bagging, Projective Adaptive Resonance Theory (PART), decision tree (C4.5), and random forest. The experimental results illustrate the superiority of IAM as it outperformed all the benchmarks in nearly all performance measures. Additionally, IAM only needs to be rerun on the new data increment rather than the entire big dataset on the arrival of new data records, which makes IAM better in time- and resource-saving. These results demonstrate the strong potential and efficiency of the IAM classifier for big data analytics in various areas.
引用
收藏
页数:17
相关论文
共 50 条
  • [21] Big data analytics for critical information classification in online social networks using classifier chains
    Douglas H. Silva
    Erick G. Maziero
    Muhammad Saadi
    Renata L. Rosa
    Juan C. Silva
    Demostenes Z. Rodriguez
    Kostromitin K. Igorevich
    Peer-to-Peer Networking and Applications, 2022, 15 : 626 - 641
  • [22] AMclr -An Improved Ant-Miner to Extract Comprehensible and Diverse Classification Rules
    Ayub, Umair
    Ikram, Asim
    Shahzad, Waseem
    PROCEEDINGS OF THE 2019 GENETIC AND EVOLUTIONARY COMPUTATION CONFERENCE (GECCO'19), 2019, : 4 - 12
  • [23] Classification of Cervical Cancer Using Ant-Miner for Medical Expertise Knowledge Management
    Wahid, Juliana
    Al-Mazini, Hassan Fouad Abbas
    PROCEEDINGS OF KNOWLEDGE MANAGEMENT INTERNATIONAL CONFERENCE (KMICE) 2018, 2018, : 393 - 397
  • [24] A case study for an incremental classifier model in big data
    Lincy S.B.T.
    Nagarajan S.K.
    International Journal of Cloud Computing, 2019, 8 (03) : 266 - 282
  • [25] Deep Incremental Learning for Big Data Stream Analytics
    Alex, Suja A.
    Nayahi, J. Jesu Vedha
    PROCEEDING OF THE INTERNATIONAL CONFERENCE ON COMPUTER NETWORKS, BIG DATA AND IOT (ICCBI-2018), 2020, 31 : 600 - 614
  • [26] 一种具有免疫特征的Ant-Miner算法
    张惠萍
    李桂成
    电脑开发与应用, 2007, (12) : 47 - 49
  • [27] A Mixed-Attribute Approach in Ant-Miner Classification Rule Discovery Algorithm
    Helal, Ayah
    Otero, Fernando E. B.
    GECCO'16: PROCEEDINGS OF THE 2016 GENETIC AND EVOLUTIONARY COMPUTATION CONFERENCE, 2016, : 13 - 20
  • [28] Rule Pruning Techniques in the Ant-Miner Classification Algorithm and Its Variants: A Review
    Al-Behadili, Hayder Naser Khraibet
    Ku-Mahamud, Ku Ruhana
    Sagban, Rafid
    2018 IEEE SYMPOSIUM ON COMPUTER APPLICATIONS & INDUSTRIAL ELECTRONICS (ISCAIE 2018), 2018, : 78 - 84
  • [29] Online big data chemical batch analytics
    Hollender, Martin
    Chioua, Moncef
    Xu, Chaojun
    CHIMICA OGGI-CHEMISTRY TODAY, 2018, 36 (05) : 33 - 35
  • [30] Big data analytics for intelligent online education
    Zhang, Rongbo
    Zhao, Weiyu
    Wang, Yixin
    JOURNAL OF INTELLIGENT & FUZZY SYSTEMS, 2021, 40 (02) : 2815 - 2825