Analyzing the potential of active learning for document image classification

被引:0
|
作者
Saifullah Saifullah
Stefan Agne
Andreas Dengel
Sheraz Ahmed
机构
[1] German Research Center for Artificial Intelligence,
[2] RPTU Kaiserslautern-Landau,undefined
[3] DeepReader GmbH,undefined
关键词
Document image classification; Document analysis; Active learning; Deep active learning;
D O I
暂无
中图分类号
学科分类号
摘要
Deep learning has been extensively researched in the field of document analysis and has shown excellent performance across a wide range of document-related tasks. As a result, a great deal of emphasis is now being placed on its practical deployment and integration into modern industrial document processing pipelines. It is well known, however, that deep learning models are data-hungry and often require huge volumes of annotated data in order to achieve competitive performances. And since data annotation is a costly and labor-intensive process, it remains one of the major hurdles to their practical deployment. This study investigates the possibility of using active learning to reduce the costs of data annotation in the context of document image classification, which is one of the core components of modern document processing pipelines. The results of this study demonstrate that by utilizing active learning (AL), deep document classification models can achieve competitive performances to the models trained on fully annotated datasets and, in some cases, even surpass them by annotating only 15–40% of the total training dataset. Furthermore, this study demonstrates that modern AL strategies significantly outperform random querying, and in many cases achieve comparable performance to the models trained on fully annotated datasets even in the presence of practical deployment issues such as data imbalance, and annotation noise, and thus, offer tremendous benefits in real-world deployment of deep document classification models. The code to reproduce our experiments is publicly available at https://github.com/saifullah3396/doc_al.
引用
收藏
页码:187 / 209
页数:22
相关论文
共 50 条
  • [41] Multi-Class Active Learning for Image Classification
    Joshi, Ajay J.
    Porikli, Fatih
    Papanikolopoulos, Nikolaos
    CVPR: 2009 IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION, VOLS 1-4, 2009, : 2364 - +
  • [42] Minimal difference sampling for active learning image classification
    Wu, Jian
    Sheng, Sheng-Li
    Zhao, Peng-Peng
    Cui, Zhi-Ming
    Tongxin Xuebao/Journal on Communications, 2014, 35 (01): : 107 - 114
  • [43] Multi-label Active Learning for Image Classification
    Wu, Jian
    Sheng, Victor S.
    Zhang, Jing
    Zhao, Pengpeng
    Cui, Zhiming
    2014 IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING (ICIP), 2014, : 5227 - 5231
  • [44] Batch Mode Active Learning for Geographical Image Classification
    Wang, Zengmao
    Du, Bo
    Zhang, Lefei
    Hu, Wenbin
    Tao, Dacheng
    Zhang, Liangpei
    WEB TECHNOLOGIES AND APPLICATIONS (APWEB 2015), 2015, 9313 : 744 - 755
  • [45] Cold-start active learning for image classification
    Jin, Qiuye
    Yuan, Mingzhi
    Li, Shiman
    Wang, Haoran
    Wang, Manning
    Song, Zhijian
    INFORMATION SCIENCES, 2022, 616 : 16 - 36
  • [46] A Novel Active Learning Algorithm for Robust Image Classification
    Xiong, Xingliang
    Fan, Mingyu
    Yu, Chuang
    Hong, Zhenjie
    IEEE ACCESS, 2020, 8 : 71106 - 71116
  • [47] A Two -Stage Active Learning Method for Image Classification
    Wang, Feiyue
    Li, Xu
    Zhang, Yifan
    Wei, Baoguo
    Li, Lixin
    2022 IEEE 17TH CONFERENCE ON INDUSTRIAL ELECTRONICS AND APPLICATIONS (ICIEA), 2022, : 1134 - 1139
  • [48] Class-Balanced Active Learning for Image Classification
    Bengar, Javad Zolfaghari
    van de Weijer, Joost
    Fuentes, Laura Lopez
    Raducanu, Bogdan
    2022 IEEE WINTER CONFERENCE ON APPLICATIONS OF COMPUTER VISION (WACV 2022), 2022, : 3707 - 3716
  • [49] COMBINING ACTIVE AND METRIC LEARNING FOR HYPERSPECTRAL IMAGE CLASSIFICATION
    Pasolli, Edoardo
    Yang, Hsiuhan Lexie
    Crawford, Melba M.
    2014 6TH WORKSHOP ON HYPERSPECTRAL IMAGE AND SIGNAL PROCESSING: EVOLUTION IN REMOTE SENSING (WHISPERS), 2014,
  • [50] Integrating Multiple Information of Active Learning for Image Classification
    Xu, Haihui
    Zhao, Pengpeng
    Wu, Jian
    Cui, Zhiming
    Li, Chengchao
    2013 IEEE INTERNATIONAL CONFERENCE ON GRANULAR COMPUTING (GRC), 2013, : 374 - 379