A Recursive Prediction-Based Feature Enhancement for Small Object Detection

被引:0
|
作者
Xiao, Xiang [1 ]
Xue, Xiaorong [1 ]
Zhao, Zhiyuan [1 ]
Fan, Yisheng [1 ]
机构
[1] Liaoning Univ Technol, Sch Elect & Informat Engn, Jinzhou 121001, Peoples R China
关键词
small object detection; SAC; DINO; NWD;
D O I
10.3390/s24123856
中图分类号
O65 [分析化学];
学科分类号
070302 ; 081704 ;
摘要
Transformer-based methodologies in object detection have recently piqued considerable interest and have produced impressive results. DETR, an end-to-end object detection framework, ingeniously integrates the Transformer architecture, traditionally used in NLP, into computer vision for sequence-to-sequence prediction. Its enhanced variant, DINO, featuring improved denoising anchor boxes, has showcased remarkable performance on the COCO val2017 dataset. However, it often encounters challenges when applied to scenarios involving small object detection. Thus, we propose an innovative method for feature enhancement tailored to recursive prediction tasks, with a particular emphasis on augmenting small object detection performance. It primarily involves three enhancements: refining the backbone to favor feature maps that are more sensitive to small targets, incrementally augmenting the number of queries for small objects, and advancing the loss function for better performance. Specifically, The study incorporated the Switchable Atrous Convolution (SAC) mechanism, which features adaptable dilated convolutions, to increment the receptive field and thus elevate the innate feature extraction capabilities of the primary network concerning diminutive objects. Subsequently, a Recursive Small Object Prediction (RSP) module was designed to enhance the feature extraction of the prediction head for more precise network operations. Finally, the loss function was augmented with the Normalized Wasserstein Distance (NWD) metric, tailoring the loss function to suit small object detection better. The efficacy of the proposed model is empirically confirmed via testing on the VISDRONE2019 dataset. The comprehensive array of experiments indicates that our proposed model outperforms the extant DINO model in terms of average precision (AP) small object detection.
引用
收藏
页数:16
相关论文
共 50 条
  • [41] Small object detection model based on feature fusion of attention mechanism
    Chen H.
    Zhen X.
    Zhao T.
    Huazhong Keji Daxue Xuebao (Ziran Kexue Ban)/Journal of Huazhong University of Science and Technology (Natural Science Edition), 2023, 51 (03): : 60 - 66
  • [42] Small Object Detection in Traffic Scenes Based on Attention Feature Fusion
    Lian, Jing
    Yin, Yuhang
    Li, Linhui
    Wang, Zhenghao
    Zhou, Yafu
    SENSORS, 2021, 21 (09)
  • [43] Dual prediction-based reporting for object tracking sensor networks
    Xu, YQ
    Winter, J
    Lee, WC
    PROCEEDINGS OF MOBIQUITOUS 2004, 2004, : 154 - 163
  • [44] Fuzzy grey prediction-based particle filter for object tracking
    Yang L.
    Lu Z.
    Mathematical and Computational Applications, 2016, 21 (03)
  • [45] Research on Small Object Detection Based on Feature Fusion and Attention Mechanism
    Liu, Jianwei
    Liu, Zheng
    Lu, Jingwen
    Li, Chuancan
    Chen, Gangqiang
    39TH YOUTH ACADEMIC ANNUAL CONFERENCE OF CHINESE ASSOCIATION OF AUTOMATION, YAC 2024, 2024, : 2285 - 2291
  • [46] A small object detection algorithm based on feature interaction and guided learning
    Shao, Xiang-Ying
    Guo, Ying
    Wang, You-Wei
    Bao, Zheng-Wei
    Wang, Ji-Yu
    JOURNAL OF VISUAL COMMUNICATION AND IMAGE REPRESENTATION, 2024, 98
  • [47] Prediction-based Object Tracking and Coverage in Visual Sensor Networks
    Chen, Tzung-Shi
    Peng, Jiun-Jie
    Lee, De-Wei
    Tsai, Hua-Wen
    2011 7TH INTERNATIONAL WIRELESS COMMUNICATIONS AND MOBILE COMPUTING CONFERENCE (IWCMC), 2011, : 278 - 284
  • [48] Object Detection Method Based on Shallow Feature Fusion and Semantic Information Enhancement
    Luo, Huilan
    Wang, Pei
    Chen, Hongkun
    Xu, Min
    IEEE SENSORS JOURNAL, 2021, 21 (19) : 21839 - 21851
  • [49] VFEDet: A Variational Information Bottleneck Based Feature Enhancement Object Detection Network
    Wu, Mingyu
    Zhu, Ming
    Tang, Ruixue
    TWELFTH INTERNATIONAL CONFERENCE ON GRAPHICS AND IMAGE PROCESSING (ICGIP 2020), 2021, 11720
  • [50] Oriented Object Detection Based on Foreground Feature Enhancement in Remote Sensing Images
    Lin, Peng
    Wu, Xiaofeng
    Wang, Bin
    REMOTE SENSING, 2022, 14 (24)