PETNet: A YOLO-based prior enhanced transformer network for aerial image detection

被引:13
|
作者
Wang, Tianyu [1 ]
Ma, Zhongjing [1 ]
Yang, Tao [1 ]
Zou, Suli [1 ]
机构
[1] Beijing Inst Technol, Sch Automat, Beijing 100081, Peoples R China
基金
中国国家自然科学基金;
关键词
Deep learning; Transformer; Small object detection; Aerial image; OBJECT DETECTION; UAV;
D O I
10.1016/j.neucom.2023.126384
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Unmanned aerial vehicles (UAVs) have been applied to inspect in various scenarios due to their high effi-ciency, low cost, and excellent mobility. However, the objects in aerial images are much smaller and den-ser than general objects, causing it difficult for current object detection methods to achieve the expected results. To solve this issue, a prior enhanced Transformer network (PETNet) based on YOLO is proposed in this paper. Specifically, a novel prior enhanced Transformer (PET) module and a one-to-many feature fusion (OMFF) mechanism are proposed to embed into the network. Two additional detection heads are added to the shallow feature maps. In this work, PET is used to capture enhanced global information to improve the expressive ability of the network. The OMFF aims to fuse multi-type features to minimize the information loss of small objects. In addition, the added detection heads provide more possibility of detecting smaller-scale objects, and the extended multi-head parallel detection is more suitable for the multi-scale transformation of objects in aerial images. On the VisDrone-2021 and UAVDT databases, the proposed PETNet achieves state-of-the-art results with average precision (AP) of 35.3 and 21.5, respectively, which indicates that the proposed network is more suitable for aerial image detection and is of a great reference value.& COPY; 2023 Elsevier B.V. All rights reserved.
引用
收藏
页数:13
相关论文
共 50 条
  • [21] A Yolo-Based Model for Breast Cancer Detection in Mammograms
    Prinzi, Francesco
    Insalaco, Marco
    Orlando, Alessia
    Gaglio, Salvatore
    Vitabile, Salvatore
    COGNITIVE COMPUTATION, 2024, 16 (01) : 107 - 120
  • [22] YOLO-based CAD framework with ViT transformer for breast mass detection and classification in CESM and FFDM images
    Hassan, Nada M.
    Hamad, Safwat
    Mahar, Khaled
    NEURAL COMPUTING & APPLICATIONS, 2024, 36 (12): : 6467 - 6496
  • [23] YOLO-based CAD framework with ViT transformer for breast mass detection and classification in CESM and FFDM images
    Nada M. Hassan
    Safwat Hamad
    Khaled Mahar
    Neural Computing and Applications, 2024, 36 : 6467 - 6496
  • [24] Image Quantization Tradeoffs in a YOLO-based FPGA Accelerator Framework
    Yarnell, R.
    Hossain, M.
    DeMara, R. F.
    2023 24TH INTERNATIONAL SYMPOSIUM ON QUALITY ELECTRONIC DESIGN, ISQED, 2023, : 654 - 660
  • [25] UAV aerial image target detection based on BLUR-YOLO
    Huang, Tongyuan
    Zhu, Jinjiang
    Liu, Yao
    Tan, Yu
    REMOTE SENSING LETTERS, 2023, 14 (02) : 186 - 196
  • [26] YOLO-Anti: YOLO-based counterattack model for unseen congested object detection
    Wang, Kun
    Liu, Maozhen
    PATTERN RECOGNITION, 2022, 131
  • [27] A Yolo-based Violence Detection Method in IoT Surveillance Systems
    Gao, Hui
    INTERNATIONAL JOURNAL OF ADVANCED COMPUTER SCIENCE AND APPLICATIONS, 2023, 14 (08) : 143 - 149
  • [28] YOLO-DA: An Efficient YOLO-Based Detector for Remote Sensing Object Detection
    Lin, Jiehua
    Zhao, Yan
    Wang, Shigang
    Tang, Yu
    IEEE GEOSCIENCE AND REMOTE SENSING LETTERS, 2023, 20
  • [29] Experimental Study on YOLO-Based Leather Surface Defect Detection
    Chen, Zhiqiang
    Zhu, Qirui
    Zhou, Xiaofan
    Deng, Jiehang
    Song, Wei
    IEEE ACCESS, 2024, 12 : 32830 - 32848
  • [30] YOLO-based Object Detection Models: A Review and its Applications
    Vijayakumar, Ajantha
    Vairavasundaram, Subramaniyaswamy
    MULTIMEDIA TOOLS AND APPLICATIONS, 2024, 83 (35) : 83535 - 83574