PETNet: A YOLO-based prior enhanced transformer network for aerial image detection

被引:13
|
作者
Wang, Tianyu [1 ]
Ma, Zhongjing [1 ]
Yang, Tao [1 ]
Zou, Suli [1 ]
机构
[1] Beijing Inst Technol, Sch Automat, Beijing 100081, Peoples R China
基金
中国国家自然科学基金;
关键词
Deep learning; Transformer; Small object detection; Aerial image; OBJECT DETECTION; UAV;
D O I
10.1016/j.neucom.2023.126384
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Unmanned aerial vehicles (UAVs) have been applied to inspect in various scenarios due to their high effi-ciency, low cost, and excellent mobility. However, the objects in aerial images are much smaller and den-ser than general objects, causing it difficult for current object detection methods to achieve the expected results. To solve this issue, a prior enhanced Transformer network (PETNet) based on YOLO is proposed in this paper. Specifically, a novel prior enhanced Transformer (PET) module and a one-to-many feature fusion (OMFF) mechanism are proposed to embed into the network. Two additional detection heads are added to the shallow feature maps. In this work, PET is used to capture enhanced global information to improve the expressive ability of the network. The OMFF aims to fuse multi-type features to minimize the information loss of small objects. In addition, the added detection heads provide more possibility of detecting smaller-scale objects, and the extended multi-head parallel detection is more suitable for the multi-scale transformation of objects in aerial images. On the VisDrone-2021 and UAVDT databases, the proposed PETNet achieves state-of-the-art results with average precision (AP) of 35.3 and 21.5, respectively, which indicates that the proposed network is more suitable for aerial image detection and is of a great reference value.& COPY; 2023 Elsevier B.V. All rights reserved.
引用
收藏
页数:13
相关论文
共 50 条
  • [31] Enhanced object detection in pediatric bronchoscopy images using YOLO-based algorithms with CBAM attention mechanism
    Yan, Jianqi
    Zeng, Yifan
    Lin, Junhong
    Pei, Zhiyuan
    Fan, Jinrui
    Fang, Chuanyu
    Cai, Yong
    HELIYON, 2024, 10 (12)
  • [32] RescueNet: YOLO-based object detection model for detection and counting of flood survivors
    B. V. Balaji Prabhu
    R. Lakshmi
    R. Ankitha
    M. S. Prateeksha
    N. C. Priya
    Modeling Earth Systems and Environment, 2022, 8 : 4509 - 4516
  • [33] RescueNet: YOLO-based object detection model for detection and counting of flood survivors
    Prabhu, B. V. Balaji
    Lakshmi, R.
    Ankitha, R.
    Prateeksha, M. S.
    Priya, N. C.
    MODELING EARTH SYSTEMS AND ENVIRONMENT, 2022, 8 (04) : 4509 - 4516
  • [34] Light-YOLO: A Lightweight and Efficient YOLO-Based Deep Learning Model for Mango Detection
    Zhong, Zhengyang
    Yun, Lijun
    Cheng, Feiyan
    Chen, Zaiqing
    Zhang, Chunjie
    AGRICULTURE-BASEL, 2024, 14 (01):
  • [35] YOLO-Based Deep Learning Model for Pressure Ulcer Detection and Classification
    Aldughayfiq, Bader
    Ashfaq, Farzeen
    Jhanjhi, N. Z.
    Humayun, Mamoona
    HEALTHCARE, 2023, 11 (09)
  • [36] YOLO-Based Object Detection in Industry 4.0 Fischertechnik Model Environment
    Schneidereit, Slavomira
    Yarahmadi, Ashkan Mansouri
    Schneidereit, Toni
    Breuss, Michael
    Gebauer, Marc
    INTELLIGENT SYSTEMS AND APPLICATIONS, VOL 2, INTELLISYS 2023, 2024, 823 : 1 - 20
  • [37] Breast Lesions Detection and Classification via YOLO-Based Fusion Models
    Baccouche, Asma
    Garcia-Zapirain, Begonya
    Olea, Cristian Castillo
    Elmaghraby, Adel S.
    CMC-COMPUTERS MATERIALS & CONTINUA, 2021, 69 (01): : 1407 - 1425
  • [38] YOLO-based lightweight traffic sign detection algorithm and mobile deployment
    WU Yaqin
    ZHANG Tao
    NIU Jianjun
    CHANG Yan
    LIU Ganjun
    Optoelectronics Letters, 2025, 21 (04) : 249 - 256
  • [39] A Yolo-based Approach for Fire and Smoke Detection in IoT Surveillance Systems
    Zhang, Dawei
    INTERNATIONAL JOURNAL OF ADVANCED COMPUTER SCIENCE AND APPLICATIONS, 2024, 15 (01) : 87 - 94
  • [40] An optimized YOLO-based object detection model for crop harvesting system
    Junos, Mohamad Haniff
    Mohd Khairuddin, Anis Salwa
    Thannirmalai, Subbiah
    Dahari, Mahidzal
    IET IMAGE PROCESSING, 2021, 15 (09) : 2112 - 2125