Encoder- and Decoder-Based Networks Using Multiscale Feature Fusion and Nonlocal Block for Remote Sensing Image Semantic Segmentation

被引:11
|
作者
Wang, Yang [1 ]
Sun, Zhaochen [2 ]
Zhao, Wei [1 ]
机构
[1] Beihang Univ, Inst Elect Informat Engn, Beijing 100191, Peoples R China
[2] Xi An Jiao Tong Univ, Sch Sci, Xian 710049, Peoples R China
关键词
Decoding; Feature extraction; Image segmentation; Convolution; Semantics; Remote sensing; Interpolation; Channel attention block (CAB); decoder; multilevel feature fusion block (MLFFB); nonlocal block; remote sensing; semantic segmentation;
D O I
10.1109/LGRS.2020.2998680
中图分类号
P3 [地球物理学]; P59 [地球化学];
学科分类号
0708 ; 070902 ;
摘要
With the development of convolutional neural networks, the semantic segmentation of remote sensing images has been widely developed, but there are still some unsolved problems in this field due to the lack of multiscale information and the feature mismatch at the upsampling process. To solve these problems, we propose a network called multiscale feature fusion and alignment network (MFANet). MFANet is composed of an encoder and a decoder. The encoder contains a fully convolutional network, a multilevel feature fusion block (MLFFB), and a multiscale feature pyramid (MSFP). These subnetworks can obtain fine-grained feature maps that are full of multiscale and global features and improve segmentation results at multiple object scales. Moreover, MFANet uses a light convolution subnetwork, called decoder, to upsample the segmentation map stage by stage. Combining three scales of features, the decoder can promote the feature alignment at the upsampling stage. Along with the decoder, MFANet utilizes a multistage supervision loss to enhance the localization performance and boundary regression ability. Benefitting from the encoder and decoder structure and the innovative components inside encoder, MFANet is very powerful for the semantic segmentation of remote sensing images and can suit the complicated environment. We evaluate our MFANet on the Vaihingen and Potsdam data sets, and it outperforms the state-of-art methods both in the metric and visual effect.
引用
收藏
页码:1159 / 1163
页数:5
相关论文
共 50 条
  • [1] Semantic Segmentation of Remote Sensing Image Based on Encoder-Decoder Convolutional Neural Network
    Zhang Zhehan
    Fang Wei
    Du Lili
    Qiao Yanli
    Zhang Dongying
    Ding Guoshen
    ACTA OPTICA SINICA, 2020, 40 (03)
  • [2] Small sample remote sensing image segmentation based on multiscale feature fusion
    Wang J.
    Zhang J.
    Huazhong Keji Daxue Xuebao (Ziran Kexue Ban)/Journal of Huazhong University of Science and Technology (Natural Science Edition), 2022, 50 (03): : 62 - 67
  • [3] Semantic Segmentation of Remote Sensing Image Based on Multi-Scale Semantic Encoder-Decoder Network
    Liang Y.
    Yi C.-X.
    Wang G.-Y.
    Hu Y.-H.
    Tien Tzu Hsueh Pao/Acta Electronica Sinica, 2023, 51 (11): : 3199 - 3214
  • [4] Remote Sensing Image Semantic Segmentation Method Based on a Deep Convolutional Neural Network and Multiscale Feature Fusion
    Zhang, Guangzhen
    Jiang, Wangyang
    INTERNATIONAL JOURNAL ON SEMANTIC WEB AND INFORMATION SYSTEMS, 2023, 19 (01)
  • [5] A ViT-Based Multiscale Feature Fusion Approach for Remote Sensing Image Segmentation
    Wang, Wei
    Tang, Chen
    Wang, Xin
    Zheng, Bin
    IEEE GEOSCIENCE AND REMOTE SENSING LETTERS, 2022, 19
  • [6] Semantic Segmentation of Remote-Sensing Images Based on Multiscale Feature Fusion and Attention Refinement
    He, Xin
    Zhou, Yong
    Zhao, Jiaqi
    Zhang, Man
    Yao, Rui
    Liu, Bing
    Li, Haichao
    IEEE GEOSCIENCE AND REMOTE SENSING LETTERS, 2022, 19
  • [7] Semantic Segmentation of Remote-Sensing Images Based on Multiscale Feature Fusion and Attention Refinement
    He, Xin
    Zhou, Yong
    Zhao, Jiaqi
    Zhang, Man
    Yao, Rui
    Liu, Bing
    Li, Haichao
    IEEE Geoscience and Remote Sensing Letters, 2022, 19
  • [8] A Semantic Segmentation Method of Remote Sensing Image Based on Feature Fusion and Attention Mechanism
    Wang, Yiqin
    Dong, Yunyun
    JOURNAL OF INFORMATION PROCESSING SYSTEMS, 2024, 20 (05): : 640 - 653
  • [9] SEMANTIC SEGMENTATION OF REMOTE SENSING IMAGERY USING AN ENHANCED ENCODER-DECODER ARCHITECTURE
    Aburaed, N.
    Al-Saad, M.
    Alkhatib, M. Q.
    Zitouni, M. S.
    Almansoori, S.
    Al-Ahmad, H.
    GEOSPATIAL WEEK 2023, VOL. 10-1, 2023, : 1015 - 1020
  • [10] MFAFNet: A Multiscale Fully Attention Fusion Network for Remote Sensing Image Semantic Segmentation
    Dang, Yuanyuan
    Gao, Yu
    Liu, Bing
    IEEE ACCESS, 2024, 12 : 123388 - 123400