Encoder- and Decoder-Based Networks Using Multiscale Feature Fusion and Nonlocal Block for Remote Sensing Image Semantic Segmentation

Yang Wang,Wei Zhao,Zhaochen Sun

doi:10.1109/lgrs.2020.2998680

Yang Wang, Wei Zhao + Show 1 more

Open Access

https://doi.org/10.1109/lgrs.2020.2998680

Copy DOI

Journal: IEEE Geoscience and Remote Sensing Letters	Publication Date: Jun 18, 2020
Citations: 10	License type: publisher-specific-oa

Affiliation: Beihang University

Abstract

With the development of convolutional neural networks, the semantic segmentation of remote sensing images has been widely developed, but there are still some unsolved problems in this field due to the lack of multiscale information and the feature mismatch at the upsampling process. To solve these problems, we propose a network called multiscale feature fusion and alignment network (MFANet). MFANet is composed of an encoder and a decoder. The encoder contains a fully convolutional network, a multilevel feature fusion block (MLFFB), and a multiscale feature pyramid (MSFP). These subnetworks can obtain fine-grained feature maps that are full of multiscale and global features and improve segmentation results at multiple object scales. Moreover, MFANet uses a light convolution subnetwork, called decoder, to upsample the segmentation map stage by stage. Combining three scales of features, the decoder can promote the feature alignment at the upsampling stage. Along with the decoder, MFANet utilizes a multistage supervision loss to enhance the localization performance and boundary regression ability. Benefitting from the encoder and decoder structure and the innovative components inside encoder, MFANet is very powerful for the semantic segmentation of remote sensing images and can suit the complicated environment. We evaluate our MFANet on the Vaihingen and Potsdam data sets, and it outperforms the state-of-art methods both in the metric and visual effect.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Encoder- and Decoder-Based Networks Using Multiscale Feature Fusion and Nonlocal Block for Remote Sensing Image Semantic Segmentation

Abstract

Talk to us

Similar Papers

More From: IEEE Geoscience and Remote Sensing Letters

Lead the way for us

Similar Papers

Many heads are better than one: A multiscale neural information feature fusion framework for spatial route selections decoding from multichannel neural recordings of pigeons
Mengmeng Li ... Hong Wan
Brain Research Bulletin | VOL. 184
Mengmeng Li, et. al.Mengmeng Li ... Hong Wan
12 Mar 2022
Brain Research Bulletin | VOL. 184

Multi-scale Adaptive Feature Fusion Network for Semantic Segmentation in Remote Sensing Images
Ronghua Shang ... Jiyu Zhang
Remote Sensing | VOL. 12
Ronghua Shang, et. al.Ronghua Shang ... Jiyu Zhang
09 Mar 2020
Remote Sensing | VOL. 12

MFEFNet: A Multi-Scale Feature Information Extraction and Fusion Network for Multi-Scale Object Detection in UAV Aerial Images
Liming Zhou ... Yadi Wang
Drones | VOL. 8
Liming Zhou, et. al.Liming Zhou ... Yadi Wang
08 May 2024
Drones | VOL. 8

Multi-Stage Multi-Scale Local Feature Fusion for Infrared Small Target Detection
Yahui Wang ... Yiping Xu
Remote Sensing | VOL. 15
Yahui Wang, et. al.Yahui Wang ... Yiping Xu
13 Sep 2023
Remote Sensing | VOL. 15

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Encoder- and Decoder-Based Networks Using Multiscale Feature Fusion and Nonlocal Block for Remote Sensing Image Semantic Segmentation

Abstract

Talk to us

Similar Papers

More From: IEEE Geoscience and Remote Sensing Letters