LOANet: a lightweight network using object attention for extracting buildings and roads from UAV aerial remote sensing images.

Han, Xiaoxiang; Liu, Yiman; Liu, Gang; Lin, Yuanjie; Liu, Qiaohong

Han, Xiaoxiang; Liu, Yiman; Liu, Gang; Lin, Yuanjie; Liu, Qiaohong.

Afiliación

Han X; School of Medical Instruments, Shanghai University of Medicine and Health Sciences, Shanghai, People's Republic of China.
Liu Y; School of Health Sciences and Engineering, University of Shanghai for Science and Technology, Shanghai, People's Republic of China.
Liu G; Department of Pediatric Cardiology, Shanghai Children's Medical Center, School of Medicine, Shanghai Jiao Tong University, Shanghai, People's Republic of China.
Lin Y; Shanghai Key Laboratory of Multidimensional Information Processing, School of Communication & Electronic Engineering, East China Normal University, Shanghai, People's Republic of China.
Liu Q; Key Laboratory of Earthquake Geodesy, Institute of Seismology, China Earthquake Administration, Wuhan, Hubei, People's Republic of China.

PeerJ Comput Sci ; 9: e1467, 2023.

Article en En | MEDLINE | ID: mdl-37547422

RESUMEN

Semantic segmentation for extracting buildings and roads from uncrewed aerial vehicle (UAV) remote sensing images by deep learning becomes a more efficient and convenient method than traditional manual segmentation in surveying and mapping fields. In order to make the model lightweight and improve the model accuracy, a lightweight network using object attention (LOANet) for buildings and roads from UAV aerial remote sensing images is proposed. The proposed network adopts an encoder-decoder architecture in which a lightweight densely connected network (LDCNet) is developed as the encoder. In the decoder part, the dual multi-scale context modules which consist of the atrous spatial pyramid pooling module (ASPP) and the object attention module (OAM) are designed to capture more context information from feature maps of UAV remote sensing images. Between ASPP and OAM, a feature pyramid network (FPN) module is used to fuse multi-scale features extracted from ASPP. A private dataset of remote sensing images taken by UAV which contains 2431 training sets, 945 validation sets, and 475 test sets is constructed. The proposed basic model performs well on this dataset, with only 1.4M parameters and 5.48G floating point operations (FLOPs), achieving excellent mean Intersection-over-Union (mIoU). Further experiments on the publicly available LoveDA and CITY-OSM datasets have been conducted to further validate the effectiveness of the proposed basic and large model, and outstanding mIoU results have been achieved. All codes are available on https://github.com/GtLinyer/LOANet.

Palabras clave

Context features; Lightweight network; Object attention; Remote sensing image; Semantic segmentation

Texto completo

Añadir a Mi BVS

Imprimir

XML

PubMed Links

Buscar en Google

Texto completo: 1 Colección: 01-internacional Base de datos: MEDLINE Idioma: En Revista: PeerJ Comput Sci Año: 2023 Tipo del documento: Article Pais de publicación: Estados Unidos

Texto completo

Añadir a Mi BVS

Imprimir

XML

PubMed Links

Buscar en Google