FMRNet: A network structure for enhanced ship detection model training using feature maps

  • role: First author第一作者
  • Affiliation:

    National Key Laboratory of Science and Technology on Multispectral Information Processing, School of Artificial Intelligence and Automation, Huazhong University of Science and Technology, Wuhan 430072, China

  • Email:m202072950@hust.edu.cn
  • Introduction:E-mail m202072950@hust.edu.cn
ZHANG Zekun1,  
  • Affiliation:

    National Key Laboratory of Science and Technology on Multispectral Information Processing, School of Artificial Intelligence and Automation, Huazhong University of Science and Technology, Wuhan 430072, China

TAN Zhenbiao1,  
  • Affiliation:

    China Railway Siyuan Survey and Design Group, Wuhan 430063, China

YU Kun3,  
  • Affiliation:

    National Key Laboratory of Science and Technology on Multispectral Information Processing, School of Artificial Intelligence and Automation, Huazhong University of Science and Technology, Wuhan 430072, China

FANG Bing1,  
  • Affiliation:

    China Ship Development, and Design Center, Wuhan 430064, China

HUANG Xiao2,  
  • role: Corresponding author通信作者
  • Affiliation:

    National Key Laboratory of Science and Technology on Multispectral Information Processing, School of Artificial Intelligence and Automation, Huazhong University of Science and Technology, Wuhan 430072, China

  • Email:majie@hust.edu.cn
  • Introduction:E-mailmajie@hust.edu.cn
MA Jie1*

реферат

With the ongoing development of artificial intelligence technology, deep learning methods have become increasingly important role in the field of ship detection. However, the false alarms and missed detections that appear in deep learning algorithms hinder the application of technology in the field of ship detection. Although the classical deep learning methods can effectively deal with a single-background sea surface, the classical models can easily yield false alarms on shore when faced with data under complex backgrounds. In custom training, the model often tends to overly emphasize some salient features, which leads to feature overfitting. Detection can easily be missed when these salient features change. In the process of forward propagation of the model to the input, different network layers in the model generate corresponding mappings or feature maps from the input. Fully utilizing the semantic and spatial information of the feature maps is an effective way to reduce false alarms and missed detections. Compared with the traditional model, our proposed Feature Map Reinforcement Network (FMRNnet) can fully utilize feature maps to generate adaptive feature map masks and water–land segmentation masks. This method ultimately reduces false alarms and missed detections by avoiding feature overfitting of the model and weakening the effects caused by complex backgrounds. In FMRNnet, we design the Self Feature-map Mask Module (SFMM), which can selectively utilize the feature map through the attention mechanism for generating an adaptive mask. The mask prevents the model from focusing on a single feature point, which prevents feature overfitting. We also propose a Feature-map Sea-Land Segmentation Module (FSSM) that is parallel to SFMM. It reduces the false alarms of ship targets appearing in the land area by introducing the fusion between the water-land segmentation mask and the feature map. The experimental results, when compared with SOTA algorithms on publicly available datasets, show that the performance of the proposed method in this study is excellent and outperforms that of other SOTA algorithms. After FMRNet is added, the 10-fold average mAP value of the detection algorithm ROI trans largely improves. This enhancement increases the mean value of baseline mAP from 86.1% to 90.8%, which surpasses that of other SOTA algorithms. Benefiting from the adaptive mask, the mAP value of the model including the SFMM module is 90.4%, which achieves a 4.2% improvement over the baseline. Owing to the priori knowledge learned from the water-land distribution, FSSM improves the precision and recall of the model, which results in a MAP value of 86.4%. For the task of ship detection, we propose a novel backbone network, that is, FMRNet based on Resnet. Our proposed SFMM module enables the model to discern the target from multiple features for avoiding the overfitting of salient features. We design the FSSM module to reduce the false alarms caused by complex backgrounds. Suppressing the non-water surface area reduces the confidence level of targets appearing in non-water surface. FSSM achieves the purpose of removing unreasonable false alarms while improving the accuracy of the model.

ключеви́че слова́

remote sensing imaging;artificial intelligence;ship detection;neural networks;feature maps;network reinforcement;rotational object detection;overfitting suppression

References

  1. 1.
    Bi F K, Hou J Y, Chen L, Yang Z H and Wang Y P. 2019. Ship detection for optical remote sensing images based on visual attention enhanced network. Sensors, 19(10): 2271
  2. 2.
    Ding G D, Kha S, Tang Z M and Porikli F. 2017. Let features decide for themselves: feature mask network for person re-identification. arXiv: 1711.07155
  3. 3.
    Ding J, Xue N, Long Y, Xia G S and Lu Q K. 2018. Learning RoI transformer for detecting oriented objects in aerial images. arXiv: 1812.00155
  4. 4.
    Girshick R. 2015. Fast R-CNN//2015 IEEE International Conference on Computer Vision. Santiago: IEEE: 1440-1448
  5. 5.
    Goetz A F H, Rock B N and Rowan L C. 1983. Remote sensing for exploration; an overview. Economic Geology, 78(4): 573-590
  6. 6.
    Gu Z Y. 2010. Research on high performance computing platform of ocean environmental information processing//2010 International Conference on Optics, Photonics and Energy Engineering. Wuhan: IEEE: 298-300
  7. 7.
    He K M, Zhang X Y, Ren S Q and Sun J. 2016. Deep residual learning for image recognition//2016 IEEE Conference on Computer Vision and Pattern Recognition. Las Vegas: IEEE: 770-778
  8. 8.
    Hu J, Shen L and Sun G. 2018. Squeeze-and-excitation networks//2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition. Salt Lake City: IEEE: 7132-7141
  9. 9.
    Lunetta R S, Congalton R C, Fenstermaker L K, Jensen J R, McGwire K C and Tinney L R. 1991. Remote sensing and geographic information system data integration: error sources and research issues. Photogrammetric Engineering and Remote Sensing, 57(6): 677-687
  10. 10.
    Nie X, Duan M Y, Ding H X, Hu B L and Wong E K. 2020. Attention mask R-CNN for ship detection and segmentation from remote sensing images. IEEE Access, 8: 9325-9334
  11. 11.
    Qian W, Yang X, Peng S L, Guo Y and Yan J C. 2019. Learning modulated loss for rotated object detection. arXiv: 1911.08299
  12. 12.
    Redmon J, Divvala S, Girshick R and Farhadi A. 2016. You only look once: unified, real-time object detection//Proceedings of the 2016 IEEE Conference on Computer Vision and Pattern Recognition. Las Vegas: IEEE: 779-788
  13. 13.
    Redmon J and Farhadi A. 2018. YOLOv3: an incremental improvement. arXiv: 1804.02767v1
  14. 14.
    Simonyan K and Zisserman A. 2015. Very deep convolutional networks for large-scale image recognition. arXiv:1409.1556
  15. 15.
    Xu K Y, Hu W H and Chen R. 2006. A vessel targets detection method based on Canny algorithm. Infrared Technology, 28(6): 315-318
  16. 16.
    Xu Y C, Fu M T, Wang Q M, Wang Y K, Chen K, Xia G S and Bai X. 2021. Gliding vertex on the horizontal bounding box for multi-oriented object detection. IEEE Transactions on Pattern Analysis and Machine Intelligence, 43(4): 1452-1459

Читать полностью

The above content is generated by Large Model Translation. The translated content is for reference only. We do not assume any commercial or legal responsibilty for any consequences arising from the use of our website