节点文献

基于注意力机制的多尺度融合航拍影像语义分割

Semantic Segmentation of Multi-Scale Fusion Aerial Image Based on Attention Mechanism

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 郑顾平王敏李刚

【Author】 ZHENG Guping;WANG Min;LI Gang;School of Computer and Control Engineering,North China Electric Power University;

【通讯作者】 李刚;

【机构】 华北电力大学控制与计算机工程学院

【摘要】 航拍影像同一场景不同对象尺度差异较大,采用单一尺度的分割往往无法达到最佳的分类效果。为解决这一问题,提出一种基于注意力机制的多尺度融合模型。首先,利用不同采样率的扩张卷积提取航拍影像的多个尺度特征;然后,在多尺度融合阶段引入注意力机制,使模型能够自动聚焦于合适的尺度,并为所有尺度及每个位置像素分别赋予权重;最后,将加权融合后的特征图上采样到原图大小,对航拍影像的每个像素进行语义标注。实验结果表明,与传统的FCN、DeepLab语义分割模型及其他航拍影像分割模型相比,基于注意力机制的多尺度融合模型不仅具有更高的分割精度,而且可以通过对各尺度特征对应权重图的可视化,分析不同尺度及位置像素的重要性。

【Abstract】 In aerial images, there is significant difference between the scales of different objects in the same scene, single-scale segmentation often hardly achieves the best classification effect. In order to solve the problem, we proposes a multi-scale fusion model based on attention mechanism. Firstly, extract multi-scale features of the aerial image using dilated convolutions with different sampling rates; then utilize the attention mechanism in the multi-scale fusion stage, so that the model can automatically focus on the appropriate scale, and learn to put different weights on all scale and each pixel location; finally, the weighted sum of feature map is sampled to the original image size, and each pixel of aerial image is semantically labeled. The experiment demonstrates that compared with the traditional FCN and DeepLab method, and other aerial image segmentation model, the multi-scale fusion model based on attention mechanism not only has higher segmentation accuracy, but also can analyze the importance of different scales and pixel location by visualizing the weight map corresponding to each scale feature.

【基金】 国家自然科学基金项目(51407076);中央高校基本科研业务费专项资金(2018MS075)
  • 【文献出处】 图学学报 ,Journal of Graphics , 编辑部邮箱 ,2018年06期
  • 【分类号】TP751;TP183
  • 【被引频次】6
  • 【下载频次】313
节点文献中: 

本文链接的文献网络图示:

本文的引文网络