MultiScale Probability Map guided Index Pooling with Attention-based learning for Road and Building Segmentation

Shirsha Bose; Ritesh Sur Chowdhury; Debabrata Pal; Shivashish Bose; Biplab Banerjee; Subhasis Chaudhuri

道路と建物のセグメンテーションのための注意ベースの学習による MultiScale Probability Map ガイド付きインデックスプーリング

衛星画像からの効率的な道路と建物のフットプリントの抽出は、多くのリモートセンシングアプリケーションで優勢です。ただし、正確なセグメンテーションマップの抽出は、樹木によってカモフラージュされた多様な建物の構造、道路と建物の間の同様のスペクトル応答、および道路上の不均一な交通による遮蔽のために、非常に困難です。既存の畳み込みニューラルネットワーク (CNN) ベースの方法は、建物抽出のための強化された空間セマンティクス学習またはきめの細かい道路トポロジ抽出のいずれかに焦点を当てています。 CNN の従来のプーリングメカニズムによる深刻なセマンティック情報の損失により、複雑な環境に密集した小さな建物の断片化された切断された道路地図と不十分にセグメント化された境界が生成されます。この論文では、2つの新しいモジュールDynamic Attention Map Guided Index Pooling（DAMIP）とDynamic Attention Map Guided Spatialを搭載した、新しい注意認識セグメンテーションフレームワーク、Multi-Scale Supervised Dilated Multi-Path Attention Network（MSSDMPA-Net）を提案します。チャネルアテンション (DAMSCA) を使用して、リモートセンシングされた画像から建物のフットプリントと道路地図を正確に抽出します。 DAMIP は、新しいインデックスプーリングメカニズムを使用して重要な幾何学的情報を保持することにより、顕著な特徴をマイニングします。一方、DAMSCA はマルチスケールの空間的およびスペクトル的特徴を同時に抽出します。さらに、MSSDMPA-Net を最適化する際に膨張畳み込みとマルチスケールの深い監視を使用すると、優れたパフォーマンスを達成するのに役立ちます。複数のベンチマーク建物および道路抽出データセットに関する実験結果により、MSSDMPA-Net が建物および道路抽出の最先端 (SOTA) 手法であることを確認できます。

Efficient road and building footprint extraction from satellite images are predominant in many remote sensing applications. However, precise segmentation map extraction is quite challenging due to the diverse building structures camouflaged by trees, similar spectral responses between the roads and buildings, and occlusions by heterogeneous traffic over the roads. Existing convolutional neural network (CNN)-based methods focus on either enriched spatial semantics learning for the building extraction or the fine-grained road topology extraction. The profound semantic information loss due to the traditional pooling mechanisms in CNN generates fragmented and disconnected road maps and poorly segmented boundaries for the densely spaced small buildings in complex surroundings. In this paper, we propose a novel attention-aware segmentation framework, Multi-Scale Supervised Dilated Multiple-Path Attention Network (MSSDMPA-Net), equipped with two new modules Dynamic Attention Map Guided Index Pooling (DAMIP) and Dynamic Attention Map Guided Spatial and Channel Attention (DAMSCA) to precisely extract the building footprints and road maps from remotely sensed images. DAMIP mines the salient features by employing a novel index pooling mechanism to retain important geometric information. On the other hand, DAMSCA simultaneously extracts the multi-scale spatial and spectral features. Besides, using dilated convolution and multi-scale deep supervision in optimizing MSSDMPA-Net helps achieve stellar performance. Experimental results over multiple benchmark building and road extraction datasets, ensures MSSDMPA-Net as the state-of-the-art (SOTA) method for building and road extraction.

updated: Sat Feb 18 2023 19:57:25 GMT+0000 (UTC)

published: Sat Feb 18 2023 19:57:25 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト