DensePASS: Dense Panoramic Semantic Segmentation via Unsupervised Domain Adaptation with Attention-Augmented Context Exchange

Chaoxiang Ma; Jiaming Zhang; Kailun Yang; Alina Roitberg; Rainer Stiefelhagen

DensePASS：注意増強コンテキスト交換を伴う教師なしドメイン適応による高密度パノラマセマンティックセグメンテーション

インテリジェント車両は、360度センサーの拡張された視野（FoV）の恩恵を明らかに受けていますが、利用可能なセマンティックセグメンテーショントレーニング画像の大部分はピンホールカメラでキャプチャされています。この作業では、ドメイン適応のレンズを通してこの問題を調べ、パノラマセマンティックセグメンテーションを設定にもたらします。この設定では、ラベル付けされたトレーニングデータが従来のピンホールカメラ画像の異なる分布に由来します。まず、パノラマセマンティックセグメンテーションの教師なしドメイン適応のタスクを形式化します。ここでは、ピンホールカメラデータのソースドメインからのラベル付きの例でトレーニングされたネットワークが、ラベルが利用できないパノラマ画像の別のターゲットドメインに展開されます。このアイデアを検証するために、DensePASSを収集して公開します。これは、クロスドメイン条件下でのパノラマセグメンテーション用の新しい高密度注釈付きデータセットであり、ピンホールからパノラマへの転送を研究するために特別に構築され、Cityscapesから取得したピンホールカメラトレーニングの例が付属しています。 DensePASSは、ラベル付きとラベルなしの両方の360度画像をカバーし、ラベル付きデータは、ソースドメイン（ピンホール）データで使用可能なカテゴリに明示的に適合する19のクラスで構成されます。ドメインシフトの課題に対処するために、注意ベースのメカニズムの現在の進歩を活用し、注意増強ドメイン適応モジュールのさまざまなバリアントに基づいて、クロスドメインパノラマセマンティックセグメンテーションの汎用フレームワークを構築します。私たちのフレームワークは、ドメインの対応を学習するときにローカルレベルとグローバルレベルでの情報交換を容易にし、2つの標準セグメンテーションネットワークのドメイン適応パフォーマンスを平均IoUで6.05％と11.26％向上させます。

Intelligent vehicles clearly benefit from the expanded Field of View (FoV) of the 360-degree sensors, but the vast majority of available semantic segmentation training images are captured with pinhole cameras. In this work, we look at this problem through the lens of domain adaptation and bring panoramic semantic segmentation to a setting, where labelled training data originates from a different distribution of conventional pinhole camera images. First, we formalize the task of unsupervised domain adaptation for panoramic semantic segmentation, where a network trained on labelled examples from the source domain of pinhole camera data is deployed in a different target domain of panoramic images, for which no labels are available. To validate this idea, we collect and publicly release DensePASS - a novel densely annotated dataset for panoramic segmentation under cross-domain conditions, specifically built to study the Pinhole-to-Panoramic transfer and accompanied with pinhole camera training examples obtained from Cityscapes. DensePASS covers both, labelled- and unlabelled 360-degree images, with the labelled data comprising 19 classes which explicitly fit the categories available in the source domain (i.e. pinhole) data. To meet the challenge of domain shift, we leverage the current progress of attention-based mechanisms and build a generic framework for cross-domain panoramic semantic segmentation based on different variants of attention-augmented domain adaptation modules. Our framework facilitates information exchange at local- and global levels when learning the domain correspondences and improves the domain adaptation performance of two standard segmentation networks by 6.05% and 11.26% in Mean IoU.

updated: Fri Aug 13 2021 20:15:46 GMT+0000 (UTC)

published: Fri Aug 13 2021 20:15:46 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト