A Joint Framework Towards Class-aware and Class-agnostic Alignment for Few-shot Segmentation

Kai Huang; Mingfei Cheng; Yang Wang; Bochen Wang; Ye Xi; Feigege Wang; Peng Chen

少数ショットセグメンテーションのためのクラス認識およびクラス非依存のアライメントに向けた共同フレームワーク

少数ショットセグメンテーション (FSS) は、少数の注釈付きサポートイメージのみが与えられた場合に、目に見えないクラスのオブジェクトをセグメント化することを目的としています。ほとんどの既存の方法は、クエリの特徴を独立したサポートプロトタイプと単純につなぎ合わせ、混合された特徴をデコーダに供給することによってクエリイメージをセグメント化します。大幅な改善が達成されましたが、既存のメソッドは、クラスのバリアントと背景の混乱により、依然としてクラスのバイアスに直面しています。このホワイトペーパーでは、セグメンテーションを容易にするために、より価値のあるクラス認識とクラスにとらわれないアライメントガイダンスを組み合わせた共同フレームワークを提案します。具体的には、対応するサポート機能から各クエリ画像に最も関連するクラス認識情報をマイニングするために、マルチスケールクエリサポート対応を確立するハイブリッドアライメントモジュールを設計します。さらに、基本クラスの知識を利用して、すべてのオブジェクト領域、特に目に見えないクラスのオブジェクト領域を強調表示することにより、実際の背景と前景を区別するクラスに依存しない事前マスクを生成することを検討します。クラス認識およびクラス非依存のアライメントガイダンスを一緒に集約することにより、クエリイメージでより優れたセグメンテーションパフォーマンスが得られます。 PASCAL-5^i および COCO-20^i データセットでの広範な実験により、提案されたジョイントフレームワークが、特に 1 ショット設定でより優れたパフォーマンスを発揮することが実証されました。

Few-shot segmentation (FSS) aims to segment objects of unseen classes given only a few annotated support images. Most existing methods simply stitch query features with independent support prototypes and segment the query image by feeding the mixed features to a decoder. Although significant improvements have been achieved, existing methods are still face class biases due to class variants and background confusion. In this paper, we propose a joint framework that combines more valuable class-aware and class-agnostic alignment guidance to facilitate the segmentation. Specifically, we design a hybrid alignment module which establishes multi-scale query-support correspondences to mine the most relevant class-aware information for each query image from the corresponding support features. In addition, we explore utilizing base-classes knowledge to generate class-agnostic prior mask which makes a distinction between real background and foreground by highlighting all object regions, especially those of unseen classes. By jointly aggregating class-aware and class-agnostic alignment guidance, better segmentation performances are obtained on query images. Extensive experiments on PASCAL-5^i and COCO-20^i datasets demonstrate that our proposed joint framework performs better, especially on the 1-shot setting.

updated: Wed Nov 02 2022 17:33:25 GMT+0000 (UTC)

published: Wed Nov 02 2022 17:33:25 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト