Res2Net: A New Multi-scale Backbone Architecture

Shang-Hua Gao; Ming-Ming Cheng; Kai Zhao; Xin-Yu Zhang; Ming-Hsuan Yang; Philip Torr

Res2Net：新しいマルチスケールバックボーンアーキテクチャ

複数のスケールで機能を表現することは、多くのビジョンタスクにとって非常に重要です。バックボーン畳み込みニューラルネットワーク（CNN）の最近の進歩は、より強力なマルチスケール表現能力を継続的に示しており、幅広いアプリケーションで一貫したパフォーマンスの向上につながります。ただし、ほとんどの既存の方法は、マルチスケール機能をレイヤーごとに表します。この論文では、単一の残余ブロック内に階層的な残余のような接続を構築することにより、CNNの新しいビルディングブロック、つまりRes2Netを提案します。 Res2Netは、マルチスケール機能を詳細なレベルで表し、各ネットワーク層の受容野の範囲を拡大します。提案されたRes2Netブロックは、ResNet、ResNeXt、DLAなどの最先端のバックボーンCNNモデルにプラグインできます。これらすべてのモデルでRes2Netブロックを評価し、CIFAR-100やImageNetなどの広く使用されているデータセットでベースラインモデルよりも一貫したパフォーマンスの向上を示しています。代表的なコンピュータービジョンタスク、つまりオブジェクト検出、クラスアクティベーションマッピング、および顕著なオブジェクト検出に関するさらなるアブレーション研究と実験結果は、最先端のベースライン手法に対するRes2Netの優位性をさらに検証します。ソースコードとトレーニング済みモデルは、https：//mmcheng.net/res2net/で入手できます。

Representing features at multiple scales is of great importance for numerous vision tasks. Recent advances in backbone convolutional neural networks (CNNs) continually demonstrate stronger multi-scale representation ability, leading to consistent performance gains on a wide range of applications. However, most existing methods represent the multi-scale features in a layer-wise manner. In this paper, we propose a novel building block for CNNs, namely Res2Net, by constructing hierarchical residual-like connections within one single residual block. The Res2Net represents multi-scale features at a granular level and increases the range of receptive fields for each network layer. The proposed Res2Net block can be plugged into the state-of-the-art backbone CNN models, e.g., ResNet, ResNeXt, and DLA. We evaluate the Res2Net block on all these models and demonstrate consistent performance gains over baseline models on widely-used datasets, e.g., CIFAR-100 and ImageNet. Further ablation studies and experimental results on representative computer vision tasks, i.e., object detection, class activation mapping, and salient object detection, further verify the superiority of the Res2Net over the state-of-the-art baseline methods. The source code and trained models are available on https://mmcheng.net/res2net/.

updated: Wed Jan 27 2021 09:55:20 GMT+0000 (UTC)

published: Tue Apr 02 2019 01:56:34 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト