Elucidating image-to-set prediction: An analysis of models, losses and datasets

Luis Pineda; Amaia Salvador; Michal Drozdzal; Adriana Romero

画像からセットへの予測の解明：モデル、損失、データセットの分析

このホワイトペーパーでは、公開されたメソッド間の適切な比較を妨げる、イメージからセットへの予測に関する文献の重要な再現性の課題を特定します。つまり、研究者はさまざまな評価プロトコルを使用して寄与を評価します。この問題を緩和するために、マルチラベル分類（VOC、COCO、NUS-WIDE、ADE20k、Recipe1M）に適した、タスクの複雑さが増している5つのパブリックデータセットの上に構築されたイメージからセットへの予測ベンチマークスイートを紹介します。ベンチマークを使用して、現在のモデルの主要なコンポーネント、つまり画像表現のバックボーンの選択とセットの予測子の設計を研究する詳細な分析を提供します。私たちの結果は、（1）より良い画像表現バックボーンを活用すると、セット予測子を強化するよりもパフォーマンスが向上し、（2）ラベルの共起と順序付けの両方をモデル化すると、パフォーマンスの点でわずかにプラスの影響があるが、明示的なカーディナリティ予測のみであるRecipe1Mなどの複雑なデータセットをトレーニングするときに役立ちます。将来の画像からセットへの予測調査を容易にするために、https：//github.com/facebookresearch/image-to-setでコード、最適なモデル、データセットの分割を公開しています。

In this paper, we identify an important reproducibility challenge in the image-to-set prediction literature that impedes proper comparisons among published methods, namely, researchers use different evaluation protocols to assess their contributions. To alleviate this issue, we introduce an image-to-set prediction benchmark suite built on top of five public datasets of increasing task complexity that are suitable for multi-label classification (VOC, COCO, NUS-WIDE, ADE20k and Recipe1M). Using the benchmark, we provide an in-depth analysis where we study the key components of current models, namely the choice of the image representation backbone as well as the set predictor design. Our results show that (1) exploiting better image representation backbones leads to higher performance boosts than enhancing set predictors, and (2) modeling both the label co-occurrences and ordering has a slight positive impact in terms of performance, whereas explicit cardinality prediction only helps when training on complex datasets, such as Recipe1M. To facilitate future image-to-set prediction research, we make the code, best models and dataset splits publicly available at: https://github.com/facebookresearch/image-to-set.

updated: Wed May 27 2020 04:02:31 GMT+0000 (UTC)

published: Thu Apr 11 2019 14:10:53 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト