A Step Toward More Inclusive People Annotations for Fairness

Candice Schumann; Susanna Ricco; Utsav Prabhu; Vittorio Ferrari; Caroline Pantofaru

公平性のためのより包括的な人々の注釈に向けた一歩

Open Images Datasetには、約900万枚の画像が含まれており、コンピュータービジョン研究で広く受け入れられているデータセットです。大規模なデータセットで一般的に行われているように、注釈は網羅的ではなく、各画像のクラスのサブセットのみに境界ボックスと属性ラベルがあります。このホワイトペーパーでは、MIAP（More Inclusive Annotations for People）サブセットと呼ばれるOpen Imagesデータセットのサブセットに新しい注釈のセットを示します。これには、これらの画像に表示されるすべての人の境界ボックスと属性が含まれます。 MIAPサブセットの属性とラベル付け方法は、モデルの公平性の調査を可能にするために設計されました。さらに、人物クラスとそのサブクラスの元の注釈方法を分析し、将来の注釈の取り組みを通知するために、結果のパターンについて説明します。元の注釈セットと網羅的な注釈セットの両方を検討することにより、研究者はトレーニング注釈の体系的なパターンがモデリングにどのように影響するかを研究することもできます。

The Open Images Dataset contains approximately 9 million images and is a widely accepted dataset for computer vision research. As is common practice for large datasets, the annotations are not exhaustive, with bounding boxes and attribute labels for only a subset of the classes in each image. In this paper, we present a new set of annotations on a subset of the Open Images dataset called the MIAP (More Inclusive Annotations for People) subset, containing bounding boxes and attributes for all of the people visible in those images. The attributes and labeling methodology for the MIAP subset were designed to enable research into model fairness. In addition, we analyze the original annotation methodology for the person class and its subclasses, discussing the resulting patterns in order to inform future annotation efforts. By considering both the original and exhaustive annotation sets, researchers can also now study how systematic patterns in training annotations affect modeling.

updated: Wed May 05 2021 20:44:56 GMT+0000 (UTC)

published: Wed May 05 2021 20:44:56 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト