Towards the Generalization of Contrastive Self-Supervised Learning

Weiran Huang; Mingyang Yi; Xuyang Zhao

対照的な自己監視学習の一般化に向けて

最近、自己監視学習は、トレーニングにラベルのないデータしか必要としないため、大きな注目を集めています。対照学習は、自己管理学習の一般的なアプローチであり、実際には経験的にうまく機能します。ただし、下流のタスクでの一般化能力の理論的理解は十分に研究されていません。この目的のために、対照的な自己監視の事前訓練されたモデルがどのように下流のタスクに一般化するかについての理論的な説明を提示します。具体的には、自己監視モデルが、クラスの中心と密接にクラスター化されたクラス内サンプルを区別する特徴空間に入力データを埋め込む場合、下流の分類タスクで一般化能力を持っていることを定量的に示します。上記の結論を踏まえて、2つの標準的な対照的な自己監視方式であるSimCLRとBarlowTwinsをさらに調査します。前述の特徴空間が任意の方法で取得できることを証明し、下流の分類タスクの一般化におけるそれらの成功を説明します。最後に、私たちの理論的発見を検証するために、さまざまな実験も行われます。

Recently, self-supervised learning has attracted great attention since it only requires unlabeled data for training. Contrastive learning is a popular approach for self-supervised learning and empirically performs well in practice. However, the theoretical understanding of its generalization ability on downstream tasks is not well studied. To this end, we present a theoretical explanation of how contrastive self-supervised pre-trained models generalize to downstream tasks. Concretely, we quantitatively show that the self-supervised model has generalization ability on downstream classification tasks if it embeds input data into a feature space with distinguishing centers of classes and closely clustered intra-class samples. With the above conclusion, we further explore SimCLR and Barlow Twins, which are two canonical contrastive self-supervised methods. We prove that the aforementioned feature space can be obtained via any of the methods, and thus explain their success on the generalization on downstream classification tasks. Finally, various experiments are also conducted to verify our theoretical findings.

updated: Mon Nov 01 2021 07:39:38 GMT+0000 (UTC)

published: Mon Nov 01 2021 07:39:38 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト