Deep Contrastive Learning is Provably (almost) Principal Component Analysis

Yuandong Tian

深い対照学習はおそらく（ほぼ）主成分分析です

損失関数のファミリー（InfoNCEを含む）の下での対照学習（CL）にはゲーム理論の定式化があり、最大プレーヤーはコントラストを最大化する表現を見つけ、最小プレーヤーは同様の表現を持つサンプルのペアに重みを付けます。表現学習を行う最大のプレーヤーが、深層線形ネットワークの主成分分析に還元され、ほとんどすべての極小値がグローバルであり、最適なPCAソリューションを回復することを示します。実験によると、InfoNCEを超えて拡張すると、CIFAR10とSTL-10で同等の（またはそれ以上の）パフォーマンスが得られ、新しい対照的な損失が得られます。さらに、理論的分析を2層ReLUネットワークに拡張し、線形ネットワークとの違いを示し、強力な拡張の下で単一の主要な機能を選択するよりも機能構成が優先されることを証明します。

We show that Contrastive Learning (CL) under a family of loss functions (including InfoNCE) has a game-theoretical formulation, where the max player finds representation to maximize contrastiveness, and the min player puts weights on pairs of samples with similar representation. We show that the max player who does representation learning reduces to Principal Component Analysis for deep linear network, and almost all local minima are global, recovering optimal PCA solutions. Experiments show that the formulation yields comparable (or better) performance on CIFAR10 and STL-10 when extending beyond InfoNCE, yielding novel contrastive losses. Furthermore, we extend our theoretical analysis to 2-layer ReLU networks, showing its difference from linear ones, and proving that feature composition is preferred over picking single dominant feature under strong augmentation.

updated: Mon Apr 11 2022 16:20:58 GMT+0000 (UTC)

published: Sat Jan 29 2022 23:08:34 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト