The Bayesian Method of Tensor Networks

Erdong Guo; David Draper

テンソルネットワークのベイズ法

ベイジアン学習は、データの外部情報（背景情報）と内部情報（トレーニングデータ）を論理的に一貫した方法で推論と予測に組み合わせる強力な学習フレームワークです。ベイズの定理により、外部情報（事前分布）と内部情報（トレーニングデータの尤度）がコヒーレントに結合され、ベイズの定理によって得られた事後分布と事後予測（周辺）分布が、推論と予測に必要な合計情報を要約します。、それぞれ。この論文では、2つの観点からテンソルネットワークのベイズフレームワークを研究します。最初に、テンソルネットワークの重みに事前分布を導入し、事後予測（周辺）分布によって新しい観測値のラベルを予測します。正規化定数計算におけるパラメーター積分の扱いやすさから、事後予測分布をラプラス近似で近似し、テンソルネットワークモデルの事後分布のヘッセ行列の結果近似を取得します。第二に、定常モードのパラメータを推定するために、テンソルネットワークが勾配降下法でより効率的かつ安定して定常経路に収束できる推論プロセスを加速するための安定した初期化トリックを提案します。 MNIST、フィッシングWebサイト、および乳がんのデータセットで作業を検証します。モデルのパラメーターと2次元合成データセットの決定境界を視覚化することにより、ベイジアンテンソルネットワークのベイジアンプロパティを研究します。アプリケーションの目的で、私たちの作業は過剰適合を減らし、通常のテンソルネットワークモデルのパフォーマンスを向上させることができます。

Bayesian learning is a powerful learning framework which combines the external information of the data (background information) with the internal information (training data) in a logically consistent way in inference and prediction. By Bayes rule, the external information (prior distribution) and the internal information (training data likelihood) are combined coherently, and the posterior distribution and the posterior predictive (marginal) distribution obtained by Bayes rule summarize the total information needed in the inference and prediction, respectively. In this paper, we study the Bayesian framework of the Tensor Network from two perspective. First, we introduce the prior distribution to the weights in the Tensor Network and predict the labels of the new observations by the posterior predictive (marginal) distribution. Since the intractability of the parameter integral in the normalization constant computation, we approximate the posterior predictive distribution by Laplace approximation and obtain the out-product approximation of the hessian matrix of the posterior distribution of the Tensor Network model. Second, to estimate the parameters of the stationary mode, we propose a stable initialization trick to accelerate the inference process by which the Tensor Network can converge to the stationary path more efficiently and stably with gradient descent method. We verify our work on the MNIST, Phishing Website and Breast Cancer data set. We study the Bayesian properties of the Bayesian Tensor Network by visualizing the parameters of the model and the decision boundaries in the two dimensional synthetic data set. For a application purpose, our work can reduce the overfitting and improve the performance of normal Tensor Network model.

updated: Fri Jan 01 2021 14:59:15 GMT+0000 (UTC)

published: Fri Jan 01 2021 14:59:15 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト