Self-Supervised Scale Recovery for Monocular Depth and Egomotion Estimation

Brandon Wagstaff; Jonathan Kelly

単眼深度と自我運動推定のための自己教師ありスケール回復

単眼画像を使用して深度と自我運動ニューラルネットワークを共同でトレーニングするための自己教師あり損失の定式化は十分に研究されており、最先端の精度が実証されています。ただし、このアプローチの主な制限の1つは、深度と自我運動の推定値が未知のスケールまでしか決定されないことです。この論文では、既知のカメラの高さと推定されたカメラの高さの間の一貫性を強制し、メトリック（スケーリングされた）深度と自我運動の予測を生成する、新しいスケール回復損失を提示します。提案された方法が、より多くの情報を必要とする他のスケール回復手法と競合することを示します。さらに、他のスケール解決アプローチではそうすることができないのに対し、私たちの方法は新しい環境内でのネットワークの再トレーニングを容易にすることを示しています。特に、私たちのエゴモーションネットワークは、テスト時にのみスケールを回復する同様の方法よりも正確な推定値を生成できます。

The self-supervised loss formulation for jointly training depth and egomotion neural networks with monocular images is well studied and has demonstrated state-of-the-art accuracy. One of the main limitations of this approach, however, is that the depth and egomotion estimates are only determined up to an unknown scale. In this paper, we present a novel scale recovery loss that enforces consistency between a known camera height and the estimated camera height, generating metric (scaled) depth and egomotion predictions. We show that our proposed method is competitive with other scale recovery techniques that require more information. Further, we demonstrate that our method facilitates network retraining within new environments, whereas other scale-resolving approaches are incapable of doing so. Notably, our egomotion network is able to produce more accurate estimates than a similar method which recovers scale at test time only.

updated: Fri Jul 16 2021 14:57:04 GMT+0000 (UTC)

published: Tue Sep 08 2020 14:30:21 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト