DeepSocNav: Social Navigation by Imitating Human Behaviors

Juan Pablo de Vicente; Alvaro Soto

DeepSocNav：人間の行動を模倣することによるソーシャルナビゲーション

社会的行動を訓練するための現在のデータセットは、通常、鳥瞰図から視覚データをキャプチャする監視アプリケーションから借用されています。これは、シーンの一人称ビューを通じてキャプチャできる貴重な関係と視覚的な手がかりを脇に置きます。この作業では、Unityなどの現在のゲームエンジンの能力を活用して、既存の鳥瞰図データセットを一人称ビュー、特に深度ビューに変換する戦略を提案します。この戦略を使用して、ソーシャルナビゲーションモデルの事前トレーニングに使用できる大量の合成データを生成できます。アイデアをテストするために、提案されたアプローチを利用して合成データを生成するディープラーニングベースのモデルであるDeepSocNavを紹介します。さらに、DeepSocNavには、補助タスクとして含まれている自己監視戦略が含まれています。これは、エージェントが直面する次の深度フレームを予測することで構成されます。私たちの実験は、ソーシャルナビゲーションスコアの点で関連するベースラインを上回ることができる提案されたモデルの利点を示しています。

Current datasets to train social behaviors are usually borrowed from surveillance applications that capture visual data from a bird's-eye perspective. This leaves aside precious relationships and visual cues that could be captured through a first-person view of a scene. In this work, we propose a strategy to exploit the power of current game engines, such as Unity, to transform pre-existing bird's-eye view datasets into a first-person view, in particular, a depth view. Using this strategy, we are able to generate large volumes of synthetic data that can be used to pre-train a social navigation model. To test our ideas, we present DeepSocNav, a deep learning based model that takes advantage of the proposed approach to generate synthetic data. Furthermore, DeepSocNav includes a self-supervised strategy that is included as an auxiliary task. This consists of predicting the next depth frame that the agent will face. Our experiments show the benefits of the proposed model that is able to outperform relevant baselines in terms of social navigation scores.

updated: Mon Jul 19 2021 21:51:06 GMT+0000 (UTC)

published: Mon Jul 19 2021 21:51:06 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト