Pedestrian Trajectory Prediction using Context-Augmented Transformer Networks

Khaled Saleh

コンテキスト拡張トランスフォーマーネットワークを使用した歩行者軌道予測

共有都市交通環境における歩行者の軌道を予測することは、自動運転車（AV）の開発が直面している困難な問題の1つと見なされています。文献では、この問題はリカレントニューラルネットワーク（RNN）を使用して対処されることがよくあります。歩行者の動きの軌跡の時間依存性をキャプチャするRNNの強力な機能にもかかわらず、より長いシーケンシャルデータを処理する場合は挑戦されると主張されました。したがって、この作業では、多くのシーケンシャルベースのタスクでRNNよりも効率的でパフォーマンスが優れていることが最近示されたトランスネットワークに基づくフレームワークを紹介します。歩行者のロバストな軌道予測を提供するために、フレームワークへの入力として、過去の位置情報、エージェントの相互作用情報、およびシーンの物理的セマンティクス情報の融合に依存しました。共有都市交通環境における歩行者の2つの実際のデータセットでフレームワークを評価し、短期および長期の両方の予測範囲で比較されたベースラインアプローチを上回りました。

Forecasting the trajectory of pedestrians in shared urban traffic environments is still considered one of the challenging problems facing the development of autonomous vehicles (AVs). In the literature, this problem is often tackled using recurrent neural networks (RNNs). Despite the powerful capabilities of RNNs in capturing the temporal dependency in the pedestrians' motion trajectories, they were argued to be challenged when dealing with longer sequential data. Thus, in this work, we are introducing a framework based on the transformer networks that were shown recently to be more efficient and outperformed RNNs in many sequential-based tasks. We relied on a fusion of the past positional information, agent interactions information and scene physical semantics information as an input to our framework in order to provide a robust trajectory prediction of pedestrians. We have evaluated our framework on two real-life datasets of pedestrians in shared urban traffic environments and it has outperformed the compared baseline approaches in both short-term and long-term prediction horizons.

updated: Sun Aug 08 2021 15:16:14 GMT+0000 (UTC)

published: Thu Dec 03 2020 08:43:12 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト