Know Your Surroundings: Panoramic Multi-Object Tracking by Multimodality Collaboration

Yuhang He; Wentao Yu; Jie Han; Xing Wei; Xiaopeng Hong; Yihong Gong

周囲を知る: マルチモダリティコラボレーションによるパノラママルチオブジェクトトラッキング

この論文では、自動運転とロボットナビゲーションの複数オブジェクト追跡 (MOT) 問題に焦点を当てます。既存の MOT メソッドのほとんどは、単一の RGB カメラを使用して複数のオブジェクトを追跡します。これは、カメラの視野が生じやすく、複雑なシナリオでは背景が乱雑で光の状態が悪いために追跡に失敗する傾向があります。これらの課題に対処するために、2D パノラマ画像と 3D 点群の両方を入力として受け取り、マルチモダリティデータを使用してターゲットの軌道を推測するマルチモダリティ PAnoramic マルチオブジェクトトラッキングフレームワーク (MMPAT) を提案します。提案された方法には、パノラマ画像検出モジュール、マルチモダリティデータ融合モジュール、データアソシエーションモジュール、および軌跡推定モデルの 4 つの主要なモジュールが含まれています。 JRDB データセットで提案された方法を評価すると、MMPAT は検出タスクと追跡タスクの両方で最高のパフォーマンスを達成し、最先端の方法を大幅に上回っています (AP と MOTA に関して 15.7 と 8.5 の改善)。、それぞれ)。

In this paper, we focus on the multi-object tracking (MOT) problem of automatic driving and robot navigation. Most existing MOT methods track multiple objects using a singular RGB camera, which are prone to camera field-of-view and suffer tracking failures in complex scenarios due to background clutters and poor light conditions. To meet these challenges, we propose a MultiModality PAnoramic multi-object Tracking framework (MMPAT), which takes both 2D panorama images and 3D point clouds as input and then infers target trajectories using the multimodality data. The proposed method contains four major modules, a panorama image detection module, a multimodality data fusion module, a data association module and a trajectory inference model. We evaluate the proposed method on the JRDB dataset, where the MMPAT achieves the top performance in both the detection and tracking tasks and significantly outperforms state-of-the-art methods by a large margin (15.7 and 8.5 improvement in terms of AP and MOTA, respectively).

updated: Mon May 31 2021 03:16:38 GMT+0000 (UTC)

published: Mon May 31 2021 03:16:38 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト