Meta-RangeSeg: LiDAR Sequence Semantic Segmentation Using Multiple Feature Aggregation

Song Wang; Jianke Zhu; Ruixiang Zhang

Meta-RangeSeg：複数の機能の集約を使用したLiDARシーケンスのセマンティックセグメンテーション

LiDARセンサーは、自動運転車やインテリジェントロボットの知覚システムに不可欠です。実際のアプリケーションでリアルタイムの要件を満たすには、LiDARスキャンを効率的にセグメント化する必要があります。以前のアプローチのほとんどは、3Dポイントクラウドを2D球面範囲画像に直接投影するため、画像のセグメンテーションに効率的な2D畳み込み演算を利用できます。有望な結果は得られましたが、球形投影では近隣情報が十分に保存されていません。さらに、時間情報は、シングルスキャンセグメンテーションタスクでは考慮されません。これらの問題に取り組むために、Meta-RangeSegという名前のLiDARシーケンスのセマンティックセグメンテーションへの新しいアプローチを提案します。ここでは、新しい範囲残差画像表現を導入して時空間情報をキャプチャします。具体的には、メタカーネルを使用してメタフィーチャを抽出します。これにより、2D範囲の画像座標入力とデカルト座標出力の間の不整合が減少します。効率的なU-Netバックボーンを使用して、マルチスケール機能を取得します。さらに、Feature Aggregation Module（FAM）は、メタ機能とマルチスケール機能を集約します。これにより、範囲チャネルの役割が強化される傾向があります。 LiDARセマンティックセグメンテーションの事実上のデータセットであるSemanticKITTIのパフォーマンス評価のために、広範な実験を実施しました。有望な結果は、提案されたMeta-RangeSegメソッドが既存のアプローチよりも効率的かつ効果的であることを示しています。私たちの完全な実装は、https：//github.com/songw-zju/Meta-RangeSegで公開されています。

LiDAR sensor is essential to the perception system in autonomous vehicles and intelligent robots. To fulfill the real-time requirements in real-world applications, it is necessary to efficiently segment the LiDAR scans. Most of previous approaches directly project 3D point cloud onto the 2D spherical range image so that they can make use of the efficient 2D convolutional operations for image segmentation. Although having achieved the encouraging results, the neighborhood information is not well-preserved in the spherical projection. Moreover, the temporal information is not taken into consideration in the single scan segmentation task. To tackle these problems, we propose a novel approach to semantic segmentation for LiDAR sequences named Meta-RangeSeg, where a novel range residual image representation is introduced to capture the spatial-temporal information. Specifically, Meta-Kernel is employed to extract the meta features, which reduces the inconsistency between the 2D range image coordinates input and Cartesian coordinates output. An efficient U-Net backbone is used to obtain the multi-scale features. Furthermore, Feature Aggregation Module (FAM) aggregates the meta features and multi-scale features, which tends to strengthen the role of range channel. We have conducted extensive experiments for performance evaluation on SemanticKITTI, which is the de-facto dataset for LiDAR semantic segmentation. The promising results show that our proposed Meta-RangeSeg method is more efficient and effective than the existing approaches. Our full implementation is publicly available at https://github.com/songw-zju/Meta-RangeSeg .

updated: Thu Mar 03 2022 09:01:17 GMT+0000 (UTC)

published: Sun Feb 27 2022 14:46:13 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト