Lifelong 3D Object Recognition and Grasp Synthesis Using Dual Memory Recurrent Self-Organization Networks

Krishnakumar Santhakumar; Hamidreza Kasaei

デュアルメモリ再帰的自己組織化ネットワークを使用した生涯にわたる3Dオブジェクト認識と把握合成

人間は、非定常的かつ連続的な条件下で以前に得られた知識を忘れることなく、生涯にわたる設定で新しいオブジェクトを認識して操作することを学びます。自律システムでは、エージェントは、新しいオブジェクトカテゴリを継続的に学習し、新しい環境に適応するために、同様の動作を軽減する必要もあります。ほとんどの従来のディープニューラルネットワークでは、これは壊滅的な忘却の問題のために不可能であり、新しく得られた知識が既存の表現を上書きします。さらに、ほとんどの最先端モデルは、オブジェクトの認識または把握予測のいずれかで優れていますが、両方のタスクは視覚的な入力を使用します。両方のタスクに取り組むための結合されたアーキテクチャは非常に限られています。この論文では、動的に成長するデュアルメモリリカレントニューラルネットワーク（GDM）と、オブジェクトの認識と把握に同時に取り組むためのオートエンコーダで構成されるハイブリッドモデルアーキテクチャを提案しました。オートエンコーダネットワークは、GDM学習の入力として機能する、特定のオブジェクトのコンパクトな表現を抽出する役割を果たし、ピクセル単位の対蹠把握構成を予測する責任があります。 GDMパーツは、インスタンスレベルとカテゴリレベルの両方でオブジェクトを認識するように設計されています。エピソード記憶が外部感覚情報の不在下で神経活性化軌道を定期的に再生する、内因性記憶再生を使用して壊滅的な忘却の問題に対処します。生涯の設定で提案されたモデルを広範囲に評価するために、シーケンシャル3Dオブジェクトデータセットがないため、合成データセットを生成します。実験結果は、提案されたモデルが継続的な学習シナリオでオブジェクト表現と把握の両方を同時に学習できることを示しました。

Humans learn to recognize and manipulate new objects in lifelong settings without forgetting the previously gained knowledge under non-stationary and sequential conditions. In autonomous systems, the agents also need to mitigate similar behavior to continually learn the new object categories and adapt to new environments. In most conventional deep neural networks, this is not possible due to the problem of catastrophic forgetting, where the newly gained knowledge overwrites existing representations. Furthermore, most state-of-the-art models excel either in recognizing the objects or in grasp prediction, while both tasks use visual input. The combined architecture to tackle both tasks is very limited. In this paper, we proposed a hybrid model architecture consists of a dynamically growing dual-memory recurrent neural network (GDM) and an autoencoder to tackle object recognition and grasping simultaneously. The autoencoder network is responsible to extract a compact representation for a given object, which serves as input for the GDM learning, and is responsible to predict pixel-wise antipodal grasp configurations. The GDM part is designed to recognize the object in both instances and categories levels. We address the problem of catastrophic forgetting using the intrinsic memory replay, where the episodic memory periodically replays the neural activation trajectories in the absence of external sensory information. To extensively evaluate the proposed model in a lifelong setting, we generate a synthetic dataset due to lack of sequential 3D objects dataset. Experiment results demonstrated that the proposed model can learn both object representation and grasping simultaneously in continual learning scenarios.

updated: Thu Sep 23 2021 11:14:13 GMT+0000 (UTC)

published: Thu Sep 23 2021 11:14:13 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト