4D-OR: Semantic Scene Graphs for OR Domain Modeling

Ege Özsoy; Evin Pınar Örnek; Ulrich Eck; Tobias Czempiel; Federico Tombari; Nassir Navab

4D-OR：ORドメインモデリングのセマンティックシーングラフ

外科的処置は、さまざまなアクター、デバイス、および相互作用で構成される非常に複雑な手術室（OR）で行われます。今日まで、医学的に訓練された人間の専門家だけが、そのような要求の厳しい環境でのすべてのリンクと相互作用を理解することができます。このホワイトペーパーは、コミュニティをORドメインの自動化された全体論的および意味論的理解とモデリングに一歩近づけることを目的としています。この目標に向けて、初めて、セマンティックシーングラフ（SSG）を使用して、手術シーンを記述および要約することを提案します。シーングラフのノードは、医療スタッフ、患者、医療機器など、部屋内のさまざまなアクターとオブジェクトを表しますが、エッジはそれらの間の関係です。提案された表現の可能性を検証するために、現実的なORシミュレーションセンターで6つのRGB-Dセンサーで記録された10のシミュレートされた人工膝関節全置換術を含む最初の公的に利用可能な4D外科SSGデータセット4D-ORを作成します。 4D-ORには6734フレームが含まれており、SSG、人間とオブジェクトのポーズ、および臨床的役割が豊富に注釈されています。エンドツーエンドのニューラルネットワークベースのSSG生成パイプラインを提案します。成功率は0.75マクロF1で、実際にORで意味論的推論を推測できます。さらに、0.85マクロF1を達成する臨床的役割予測の問題に使用することにより、シーングラフの表現力を示します。コードとデータセットは、承認されると利用できるようになります。

Surgical procedures are conducted in highly complex operating rooms (OR), comprising different actors, devices, and interactions. To date, only medically trained human experts are capable of understanding all the links and interactions in such a demanding environment. This paper aims to bring the community one step closer to automated, holistic and semantic understanding and modeling of OR domain. Towards this goal, for the first time, we propose using semantic scene graphs (SSG) to describe and summarize the surgical scene. The nodes of the scene graphs represent different actors and objects in the room, such as medical staff, patients, and medical equipment, whereas edges are the relationships between them. To validate the possibilities of the proposed representation, we create the first publicly available 4D surgical SSG dataset, 4D-OR, containing ten simulated total knee replacement surgeries recorded with six RGB-D sensors in a realistic OR simulation center. 4D-OR includes 6734 frames and is richly annotated with SSGs, human and object poses, and clinical roles. We propose an end-to-end neural network-based SSG generation pipeline, with a rate of success of 0.75 macro F1, indeed being able to infer semantic reasoning in the OR. We further demonstrate the representation power of our scene graphs by using it for the problem of clinical role prediction, where we achieve 0.85 macro F1. The code and dataset will be made available upon acceptance.

updated: Tue Mar 22 2022 17:59:45 GMT+0000 (UTC)

published: Tue Mar 22 2022 17:59:45 GMT+0000 (UTC)

arXiv

参考文献 (このサイトで利用可能なもの) / References (only if available on this site)

被参照文献 (このサイトで利用可能なものを新しい順に) / Citations (only if available on this site, in order of most recent)

Amazon.co.jpアソシエイト