arXiv論文メモ
新着一覧
cs.CV · 査読状況未確認

献体の内視鏡映像を生体に近い見た目へリアルタイム変換

EndoLive: Real-Time Style Transfer for Endoscopic Endonasal Skull Base Surgical Video

Griffin Hurt, Calvin Brinkman

この論文をやさしく読む

ひとことで言うと

献体を使う手術練習の内視鏡映像を、生体の組織に近い見た目へその場で変換する映像処理の研究です。

何に役立つ?

考えられる用途は、献体による訓練と実際の手術の見た目の違いを小さくすることです。要旨で検証したのは映像変換の速度と構造の整合性です。

この研究の面白いところ

写実的な手術画像変換のモデルと、速い変換ができるモデルを組み合わせ、対になっていない献体・生体画像で学習しています。

どこまで分かった?

外科医の学習効果や患者の転帰が改善したという結果は要旨にありません。映像の見た目を変えるもので、組織の物理的な性質や実際の出血を再現したという主張ではありません。具体的な速度値も示されていません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

内視鏡下経鼻頭蓋底手術のように、重要な解剖学的構造の周囲を扱う複雑な手術では、生きた患者に手術を行うことが認められる前に、外科医は十分な練習と訓練を行う必要がある。通常、この訓練には献体標本を用いる。生きた人と同じ重要な構造を含んでいるためである。しかし、献体は生体患者の完全な代わりではない。死後に保存された組織の色は生体とはまったく異なり、複雑で高価な送液装置がなければ、同じようには出血しない。そのため、手術を複雑にする重要な解剖学的構造を見分けることは、生体での手術と献体での練習とで大きく異なる場合がある。 本論文では、献体の内視鏡映像と生体の内視鏡映像の間でリアルタイムのスタイル変換を行う枠組みEndoLiveを提案する。本手法は、手術用途の写実的なスタイル変換を行うConStructS GANモデルと、複雑な変換を学習してリアルタイムに実行できるHyPER-GANモデルを組み合わせる。対応付けのない献体と生体の内視鏡画像でEndoLiveを学習し、学習済みモデルを献体映像とさまざまな機器を用いて試験する。 実験結果は、EndoLiveが重要な解剖学的構造の意味的な整合性を保ちながら、リアルタイムに必要な最低速度を十分に上回る速さで、献体から生体への映像変換を行えることを示す。ソースコードはhttps://github.com/griffhurt/endoliveで公開されている。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-10-01(UTC)
最新改訂
2026-10-01 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Complex surgical procedures around critical anatomy, such as the endoscopic endonasal skull base surgery, requires significant practice and training on the part of the surgeon before they are allowed to perform the operation on a live patient. This training in typically done in cadaveric specimens, due to them containing the same critical structures as a living human. However, cadavers are not a perfect 1-to-1 substitute for a living patient. The dead and preserved tissues of a cadaver are colored completely differently than a living human, and -- without complex and expensive pumping systems -- do not bleed in the same way. As a result, identifying the critical pieces of anatomy that make this procedure so complex can be quite different in a live case than in a surgeon's cadaveric practice. This paper presents EndoLive, a framework for real-time style transfer between cadaveric endoscopic video and living human endoscopic video. Our method combines the ConStructS GAN model for realistic style transfer for surgical applications, with the HyPER-GAN model that can learn complex translations and perform them in real-time. We train EndoLive on unpaired cadaveric and live images taken from an endoscope, and test the trained model with cadaveric video, on a variety of devices. Experimental results demonstrate that EndoLive can perform cadaveric-to-live translation at speeds well above the minimum necessary for real-time, while maintaining semantic consistency of critical anatomical structures. Our source code is available at https://github.com/griffhurt/endolive.

arXiv ID: 2610.01956 / 要約の誤りについて