arXiv論文メモ
新着一覧
cs.RO · 査読状況未確認

腹腔鏡器具から集めた実演で虫垂切除を学習

From Instrument-Mounted Demonstrations to In-Vivo Execution: Learning Bimanual Laparoscopic Appendectomy Without Robot-Collected Demonstrations

Dongho Yee, Juahn Oh, Jinseok Lee, Jiyul Lee, Yechan Seo, Seong Jeong, Minsung Kim, Seonho Shim, Younghoon Noh, Hyuk Choi, Youngbin Kong, Kyu Eun Lee, Hyoun-Joong Kong

この論文をやさしく読む

ひとことで言うと

外科医の腹腔鏡器具に付けたセンサーで手術中の動きを記録し、その実演からロボットの虫垂切除動作を学習させた。

何に役立つ?

手持ちの手術器具から学習用の動作データを集める方法の研究に役立つ可能性がある。人の手術で使えると示した結果ではない。

この研究の面白いところ

ロボットに実演をさせず、器具側のセンサーで849件の生体内実演を集めた。別の生きたウサギ4匹で試し、外科医が手術段階を選ぶ条件で3匹の虫垂切除を完了した。

どこまで分かった?

評価はウサギで行われ、実行時には外科医が手術段階を選んだ。4匹中1匹では完了しておらず、人に対する安全性や有効性は要旨に記載されていない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

低侵襲手術の多くは今も手持ちの腹腔鏡器具で行われており、手術が終わると外科医の器具の動きは失われ、内視鏡映像だけが残る。本論文は、この動きを手術室で記録し、手術ロボットの方策学習に用いる一連の工程を提示し、生きた動物で検証する。標準的な腹腔鏡器具のシャフトに取り付ける状態記録装置を導入する。慣性センサー、飛行時間センサー、ホールセンサーから器具の位置・姿勢と鉗子の開閉状態を復元し、外部カメラや追跡装置は使わない。データ処理工程では、各センサーの時間遅れをロボットの基準値と比較して測定し、観測と行動の組を作る前に各チャンネルを時間的に揃える。これらの実演を使い、微調整したDINOv3を基盤にした拡散方策を訓練する。方策の設計は、体外に取り出したウサギの虫垂の深度マップから再構成した物理シミュレーターで、閉ループの動作を繰り返して選ぶ。 その後、生きたウサギ4匹から得た849件の実演で方策を再訓練し、別の生きたウサギ4匹で電気手術器を使用可能な状態にして実行した。外科医が手術の段階を選ぶ条件で、方策は4匹中3匹の虫垂切除を完了した。結果は、外科医自身の器具から記録した実演によって、生体内で使う両手操作の手術方策を訓練、選択、実行できることを示す。ロボットはセンサー校正の時間基準と実行器としてのみ用い、実演は収集していない。両方の実演データ集は、今後の手術ロボット学習研究を支えるため公開されている。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-22(UTC)
最新改訂
2026-09-22 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Most minimally invasive surgery is still performed with hand-held laparoscopic instruments, and the surgeon's instrument kinematics are lost when the operation ends; only the endoscope video is kept. This paper presents an end-to-end pipeline that captures this motion in the operating room and uses it to train a surgical robot policy, validated on live animals. We introduce a surgical instrument-state logger that mounts on the shaft of a standard laparoscopic instrument and recovers its pose and jaw state from an inertial sensor, a time-of-flight sensor and a Hall sensor, with no external camera or tracker. A data pipeline measures the latency of every sensor channel against a robot ground truth and aligns the channels before forming observation-action pairs. On these demonstrations we train a diffusion policy with a fine-tuned DINOv3 backbone, selecting its design by closed-loop rollouts in a physics simulator reconstructed from depth maps of an ex-vivo rabbit appendix. The policy is then retrained on 849 in-vivo demonstrations from four live rabbits and deployed on four additional live rabbits with electrosurgery armed. With the surgeon selecting the surgical phase, the policy completed the appendectomy in three of the four animals. The results show that demonstrations recorded from a surgeon's own instruments are sufficient to train, select and deploy a bimanual surgical policy in vivo. The robot serves only as the timing reference for sensor calibration and as the executor, and collects no demonstrations. Both demonstration corpora are released to support future surgical robot learning research.

著者のコメント

Submitted to IEEE ICRA 2027

arXiv ID: 2609.25625 / 要約の誤りについて