人型ロボットの道具選びから移動・作業までを評価
HumanoidToolBench: Benchmarking Humanoid Tool Use from Selection to Mobile Execution
この論文をやさしく読む
ひとことで言うと
人型ロボットについて、道具を正しく選ぶ能力と、その道具で移動しながら作業を終える能力をまとめて測る評価セットです。
何に役立つ?
道具使用のどの段階でロボットが失敗するかを比較するために役立ちます。シミュレーションと実機の実演データも提供しています。
この研究の面白いところ
道具の選択が正しくても作業を完了できるとは限らない点を評価で切り分けています。未見の道具や無関係な指示への反応も調べています。
どこまで分かった?
評価範囲は18課題で、方策数はシミュレーション7種類、実機3種類です。未見道具と無関係な指示に関する追加検証はGR00T N1.7に絞られており、全方策に共通する結果とは要旨から言えません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
ロボットのハードウェアと学習手法が進歩するにつれ、人型ロボットには、自身の身体的限界を超えた作業を行うために道具が必要になる。道具をうまく使うには、適切な道具を選び、物体操作と、必要な場合には移動を協調させて作業を完了する必要がある。既存のベンチマークは、人型ロボット上でこれらの能力をまとめて評価していない。 本研究では、3種類の場面、3段階の実行レベル、2種類の道具セットモードにまたがる18課題のベンチマークHumanoidToolBenchを導入する。併せて、シミュレーションと実機のUnitree G1で収集した3,100件の実演データセットToolBookを提供する。シミュレーションで7種類、実機で3種類の方策を評価すると、適切な道具の選択と作業の完了の間に大きな隔たりが見られる。GR00T N1.7に絞った検証では、未見の道具で選択精度が低下し、無関係な指示のもとでも作業の実行が続くことが示される。コードとデータは https://snu-pi.github.io/HumanoidToolBench/ で公開している。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-10-01(UTC)
- 最新改訂
- 2026-10-01 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-10-01 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
As robotic hardware and learning methods advance, humanoids need tools to perform tasks beyond their inherent physical limits. Successful tool use requires selecting a suitable tool and coordinating manipulation and, when needed, locomotion to complete the task. Existing benchmarks do not jointly evaluate these capabilities on a humanoid. We introduce HumanoidToolBench, an 18-task benchmark spanning three scenarios, three execution levels, and two tool-set modes, together with ToolBook, a dataset of 3.1k demonstrations collected in simulation and on a real Unitree G1. Evaluation of seven policies in simulation and three on the real robot reveals substantial gaps between selecting a suitable tool and completing the task. Focused GR00T N1.7 probes show reduced selection accuracy on unseen tools and continued task execution under unrelated instructions. Code and data are available at https://snu-pi.github.io/HumanoidToolBench/.
著者のコメント
9 pages, 7 figures
arXiv ID: 2610.02089 / 要約の誤りについて