arXiv論文メモ
新着一覧
cs.RO · 査読状況未確認

推定の不確かさを含めて移動量推定を評価する方法

You Should Be Properly Scoring Your Odometry

Ola Rønning, Usama Saqib, Andrzej Wąsowski

この論文をやさしく読む

ひとことで言うと

移動量推定の誤差だけでなく、推定器自身が報告する不確かさが妥当かも採点する方法を提案した。

何に役立つ?

ロボットなどの位置推定器で、誤差が小さく見えても不確かさの報告が過信になっていないかを調べるのに役立つ。

この研究の面白いところ

正解の軌跡がなくても二つの推定器を比べ、少なくとも一方の過信を検出できるという検定を示した。

どこまで分かった?

要旨の事例評価は地上LiDAR・慣性移動量推定の並進成分と四つのフィルタに関する。全種類の推定器に同じ過信機構があるとは示していない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

移動量推定の性能を評価するとき、推定した軌跡を正解の軌跡と照らし合わせて採点するのが一般的である。しかし、二乗平均平方根誤差などの点推定の指標は、フィルタや平滑化器が既に出力している共分散行列を無視する。共分散が重要な理由は二つある。第一に、推定器自身の不確かさを表し、推定器がその出力をどれほど信じているかを示す。過度に自信のある推定器は、自分が位置を見失ったと報告しない。第二に、共分散は推定の各方向の誤差に重みを付ける。これを無視すると、不確かな方向での大きな誤差が過度に罰せられる。そこで著者らは、点推定の指標の代わりに、推定値と報告された不確かさを併せて採点する厳密に適正なスコアリング規則を使うことを提案する。共分散を報告しない場合は点推定の指標に戻り、報告する場合は共分散の整合性の欠如を診断できる。片側のペア比較検定を用い、正解軌跡がなくても、二つの推定器の少なくとも一方の過信を明らかにできることを示す。このスコアリング規則とペア比較検定は、公開ソフトウェアsmfevalで利用できる。事例として、地上のLiDARと慣性センサーを使う移動量推定について、並進成分の不確かさの品質をsmfevalで評価した。四つのフィルタすべてで過信が見られ、最も極端な例では、誤差がキロメートル単位なのにセンチメートル単位の確かさを報告していた。過信の仕組みを調べると、フィルタがLiDAR測定に実際以上の新情報を与えたと見なしていることに行き着いた。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-22(UTC)
最新改訂
2026-09-22 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

When we evaluate the performance of our odometry, it is common practice to score the estimated track against a ground truth. Unfortunately, scoring uses point metrics, such as the root mean square error, that ignore the covariance matrix which estimators like filters and smoothers already report. Using the covariance matters for two reasons. First, the covariance encodes the estimator's uncertainty, so it tells us whether the estimator trusts its own output. An overconfident estimator will not report itself lost. Second, the covariance weights the error in each direction of the estimate. Without the covariance, an estimator is unduly penalized for a high error in an uncertain direction. Instead of point metrics, we should use strictly proper scoring rules. These rules score the estimate together with its reported uncertainty. Strictly proper scoring rules recover the point metrics when no covariance is reported, and they diagnose covariance inconsistency when covariance is reported. Using a one-sided pairwise test, we show that two estimators can expose overconfidence in at least one of them without a ground truth. Strictly proper scoring rules and our pairwise test are available in our open-source framework smfeval. As a case study, we use smfeval to assess the uncertainty quality of the translational component of ground-based LiDAR-inertial odometry. Across four filters we find overconfidence - the worst case reports centimeter certainty with kilometer error. Knowing the filters are overconfident, we investigate the mechanism. The investigation traces overconfidence to filters crediting LiDAR measurements with more new information than they carry.

arXiv ID: 2609.25900 / 要約の誤りについて