探査地震データ処理の学習手法を比較するSPBench
SPBench: A Multi-Task Evaluation Benchmark for Exploration Seismic Processing
この論文をやさしく読む
ひとことで言うと
地震データ処理の学習手法を、共通のデータ・設定・評価方法で比較するベンチマーク。
何に役立つ?
探査地震データ向け手法の性能差が、モデルそのものか実験条件かを調べるのに役立つ。
この研究の面白いところ
合成データの順位が実地データの順位を安定して予測しないことや、全体スコアに隠れる成分ごとの差を示す。
どこまで分かった?
順位や指標の一致に関する結論は収録した6課題、10データセット、43条件での比較に基づく。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
探査地震データの処理は地下の画像化と資源探査の基礎だが、学習を用いた手法は研究間で比較しにくい。368本の論文を調査すると、非公開または再現が難しいデータセットへの依存が広く見られ、公開コードがあるのは25本だけだった。そのため、報告された改善がモデル設計によるものか実験条件によるものかが分かりにくい。本研究は地震データ処理のベンチマーク SPBench を導入する。対象はランダムノイズの抑制、トレース補間、表面波の抑制、多重反射の抑制、混合データの分離、初動到達時刻の抽出の6課題である。10データセット、43通りの標準化した条件の下で、教師あり学習の24手法を再現し、データ、実装、設定、評価コード、結果を公開する。全体の得点やトレースごとの初動誤差を補うため、再構成を信号成分ごとに評価する SCoRE と、初動抽出用の参照値を必要としない尾根の曲率スコア RC_norm を導入する。分析では、各条件内でモデルを学習した場合、合成データでの順位は実地データでの順位を安定して予測せず、一致の程度は課題に依存した。劣化が強まると、ランダムに近い妨害より、まとまった表面波による妨害の方が順位の入れ替わりが大きい。評価条件では、尾根スコアは平均絶対誤差に基づく順位と一致し、3つの実地調査を通じた Kendall 相関の平均は0.881だった。SCoRE は全体スコアでは隠れる周波数とエネルギーに依存した差を示した。SPBench は学習を用いた地震データ処理を再現可能な形で比較し、データ条件、劣化の強さ、評価基準によって手法の相対的な利点がどう変わるかを示す。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-24(UTC)
- 最新改訂
- 2026-09-24 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-24 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Exploration seismic processing underpins subsurface imaging and resource exploration, but learning-based methods remain difficult to compare across studies. Our survey of 368 papers finds widespread reliance on private or difficult-to-reproduce datasets, with only 25 providing public code. This obscures whether reported gains arise from model design or experimental settings. We introduce the Seismic Processing Benchmark (SPBench), covering six tasks: random noise attenuation, trace interpolation, ground-roll suppression, multiple suppression, deblending, and first-arrival picking. We reproduce 24 supervised methods on 10 datasets under 43 standardized settings and release datasets, implementations, configurations, evaluation scripts, and results. To complement global scores and per-trace pick errors, we introduce signal-component-resolved evaluation (SCoRE) for reconstruction and a reference-free ridge-curvature score (RC_norm) for first-arrival picking. Our analyses show that synthetic rankings do not reliably predict field rankings, with task-dependent agreement when models train within each setting. As degradation strengthens, rankings reorder more under coherent ground roll than under random-like interference. The ridge score agrees with MAE-based model rankings in the evaluated settings, with a mean Kendall correlation of 0.881 across three field surveys, while SCoRE reveals frequency- and energy-dependent differences hidden by global scores. SPBench provides a reproducible basis for comparing learning-based seismic processing methods and characterizes how their relative advantages vary across data settings, degradation strengths, and evaluation criteria.
arXiv ID: 2609.28925 / 要約の誤りについて