arXiv論文メモ
新着一覧
astro-ph.GA · 査読状況未確認

複数AIの一般化平均で重力レンズ候補の誤検出を削減

Reducing False Positives in Strong-Lens Searches with Generalized-Mean Consensus of Machine-Learning Ensembles in the Kilo-Degree Survey

Ziqi Li, Rui Li, Xu Huang, Hui Li, Pufan Liu, Liang Gao, Crescenzo Tortora, Nicola N. Napolitano, Xiaoyue Cao, Ran Li, Liqing Chen, Kang Jiao, Valerio Busillo and Yue Dong

この論文をやさしく読む

ひとことで言うと

複数の画像分類AIの出力を一般化平均でまとめ、重力レンズらしく見えるだけの候補を減らして、人が確認する量を抑えています。

何に役立つ?

広域天文サーベイで、有望な候補をできるだけ落とさずに追加観測や目視確認の対象を絞るために役立ちます。

この研究の面白いところ

シミュレーションでは最良の単独モデルを上回らないのに、実データでは組合せが有効です。完全性90%を保った比較で偽陽性率を0.007まで減らしています。

どこまで分かった?

新たな170個は高品質な候補であり、すべて重力レンズと確定したという意味ではありません。50%・70%の削減は抽出候補数の比較で、真のレンズ数を減らした割合ではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

背景:広域サーベイでの主要な課題は、分類器の感度だけでなく、圧倒的に多い偽陽性である。数百万から数十億の銀河の中で強い重力レンズを探すと多数の混入天体が生じ、追加の目視確認と、統計的に有用なレンズ標本の構築の障害になる。目的:複数の分類器を組み合わせ、KiDS DR4の強い重力レンズ候補選択の純度を改善する。既知の候補に対する高い完全性を保ちつつ、非レンズの割合を大きく減らすことを目指す。 方法:Li ResNet+、Swin Transformerの各種モデル、Swin-MLP、DemiLensNetを含む、畳み込み型、Transformer型、およびハイブリッド型の分類器を学習した。確率出力を、平均と一般化平均による合意を用いてスコアの段階で組み合わせた。KiDSに似せたレンズ画像のシミュレーションでモデルを試験した後、非レンズ標本に実際のKiDS DR4のレンズ候補を加えたデータで評価した。 結果:シミュレーションのテスト集合では、アンサンブルは最良の単独モデルを上回らない。実データを混合したKiDSテスト集合では、完全性90%における偽陽性率が、良い方の二つの単独モデルの範囲0.016~0.020から、7モデルの算術平均アンサンブルでは0.011へ低下する。一般化平均は、これをさらに0.007へ下げる。同じ完全性90%でLRGおよびBGの全標本へ適用すると、最良の単独モデルに比べて、一般化平均による抽出候補数はLRGで約50%、BGで約70%減少する。目視確認の後、新たな高品質候補170個(Class Aが24個、Class Bが146個)と、Class C候補1,706個を得た。 結論:これらの結果は、機械学習アンサンブルの一般化平均による合意を使う戦略が、有望な強い重力レンズ候補の高い回収率を保ちながら、目視確認の負担を減らす実用的な道筋となることを示す。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-21(UTC)
最新改訂
2026-09-21 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Context. In wide-field surveys, the main challenge is not just classifier sensitivity, but the overwhelming number of false positives. Searching for strong lenses among millions to bilions of galaxies produces many contaminants, making the bottleneck for follow-up inspection and building statistically useful lens samples. Aims. We aim to improve the purity of strong-lens candidate selection in KiDS DR4 by combining several classifiers. The objective is to retain high completeness for known candidates while substantially reducing the fraction of non-lenses. Methods. We trained convolutional, Transformer-based, and hybrid classifiers, including Li ResNet+, Swin Transformer variants, Swin-MLP, and DemiLensNet. Their probabilistic outputs were combined at score level using averaging and a generalized mean consensus. The models were tested on simulated KiDS-like lens images and then evaluated on real KiDS DR4 lens candidates embedded in a non-lens sample. Results. On the simulated test set, ensembles show no advantage over the best single models. On the mixed real KiDS test set, the arithmetic mean reduces the false-positive rate at 90% completeness from 0.016-0.020 (the range spanned by the two best individual models) to 0.011 for the seven-model ensemble. The generalized mean reduces it further, to 0.007. Applied to the full LRG and BG samples at the same 90% completeness level, the generalized mean reduces returned candidates by roughly 50% for LRGs and 70% for BGs, relative to the best single model. After visual inspection, we obtain 170 new high-quality candidates (24 Class A and 146 Class B), together with 1706 Class C candidates. Conclusions. Our results demonstrate that the generalized mean consensus of an ML ensemble strategy provides a practical route to reducing the visual inspection workload while preserving a high recovery rate of promising strong-lens candidates.

著者のコメント

16 pages, 4 tables, 8 figures. Submitted to A&A

arXiv ID: 2609.24891 / 要約の誤りについて