arXiv論文メモ
新着一覧
cs.LG · 査読状況未確認

候補同士の順位比較で科学実験の対象を選ぶ

Scientific Discovery under Validation Congestion via Multi-Fidelity Pairwise Rankings

Kevin Tirta Wijaya, Alston Lo, Michael Sun, Wojciech Matusik, Vahid Babaei

この論文をやさしく読む

ひとことで言うと

大量の科学的候補から実験対象を選ぶとき、各候補の正確な点数を予測する代わりに、二つを比べてどちらが有望かという判断を使います。

何に役立つ?

実験データが少なく、候補の物理的検証に時間や費用がかかる場面での候補選別が想定されます。評価では創薬ライブラリの探索回数と、新規設計最適化のハイパーボリュームで比較しています。

この研究の面白いところ

すべての比較を高価な専門家に任せず、Fisher情報を使って必要な問い合わせだけ精度の高い判断へ回します。精度と費用の違う人や計算手法を一つの選別過程に組み込む点が特徴です。

どこまで分かった?

約42%と約15%は上位10候補の発見再現率50%に達するまでの反復回数の削減です。約18.8%は別の最適化実験のハイパーボリューム改善で、薬効や臨床成績の改善率ではありません。要旨に実験室での新規発見の実証は記載されていません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

現代の計算手法は、分子、材料、その他の科学的設計候補をかつてない規模で提案できる。その結果、候補は豊富でも、それらを物理的に評価する実験能力が不足する「検証の混雑」が生じている。したがって、新しい科学的設計の発見は、時間と費用のかかる実験へ進める有望な少数の設計を選ぶ、候補の選別にますます依存するようになった。既存の選別法は通常、絶対的なスコアを予測するデータ駆動型の回帰モデルに頼るが、その学習にはそもそも大量の実験データが必要である。しかし、科学的設計の発見は比較によることが多く、有用な選別の手掛かりは絶対的な測定値である必要はない。 本研究では、より収集しやすい専門家による候補対の順位付けを、選別の主な駆動力にすることを提案する。この専門知識は、経験則からエージェントを用いる処理手順、熟練した科学者まで、精度が異なる計算ツールや人の入力から得られる。提案するPRISMSは、専門性の異なる場合も含め、1人以上の専門家からの候補対の順位を用い、大量のデータを必要とする回帰器に頼らず最も有望な候補を特定するフレームワークである。専門家の精度と費用が異なる場合、Fisher情報に基づく基準で、候補対への問い合わせを低精度の順位付け器から高精度のものへ段階的に引き上げる。 固定された創薬ライブラリから設計を選ぶ反復スクリーニングでは、上位10候補の発見再現率50%へ達するのに必要な反復回数を、回帰だけを用いる能動学習より約42%、選択的な問い合わせ格上げをしない順位ベースの手法より約15%減らした。事前定義されたライブラリに制限せず新しい設計を生成する最適化では、ベイズ最適化の比較手法より約18.8%高いハイパーボリュームを達成した。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-10-01(UTC)
最新改訂
2026-10-01 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Modern computational methods can now propose candidate molecules, materials, and other scientific designs at an unprecedented scale, creating a validation congestion where candidates are abundant, but experimental capacity to physically evaluate them remains scarce. Discovering novel scientific designs has therefore become increasingly dependent on curation: selecting a small set of promising designs for slow and costly experiments. Existing curation methods typically rely on data-driven regression models that predict absolute scores, but training these models requires substantial experimental data to begin with. Yet, useful curation signals do not have to take the form of absolute measurements, as scientific design discovery is often comparative in nature. Here, we propose that curation can instead be primarily driven by expert pairwise rankings, which are substantially easier to gather. The expertise can come from computational tools or human input of multiple levels of fidelity, ranging from empirical rules of thumb to agentic workflows and experienced scientists. We introduce PRISMS, a framework that uses pairwise rankings from one or more experts, potentially spanning multiple levels of expertise, to identify the most promising candidates without relying on data-hungry regressors. When experts differ in fidelity and cost, PRISMS escalates pairwise queries from lower- to higher-fidelity rankers based on a Fisher-information criterion. In iterative screening that selects designs from fixed drug discovery libraries, PRISMS achieves 50% top-10 discovery recall in ~42% fewer rounds than regression-only active learning, and in ~15% fewer rounds than the ranking-based method with no selective escalation. In optimization that generates new designs without restriction to a predefined library, PRISMS achieves ~18.8% higher hypervolume than the Bayesian optimization baseline.

arXiv ID: 2610.01827 / 要約の誤りについて