arXiv論文メモ
新着一覧
cs.LG / cond-mat.mtrl-sci / cs.AI · 査読状況未確認

高分子物性予測を比べる公開ベンチマークPolyBench26

An open benchmark for machine learning-based polymer property prediction

Robert W. Learsch, Nicholas Liesen, Daniel S. Levine, Anna M. Hiszpanski, Evan R. Antoniuk

この論文をやさしく読む

ひとことで言うと

さまざまな構造の高分子について、物性予測モデルを同じ条件で比較できる公開データセット。

何に役立つ?

高分子設計向けの機械学習モデルを選び、学習データ量や未知の構造への対応を評価するのに役立つ。

この研究の面白いところ

約25万データ点、8物性、4評価課題を用意し、実験・理論計算・分子動力学のデータを含める。

どこまで分かった?

グラフモデルの優位性は収録された物性・構造と比較手法についての結果。すべての高分子物性で最良とは要旨にない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

高分子の物性予測では、機械学習手法を厳密に比較できる公開・標準化されたベンチマークが不足している。既存の資料は、単独重合体など、高分子の構造の一部しか扱わない。そこで、8種類の物性について約25万件の高分子と物性のデータ点を含む公開データセットPolymer Benchmark 2026(PolyBench26)を紹介する。データには実験測定、密度汎関数理論、分子動力学から得たものが含まれる。 ベンチマークでは、単独重合体に加え、交互、ランダム、ブロック共重合体を対象として、同じ分布内での物性予測、データ量を変えた評価、繰り返し単位の複雑さ、学習に使わない高分子構造への移行という四つの課題を扱う。言語モデル、グラフに基づくモデル、記述子に基づくモデルを比較した結果、グラフモデルが物性予測の誤差を最も低く抑え、評価した学習データ量全体で優位性を保ち、繰り返し単位が複雑になっても頑健だった。PolyBench26は、複雑さが増す高分子設計のためのモデル開発を再現可能にする基盤となる。ベンチマークは公開ソースで提供される。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-22(UTC)
最新改訂
2026-09-22 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Polymer property prediction lacks open, standardized benchmarks that enable rigorous comparison of machine-learning methods, with existing resources covering only a narrow fraction of polymer architectures, such as homopolymers. We introduce Polymer Benchmark 2026 (PolyBench26), an open dataset comprising nearly 250,000 polymer-property datapoints across eight physical properties, including data from experimental measurements, density functional theory, and molecular dynamics. The benchmark supports four evaluation tasks across homopolymers and alternating, random, and block copolymers: in-distribution property prediction, dataset-size scaling, repeat-unit complexity, and transfer to held-out polymer architectures. We compare language model, graph-based, and descriptor-based approaches and find graph-based models provide the lowest errors in property prediction, retain their advantage across the evaluated training-set sizes, and remain robust to increasing repeat-unit complexity. PolyBench26 provides a reproducible foundation for developing models for the increasingly complex polymer design space. The PolyBench26 benchmark is available open-source at https://github.com/rlearsch/PolymerBenchmark2026.

arXiv ID: 2609.27036 / 要約の誤りについて