arXiv論文メモ
新着一覧
physics.comp-ph / cs.LG / physics.flu-dyn · 査読状況未確認

流体予測モデルの事前学習は分布変化でどう変わるか

How Does Distribution Shift Shape Pretraining Gains in Neural PDE Surrogates?

Pochinapeddi Sai Bhargav, Nithin Somasekharan, Rohit Sunil Kanchi, Sicheng He, Shaowu Pan

この論文をやさしく読む

ひとことで言うと

ニューラルPDEサロゲートの事前学習効果が、空力データの分布変化や対象側の物理モデルの違いでどう変わるかを調べた研究です。

何に役立つ?

CFDデータを追加して微調整する際、事前学習の価値をサンプル数、対象範囲、物理モデルの一致度で見積もるのに役立ちます。

この研究の面白いところ

254,909件のRANS解で事前学習し、同じSAモデルと遷移モデル付きSAへ微調整しました。1000件と5000件で優位性が入れ替わり、被覆の効果も物理モデルで異なりました。

どこまで分かった?

検証は翼型、RANS、SA系の設定に限られます。別のPDEや形状への一般化は要旨からは分かりません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

ニューラルPDE代理モデルを事前学習すると、形状やモデル化する物理が変わった際に必要となる新しい数値流体力学(CFD)データを減らせる。しかし、分布変化の異なる要素がこの利点にどう影響するかは明らかでない。 本研究では、ある翼型群に属する254,909件のレイノルズ平均ナビエ–ストークス(RANS)解で代理モデルを事前学習し、自由流の範囲をそろえた二つの対象設定で、新しい翼型群へ微調整する。一方は元と同じSpalart–Allmaras(SA)モデル、もう一方はSAにe^N遷移モデルを加えた設定である。N=1000では、事前学習モデルは、同じSAを使う対象では3.25倍の標本数でゼロから学習したモデルと同等の精度に達するが、遷移をモデル化した対象では2.58倍である。N=5000になると、この大小関係は逆転する(1.56倍対1.86倍)。 N=1000では、より多くの異なる翼型をサンプリングすると両対象で誤差が減るが、利得の増加が観測された抽出ごとの変動を上回るのは同じSAの対象だけである(3.3倍から4.0倍)。これらの結果は、事前学習の価値が、対象データの量、対象データの被覆範囲、そして事前学習側と対象側でモデル化する物理が異なるかどうかに、共同で依存することを示している。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-17(UTC)
最新改訂
2026-09-17 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Pretraining a neural PDE surrogate can reduce the amount of new CFD data needed when geometry or modeled physics changes. However, it remains unclear how different components of distribution shift affect this benefit. We pretrain a surrogate on 254,909 RANS solutions from one airfoil family and fine-tune it on a new family under two target settings with matched freestream ranges: the same Spalart-Allmaras (SA) modeling and SA with added $e^N$ transition modeling. At $N=1000$, the pretrained model matches the accuracy of a model trained from scratch on $3.25\times$ as many samples for the same-SA target, but $2.58\times$ as many for the transition-modeled target. By $N=5000$, this ordering reverses ($1.56\times$ versus $1.86\times$). At $N=1000$, sampling more distinct airfoils lowers error on both targets, but only for the same-SA target is the gain increase larger than the observed draw-to-draw variation ($3.3\times$ to $4.0\times$). These results show that pretraining value depends jointly on target-data budget, target-data coverage, and whether source and target differ in modeled physics.

著者のコメント

15 pages, 5 figures. Representations for the Physical Sciences Workshop, NeurIPS 2026

arXiv ID: 2609.20814 / 要約の誤りについて