AIモデルの世代継承を集団遺伝学で説明する
The evolution of sex for artificial intelligence: a population-genetic framework for multigenerational model populations
この論文をやさしく読む
ひとことで言うと
AIが別のAIの出力を学び、複数のモデルを統合しながら世代交代するときの振る舞いを、集団遺伝学の数理と対応付けた研究です。
何に役立つ?
生成データを繰り返し学習する際のモデル崩壊や、モデル統合で能力が残る条件を考えるのに役立ちます。検証済み実データの追加や統合方式の違いを、継承の問題として比較できます。
この研究の面白いところ
複数の親を使うだけでは十分でなく、出力を平均すると利点が消える一方、各親の強い寄与を残す統合では利点が保たれると報告しています。また、規約の衝突と単なる変化の蓄積を区別します。
どこまで分かった?
Wright–Fisher過程との厳密な一致は最小継承モデルについての結果です。ネットワークには構造固有の偏りも報告されています。実データは比率より絶対数が重要という知見や統合の成功を、あらゆるAI学習条件の保証として一般化することはできません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
AI開発の一部は、モデルが専門化され、他のモデルの出力で再学習され、あるいは重みの平均によって組み合わされるという集団過程に似ている。こうした実践は、集団遺伝学が研究する生物学的な意味でのモデルの世代を生む。本研究では、この対応関係を発展させ、多世代のモデル集団を有性生殖と無性生殖の観点から解釈することで、2つの分野を形式的に結び付ける。厳密な継承モデル、学習済みネットワーク(再帰型、順伝播型、変分オートエンコーダー型の生成器)、大規模言語モデルでこれらの類比を検証し、構造に固有の測定可能な偏りはあるものの、一般に成り立つことを示す。 モデルの出力を再帰的に学習するとモデル崩壊に至ることは知られており、この過程は以前から遺伝的浮動に似ると説明されてきた。本研究では、その類比から導かれる帰結を展開する。親の出力で再学習する学習器の最小モデルは、Wright–Fisher過程を厳密に再現する。各世代に加える検証済みの実データは移入の役割を果たし、集団遺伝学とまったく同じように、重要なのは実データの比率ではなく絶対的なサンプル数であるという意外な結果が得られる。 複数の親の出力の平均で子を学習すると、複数の親を持つ利点が失われる。これは混合遺伝に対応し、ダーウィンに対するJenkinの反論を想起させる。一方、それぞれの親の最も強い寄与が残るように組み合わせると、その利点は保たれる。統合した専門言語モデルは、異なる乱数シードにわたってすべての親を上回り、Fisher–Muller効果に対応する結果を示した。また、系統が統合能力を完全に失って生殖的に隔離されるのは、単に浮動によって離れた場合ではなく、互いに矛盾する規約を学習した場合である。AIの社会が空間だけでなく時間にもまたがる社会になるにつれ、その継承を記述する数学的枠組みは予測力を持つようになる。注目すべきことに、この枠組みは生物学からほぼそのまま適応できる。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-16(UTC)
- 最新改訂
- 2026-09-16 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-16 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Some aspects of AI development resemble a population process in which models are specialised, retrained on the output of peers, or combined by averaging weights. These practices lead to generations of models, in the biological sense studied by population genetics. Here, I develop this parallelism and interpret multigenerational model populations in terms of sexual and asexual reproduction, formally recombining the two fields. I test these analogies in an exact inheritance model, in trained networks (recurrent, feedforward and variational autoencoder generators) and in large language models, and show that they hold generally, with some measurable architecture-specific biases. Training recursively on model output is known to lead to model collapse, a process previously described as akin to genetic drift; I develop all that follows. A minimal model of a learner retrained on its parent's output reproduces the Wright-Fisher process exactly; verified real data added to each generation play the role of immigration, with the surprising finding that the absolute number of real data samples matters, not their share, exactly as in population genetics. Training a child on the average of its parents' outputs cancels the benefit of having several parents, matching blending inheritance (and reviving Jenkin's objection to Darwin), whereas combining parents so that each keeps its strongest contribution preserves it; merged language-model specialists exceeded every parent across seeds (the Fisher-Muller effect); and lineages become reproductively isolated, losing the ability to merge at all, when they have learned conflicting conventions and not when they have merely drifted apart. As AI societies become societies in time as well as in space, a mathematical framework for their inheritance acquires predictive power. Remarkably, that framework can be adapted almost wholesale from biology.
著者のコメント
22 pages, 5 figures, 1 table. Supplementary Information (26 pp) and a plain-language figure appendix for readers from biology (23 pp) are included as ancillary files. Code, configs and seeds: https://git.lab.gilest.ro/giorgio/MachineSex
arXiv ID: 2609.18560 / 要約の誤りについて