arXiv論文メモ
新着一覧
cs.CL · 査読状況未確認

異なる分野の世界モデルを共通の予測原理で学習

JEPA-Anything: Learning Predictive Models across Different Worlds

Taoyong Cui, Zhongyao Wang, Xinyue Xu, Weiyang Liu, Zhaochen Yu, Yuying Zhang, Qiang Gao, Mengyue Yang, Wanli Ouyang, Pheng Ann Heng, Yingcheng Wu, Zhenfei Yin, Ling Yang

この論文をやさしく読む

ひとことで言うと

異なる分野の未来予測を、潜在表現を相補的な要素に分ける共通の方法で学習します。視覚、生物、臨床時系列、制御、分子、物理場、天気の7分野で評価します。

何に役立つ?

多様な系で予測や介入の影響を学ぶための手法です。生物学的介入の候補については、細胞系やオルガノイド、組織片、マウスで実験的な支持も得ています。

この研究の面白いところ

同条件のJEPA比較に対し10の動力学課題すべてで指標が改善し、Pongの単一介入誤差は34.8%低下しています。軌道の潜在モードからケプラー則に近い指数も得ています。

どこまで分かった?

性能は報告された課題と比較条件での結果です。生物学的介入の実験的支持を、人での治療効果の実証や全分野での一般的成功へ広げることはできません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

世界モデルは、結果を予測し、介入を導き、相互作用から学ぶことを知的システムに可能にする。しかし、予測モデルは依然として分野ごとに作られている。大きく異なる系を対象とした世界モデルを、共通の学習原理で支えられるだろうか。 本研究では、直交予測因子分解(OPF)に基づく、分野に依存しない枠組みJEPA-Anythingを導入する。OPFは共同埋め込み予測アーキテクチャを拡張し、潜在空間の予測対象を相補的な因子へ分解し、専用の経路で学習して、共通の予測設計の中で再結合する。視覚、生物学、臨床経過、制御、分子動力学、物理場、気象の七分野で評価する。実験は表現学習、介入予測、分布外汎化、長期動力学に及び、条件をそろえた10の動力学課題、1000種類を超える臨床イベントの予測、4系にわたる100ステップの分子ロールアウトを含む。 対応するJEPAベースラインに比べ、JEPA-Anythingは10の動力学課題すべてで報告指標を改善し、Interventional Pongの単一介入予測誤差を34.8%減らす。4系すべてで、比較した手法の中で1ステップおよび100ステップの分子予測誤差が最も小さい。予測にとどまらず、因子から選ばれた生物学的介入は、細胞共培養、患者由来オルガノイド、腫瘍断片、マウスで実験的な支持を得た。また、潜在軌道モードからは、当てはめた傾き−1.4991としてケプラー則のスケーリング指数が再現された。 これらの結果は、異種の世界に共通する因子分解型の予測原理を支持し、世界モデルを介入や実験に裏付けられた科学的発見へ結び付ける。コードは https://github.com/Gen-Verse/JEPA-Anything 。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-17(UTC)
最新改訂
2026-09-17 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

World modeling enables intelligence to anticipate consequences, guide interventions, and learn from interaction. Yet predictive models remain domain-specific: can a common learning principle support world modeling across radically different systems? We introduce JEPA-Anything, a domain-agnostic framework based on orthogonal predictive factorization (OPF). Extending joint-embedding predictive architectures, OPF decomposes latent targets into complementary factors, learns them through dedicated pathways, and recombines them within a shared predictive design. We evaluate JEPA-Anything across seven domains: vision, biology, clinical trajectories, control, molecular dynamics, physical fields, and weather. Experiments span representation learning, intervention prediction, out-of-distribution generalization, and long-horizon dynamics, including 10 matched dynamics tasks, forecasting of over 1,000 clinical events, and 100-step molecular rollouts across four systems. Against matched JEPA baselines, JEPA-Anything improves reported metrics on all 10 dynamics tasks and reduces single-intervention prediction error on Interventional Pong by 34.8%. It achieves the lowest one-step and 100-step molecular errors among compared methods in all four systems. Beyond prediction, a factor-nominated biological intervention receives experimental support in cell co-cultures, patient-derived organoids, tumor fragments, and mice; latent orbital modes recover the Keplerian scaling exponent with a fitted slope of -1.4991. These results support a common factorized predictive principle across heterogeneous worlds, connecting world modeling with intervention and experimentally grounded scientific discovery. Code: https://github.com/Gen-Verse/JEPA-Anything

著者のコメント

Code: https://github.com/Gen-Verse/JEPA-Anything

arXiv ID: 2609.20800 / 要約の誤りについて