arXiv論文メモ
新着一覧
cs.LG · 査読状況未確認

多様なグラフを学ぶClifford表現モデルICE

ICE: Task-Aligned Clifford Latent Fields for Multimodal Graph Foundation Models

Xunkai Li, Xu Wang, Yinlin Zhu, Xiong Yongfu, Yi Liu, Rong-Hua Li, Guoren Wang

この論文をやさしく読む

ひとことで言うと

文章・画像・接続関係を持つグラフに対し、Clifford代数の複数のグレードを使って実体の意味と関係を表す基盤モデルです。

何に役立つ?

ノード分類やリンク予測に共通の表現を使う研究の参考になります。報告された11グラフと各評価タスクでの性能であり、あらゆるグラフへの汎化が確認されたわけではありません。

この研究の面白いところ

実体の意味を守る経路を設けながら高次の関係と複数層の情報を残し、30の報告比較すべてで首位だったとしています。

どこまで分かった?

性能の根拠は要旨に記載されたデータセットと比較設定です。理論的性質も示していますが、未評価のグラフや用途での性能は要旨から判断できません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

複数の形式の属性を持つグラフは、実体、画像、言語、観測された関係を結び付ける。こうしたグラフを共通の基盤モデルで学ぶには、各ノードを融合したユークリッド空間のベクトルに圧縮するだけでは足りない。表現は、実体の意味を保ち、グラフ近傍から相互作用の状態を作り、幾何学的な構造の異なる予測器にその状態を渡せなければならない。著者らの実証的な検討では、これらの条件が切り離せないことが示された。高次のグレードのチャネルは基盤グラフ間の対関係を捉え、専用のクエリは一般的な読み出しでは隠れる情報を明らかにし、ブレードを厳格に分離するとグレード間の表現能力が失われる。 そこで、ノードを添字とするClifford潜在場を用いた、マルチモーダルグラフ基盤モデルICE(Interaction-aware Clifford Encoder)を導入する。グラフの接続構造、文章、画像をCl(3)の明示的なアドレスに入力する。辺を考慮した幾何積によって、これらの方向を、観測された近傍にわたるスカラー、双ベクトル、三ベクトルの関係へ変換する。保護されたGrade-1経路が実体の意味を維持し、一方で全グレードと深さの表現群を新たなノード予測器とリンク予測器が利用できる。 著者らは、グレードをまたぐ到達可能性の厳密な性質、ノードの並べ替えに対する同変性、意味スコアの周りのタスク残差に対する上界を示す。評価は、11グラフに共通の一つの基盤、ノード分類6データセット、リンク予測3データセット、条件を揃えた少数例学習のタスクに及ぶ。報告された教師あり学習と少数例学習の30比較すべてでICEが1位となった。主要部分を除くと各タスクの要約指標が低下し、機構を調べる対照実験は、改善が高次の情報伝播、多層の構造保持、意味情報の保護、潜在場への直接アクセスに関係することを示した。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-24(UTC)
最新改訂
2026-09-24 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Multimodal attributed graphs connect entities, visual content, language, and observed relations. Learning one foundation across such graphs requires more than compressing each node into a fused Euclidean vector. The representation must preserve entity semantics, construct interaction state from graph neighborhoods, and expose that state to prediction units with different geometry. Our empirical study shows why these requirements are inseparable. Higher-grade channels recover pair relations across the foundation graphs, specialized queries reveal information hidden by a generic readout, and rigid blade isolation removes cross-grade capacity. We therefore introduce ICE (Interaction-aware Clifford Encoder), a multimodal graph foundation model built on a node-indexed Clifford latent field. Topology, text, and images enter explicit Cl(3) addresses. Edge-aware geometric products transform these directions into scalar, bivector, and trivector relations over observed neighborhoods. A protected Grade-1 route preserves entity semantics, while the full grade and depth bank remains available to fresh node and link heads. We establish exact cross-grade reachability, node-permutation equivariance, and a bound on the task residual around the semantic score. Experiments span one shared foundation over eleven graphs, six node-classification datasets, three link-prediction datasets, and matched few-shot tasks. ICE ranks first in all 30 reported supervised and few-shot comparisons. Core removals reduce every task summary, and mechanism controls connect the gains to higher-order transport, retained multidepth structure, semantic protection, and direct field access.

arXiv ID: 2609.29398 / 要約の誤りについて