arXiv論文メモ
新着一覧
stat.ME / stat.AP / stat.CO · 査読状況未確認

連続値と離散値が混じるデータのグラフ構造をベイズ推論で学ぶ

Bayesian Posterior Learning of Mixed Graphical Models

Erdong Guo, Alexandros Beskos, and Maria De Iorio

この論文をやさしく読む

ひとことで言うと

連続値と離散値が混じるデータについて、変数同士の関係をグラフとして推定するベイズ法。

何に役立つ?

異種データの依存関係を調べるのに役立つ。要旨では模擬データと乳がん遺伝子発現データで評価した。

この研究の面白いところ

離散変数を潜在ガウス変数で表し、二種類の尤度に対応したMCMC手法を設計する。

どこまで分かった?

精度比較は記載されたシミュレーション条件に基づく。PAM50で見いだした依存関係から因果関係が示されたとは述べていない。

v2のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

混合グラフモデル(MGM)は、連続値と離散値の両方を含む異種データから構造を学ぶ柔軟な枠組みを与える。しかし、取り得るグラフを組合せ的に探索し、事後分布を計算することが難しいため、MGMのベイズ推論は依然として課題である。本論文では離散成分を潜在ガウス変数で表し、二つの尤度の指定を考える。一つはコピュラに基づく順位尤度で、copula-MGMを与える。もう一つはしきい値に基づくprobit形式で、probit-MGMを与える。 次に、ベイズMGMの事後分布をシミュレートするMCMC手法群Mixed Graph WWAを提案する。WWAアルゴリズムを土台に、copula-MGM用のcopula-WWAとprobit-MGM用のprobit-WWAという二つの専用アルゴリズムを開発する。どちらも潜在ガウス表現を使い、潜在変数の追加とグラフ構造の更新を交互に行うGibbsサンプリングで事後推論を行う。 広範なシミュレーションでは、提案法のグラフ回復精度はBirth–Death MCMCに基づくcopula-BD MCMCやprobit-BD MCMCなどの既存法と同等以上であり、事後分布の効率的な探索と、計算時間当たりの有効標本数の良さも保った。さらに、乳がんの遺伝子発現データPAM50への適用では、推定されたMGMから遺伝子発現の特徴とがんの亜型との間に意味のある依存関係が見いだされた。著者らは、Mixed Graph WWAがMGMでのベイズ的な構造学習に拡張可能で原理的な道具だと述べる。

v2の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-20(UTC)
最新改訂
2026-09-22 · v2
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Mixed Graphical Models (MGMs) provide a flexible framework for structure learning from heterogeneous data by treating sets of both continuous and discrete. Bayesian inference for MGMs remains challenging due to the combinatorial complexity of graph space exploration and posterior computation. In this paper, we model discrete components through latent Gaussian variables and consider two likelihood specifications: a copula-based ranked likelihood, yielding the copula-MGM, and a probit formulation based on cut-off points, yielding the probit-MGM. We then propose the Mixed Graph WWA, a class of MCMC methods for posterior simulation in Bayesian MGMs. Building upon the WWA algorithm, we develop two specialized algorithms: copula-WWA for copula-MGMs and probit-WWA for probit-MGMs. Both methods exploit the latent Gaussian representations to perform posterior inference through a Gibbs sampling scheme that alternates between latent-variable augmentation and graph-structure updates. Through extensive simulation studies we demonstrate that the proposed methods achieve graph recovery accuracy comparable to or better than existing approaches, including copula-BD MCMC and probit-BD MCMC based on the Birth-Death MCMC methodology, while maintaining efficient posterior exploration and favorable effective sample size per unit computational time. We further illustrate the practical utility of our approach through an application to the PAM$50$ breast cancer gene expression dataset, where the inferred MGMs reveal meaningful dependencies between gene expression profiles and cancer subtypes. These results highlight the effectiveness of the Mixed Graph WWA method as a scalable and principled tool for Bayesian structure learning in MGMs.

著者のコメント

29 pages, 10 figures, 1 table

arXiv ID: 2609.23857 / 要約の誤りについて