深層クラスタリングで見つけた群の差を統計的に検定
Selective Inference for Deep Clustering in Latent Spaces
この論文をやさしく読む
ひとことで言うと
AIがデータから見つけたグループの違いが、統計的に確かかを検定する方法。
何に役立つ?
ゲノムデータなどの深層クラスタリング結果を解釈する際、同じデータで群を発見・検定することによる偏りを補正するのに役立つ。
この研究の面白いところ
非線形なエンコーダーを通した群の選択過程を考慮し、合成実験で過誤率の制御と基準法より高い検出力を両立した。
どこまで分かった?
要旨で扱うのは固定した事前学習済みエンコーダーによる群分け。エンコーダーも同じデータで学習する場合の妥当性は示されていない。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
深層クラスタリングは、高次元データから低次元の潜在表現を学んでから群分けし、意味のある構造を発見する方法である。しかし、得られた群の統計的な信頼性を評価することは難しい。同じデータで群を見つけて差を検定すると選択バイアスが生じ、通常のp値は妥当でなくなる。選択的推論(SI)はこのバイアスを補正する枠組みだが、従来法は観測した特徴を直接クラスタリングする場合に重点を置いていた。 本研究では、事前学習済みのエンコーダーを固定した深層クラスタリングに対するSIの枠組みを開発する。群の割り当ては元データ空間から潜在空間への非線形変換を経て決まるため、通常のクラスタリングより選択過程が大幅に複雑になる。提案法は、この過程を計算可能な形で考慮し、潜在空間で特定された群間の差を妥当な形で統計検定できるようにする。合成データの実験では、第1種の過誤率を制御しつつ、妥当だが保守的な基準法より高い検出力を示した。ゲノムデータへの適用では、選択バイアスを適切に考慮しながら、有意な群間差を特定できることを示した。この枠組みは、深層クラスタリングで見つかった構造の統計的な信頼性を評価する方法を提供する。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-23(UTC)
- 最新改訂
- 2026-09-23 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-23 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Deep clustering is a powerful approach for discovering meaningful structures in high-dimensional data by learning a low-dimensional latent representation prior to clustering. Despite its empirical success, assessing the statistical reliability of the resulting clusters remains challenging. Testing discovered clusters on the same data induces selection bias and invalidates classical $p$-values. Selective inference (SI) provides a principled framework for correcting this bias, but existing methods focus on clustering performed directly on the observed features. In this work, we develop an SI framework for deep clustering with a fixed pretrained encoder. The key challenge is that cluster assignments are determined through a nonlinear transformation from the original data space to the latent space, resulting in a substantially more complex selection process than in conventional clustering. Our method provides a computationally tractable way to account for this process and enables valid statistical testing of differences between clusters identified in the latent space. Synthetic experiments demonstrate that the proposed method controls the Type I error rate while achieving higher power than valid but conservative baselines, and genomic applications show that it can identify significant cluster differences while appropriately accounting for selection bias. Our framework provides a principled approach to quantifying the statistical reliability of structures discovered by deep clustering.
著者のコメント
62 pages, 7 figures
arXiv ID: 2609.28756 / 要約の誤りについて