arXiv論文メモ
新着一覧
stat.ME / stat.CO · 査読状況未確認

角度データの欠損をトーラス上の確率モデルで補う

Model-based estimation and imputation with torus missing values

Luca Greco, Lucia Filippozzi, Claudio Agostinelli

この論文をやさしく読む

ひとことで言うと

角度のように一周すると元に戻るデータは、普通の数値と同じ補い方では不都合が生じます。その周期性を確率分布に組み込み、欠損値と分布のパラメータを推定する研究です。

何に役立つ?

複数の角度を同時に扱うデータで、欠損がある場合の統計解析に役立ちます。円周に沿った構造を保ちながら、モデルに基づいて補完する枠組みを提供します。

この研究の面白いところ

欠けた値だけでなく、値を何周分巻き戻すかに関わる係数も潜在変数としてEM法で扱います。巻き付け前の分布の性質を利用して計算を組み立てています。

どこまで分かった?

欠損メカニズムが無視可能であることが条件です。有限標本評価は巻き付け正規分布によるシミュレーションで、例示データの欠損も人為的に導入したものです。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

本論文は、欠損値を含むp次元トーラス上の多変量円周データについて、パラメータ推定とモデルに基づく補完の問題を扱う。標本空間の周期性により、ユークリッド空間のデータ向けに設計された通常の補完手法は適用できない。そこで、欠損メカニズムが無視可能な場合に、巻き付け楕円対称分布族のもとで最尤推定を行う一般的枠組みを提案し、特に多変量巻き付け正規分布に注目する。 この方法は、巻き付ける前の空間における楕円対称分布の条件付きの性質を利用し、巻き付け係数と欠損要素の双方を潜在変数として扱う期待値最大化(EM)アルゴリズムに、欠損したトーラスデータの補完を組み込む。EステップとMステップの両方の導出を詳述し、実行可能なアルゴリズムを議論する。補完方法についても検討する。無視可能な欠損のもとでの最尤推定量の有限標本性能を、巻き付け正規分布の設定におけるモンテカルロシミュレーションで評価する。また、人為的に欠損を導入したデータで手法を例示する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-18(UTC)
最新改訂
2026-09-18 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

This paper addresses the problem of parameter estimation and model-based imputation for multivariate circular data lying on a p-dimensional torus in the presence of missing values. Actually, the periodic nature of the sample space invalidates conventional imputation techniques designed for Euclidean data. Then, we propose a general framework for maximum likelihood estimation under the wrapped elliptically symmetric family of distributions, with particular interest in the multivariate wrapped normal distribution, when the missing data mechanism is ignorable. The methodology leverages the conditional properties of the elliptically symmetric distributions on the unwrapped space, embedding the imputation of missing torus data into an Expectation-Maximization algorithm that treats both the wrapping coefficients and the missing entries as latent variables. Derivation of both the E and M steps is detailed and a working algorithm is discussed. Imputation methods are also taken into account. The finite-sample performance of the maximum likelihood estimator under ignorable missingness is assessed through Monte Carlo simulations under the wrapped normal specification. The methodology is also illustrated on data with artificially introduced missingness.

arXiv ID: 2609.21549 / 要約の誤りについて