arXiv論文メモ
新着一覧
cs.LG / stat.CO · 査読状況未確認

A/Bテストのばらつきを減らす多分岐の層化決定木

Optimal Multi-way Decision Trees for Stratified Sampling in Online Controlled Experiments

Tomoka Takei, Shunnosuke Ikeda, Yuichi Takano

この論文をやさしく読む

ひとことで言うと

A/Bテストの対象者を分ける層を決定木で選び、推定値のばらつきを減らした。

何に役立つ?

標本数を増やさずに実験の検出力を高めるため、層化規則を設計する際に役立つ。

この研究の面白いところ

解釈しやすい多分岐木を、分散を最小にする二値最適化で選び、候補削減で計算量も抑えた点。

どこまで分かった?

実験は実データと模擬データの指定された条件で行われ、要旨に改善量の具体的な数値はない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

オンラインでの対照実験、すなわちA/Bテストは、デジタルプラットフォームで因果効果を推定するために広く使われる。課題の一つは標本数を増やさずに検出力を高めることである。層化抽出は古典的な分散削減法だが、その効果は層の作り方に大きく左右される。本研究は、最適な多分岐決定木を用いて層を作る、最適化に基づく枠組みOMSTを提案する。特徴量のグラフ上の経路選択問題として層化を定式化し、選んだ経路を解釈可能な層化規則とする。連続的な比例配分とNeyman型の最適配分の下で、分散を厳密に最小化する二値最適化で経路を選ぶ。 数値の特徴量については、結果と関係する候補分割を作るため、教師ありの最適な区間分けを導入する。また、重複する候補経路と割当制約を減らす手順を加え、最適化問題の規模を大きく縮める。実データと模擬データでの実験では、OMSTは浅く解釈しやすい層化木を保ちながら、既存の方法と同等かそれ以上に分散を減らした。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-20(UTC)
最新改訂
2026-09-20 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Online controlled experiments, or A/B tests, are widely used to estimate causal effects on digital platforms. A central challenge is to improve experimental sensitivity, or statistical power, without increasing the experimental sample size. Stratified sampling is a classical variance reduction technique; however, its effectiveness depends critically on how the strata are constructed. We thus propose an optimization-based stratification framework for stratified sampling using optimal multi-way decision trees. Our method, called Optimal Multi-way Stratification Trees (OMST), formulates stratification as a path-selection problem over a feature graph. The selected paths define interpretable stratification rules and are optimized using an exact variance-minimizing binary optimization formulation under continuous proportional allocation and a Neyman-type optimal allocation. We incorporate supervised optimal binning to generate outcome-relevant candidate splits for numerical features. Furthermore, we introduce reduction procedures for redundant candidate paths and assignment constraints, substantially reducing the optimization problem size. Experiments on both a real-world and a simulated dataset demonstrate that OMST achieves comparable or superior variance reduction to existing methods while maintaining shallow and interpretable stratification trees.

著者のコメント

16 pages, 4 figures, The 23rd Pacific Rim International Conference on Artificial Intelligence 2026 (PRICAI 2026)

arXiv ID: 2609.23308 / 要約の誤りについて