AI検索でブランド名が挙がる条件を観測データで分析
From Prompt to Recommendation: A Fitted Stage Model of Brand Visibility in AI Search
この論文をやさしく読む
ひとことで言うと
AI検索であるブランド名が回答に出るかどうかと、検索経路での自社サイト露出や過去の言及との関係を分析した研究。
何に役立つ?
ブランドの検索上の可視性を観測する指標の設計に役立つ可能性がある。因果効果や検索エンジンの内部動作を証明したものではない。
この研究の面白いところ
同じプロンプトを繰り返した場合でも、現在の検索経路と前回までの言及履歴を合わせると予測性能が高くなる。
どこまで分かった?
観測データと予測モデルに基づく関係であり、要旨も因果的な説明ではないと明記する。対象は指定のプロジェクト、期間、GPT・Geminiの実行に限られる。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
匿名化されたAisoプロジェクト75件から、ブランド名を含まないプロンプトと検索エンジンの観測結果3万4960件を分析した。対象は異なる監視プロンプト2854件で、2026年6~9月にGPTとGeminiで繰り返し実行された。観測できるその場の検索経路に、対象ブランドもそのブランド自身のドメインも現れない場合、ブランド言及率はGPTで2.8%、Geminiで3.8%だった。自社ドメインの引用はあるがブランド名を含む検索の展開がない場合、49.0%と58.4%に上がった。自社ドメインの露出とブランド名を含む検索の展開が両方ある場合は91.4%と100%だった。 同じプロジェクト、プロンプト、エンジンを繰り返し観測しても関係は残る。ブランド名を含む検索の展開がなく、自社ドメインの露出だけが変わるプロンプト群では、露出はGPTで平均40.2パーセントポイント、Geminiで49.0ポイント高い言及率と関連した。過去の可視性にも独立した持続性がある。前回は非言及で今回も自社ドメインの露出がない場合、次回の言及率は1.6%と1.9%だった。一方、前回は言及され、今回も露出がある場合は80.5%と83.7%だった。 過去の実行履歴と現在の検索指標を使い、時間順に診断するロジスティックモデルを当てはめた。式はlogit P(M_t=1)=α_e+β_e logit(P̃_(t−1))+γ_e E_t+δ_e F_t+θ_eᵀXである。最新の30%を保留した評価では、完全なモデルのAUCはGPTで0.963、Geminiで0.942であり、過去履歴だけの場合の0.937/0.917、現在の指標だけの場合の0.880/0.840を上回った。手作業で選んだプロンプトでの感度分析でも、AUCはほぼ同じ0.960と0.943だった。別の199プロンプトによるページ集合の検証では、プロンプトとページの一致がGeminiの露出を予測する程度はAUC 0.641で、GPTの0.545より明瞭だった。これは、関連性より下流に、エンジンを介したより大きな露出効果があることを示す。式は予測のための観測モデルであり、非公開のエンジン内部に関する因果的な説明ではない。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-19(UTC)
- 最新改訂
- 2026-09-19 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-19 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
We analyze 34,960 unbranded prompt-engine observations from 75 anonymized Aiso projects, covering 2,854 distinct monitored prompts and repeated GPT and Gemini runs from June-September 2026. When neither the target brand nor its own domain appears in the observable live retrieval path, target mention rates are 2.8% for GPT and 3.8% for Gemini. With an own-domain citation but no branded fan-out, they rise to 49.0% and 58.4%. When both own-domain exposure and a branded fan-out occur, mention rates reach 91.4% and 100%. The relationship persists within the same project, prompt, and engine across repeated runs: among prompt cells that vary in own-domain exposure while holding branded fan-out absent, exposure is associated with a mean mention-rate increase of 40.2 percentage points on GPT and 49.0 points on Gemini. Prior visibility is independently persistent. A previous non-mention plus no current own-domain exposure yields next-run mention rates of 1.6% and 1.9%; previous mention plus current exposure yields 80.5% and 83.7%. We fit a chronological diagnostic model using prior-run history and contemporaneous retrieval indicators: $ \operatorname{logit}P(M_t=1)=\alpha_e+\beta_e\operatorname{logit}(\widetilde P_{t-1})+\gamma_e E_t+\delta_e F_t+\theta_e^\top X. $ On the latest 30% holdout, the full model achieves AUC 0.963 on GPT and 0.942 on Gemini, compared with 0.937/0.917 for prior history alone and 0.880/0.840 for live signals alone. A manually curated prompt sensitivity gives nearly identical AUCs (0.960 and 0.943). A separate 199-prompt page-corpus validation finds that prompt-page match predicts Gemini exposure (AUC 0.641) more clearly than GPT exposure (0.545), placing relevance upstream of a larger engine-mediated exposure effect. The equation is predictive and observational, not a causal description of proprietary engine internals.
著者のコメント
29 pages, 10 figures. Includes aggregate results and figure-reproduction code
arXiv ID: 2609.23162 / 要約の誤りについて