PUBGで音声会話しながら行動するAIのチームメート
PUBG Ally: A Conversational Embodied Agent as an AI Teammate
この論文をやさしく読む
ひとことで言うと
PUBGでプレイヤーと会話しながら一緒に動くAIのチームメートを作り、実際のプレイデータで学習・評価した研究です。
何に役立つ?
リアルタイムのゲーム行動と音声会話を組み合わせるエージェントの設計や、実際の利用者の好みを反映した評価に役立ちます。
この研究の面白いところ
約3万9,000セッションの実プレイ記録を使い、言語モデルの高水準の判断と、動作を担う速い制御層を分けています。141か国での調査も報告しています。
どこまで分かった?
推薦意向の25.1ポイント差は、ゲーム記録でAllyとのプレイを確認できた調査回答者についての値です。全プレイヤーにそのまま当てはまるとは要旨からは分かりません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
PUBG: BATTLEGROUNDSでプレイヤーと共に遊び、推論、自律行動、音声会話ができる身体性エージェントPUBG Allyを紹介する。このようなチームメートには、厳しい遅延の制約の下で絶えず変化するゲーム世界を知覚して対応することと、発話を自分の行動に同期させながらプレイヤーと自然に会話することが同時に必要になる。そこでAllyは、エージェントによるツールの利用とリアルタイムのゲーム操作を組み合わせる。言語モデルのエージェントは、制御されたインターフェースを通してゲーム情報を調べ、プレイヤーの発話を解釈し、文脈を保ち、話す内容を決め、移動、戦闘、回復を担うより速い制御層を導く高水準の行動を選ぶ。 プレイヤーとAllyの発話や行動は互いに影響し、試合の進行も変えるため、学習には実際のプレイデータが必要となる。このため、実際のプレイヤーがAllyと遊んだ約3万9,000セッションから、ゲームの進行、プレイヤーの発話、エージェントの判断、ツールの使用、行動、プレイヤーの意見を記録して、反復的な学習に使う。チームメートとしての質を評価するため、プレイヤーの意見と好みの比較から、オフライン評価と実際の好みのずれを見つけ、評価基準を反復的に改善する。実サービスで動かすためには、端末上での低遅延の実行と、プレイヤー向けの会話の保護策も必要であり、モデル圧縮、文脈の圧縮、対象を絞った安全性の学習、実行時の制約、記憶の伏せ字処理で対応する。実サービス中には141か国のプレイヤーに調査を行った。ゲーム記録でAllyとのプレイを確認できた回答者では、Allyを勧めたいかという質問で肯定的な回答が否定的な回答を25.1ポイント上回った。プレイヤーはAllyを道具だけでなく、チームメートや仲間としても説明した。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-24(UTC)
- 最新改訂
- 2026-09-24 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-24 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
We introduce PUBG Ally, an embodied agent for PUBG: BATTLEGROUNDS that can reason, act autonomously, and play alongside players as a voice-enabled teammate. Building such a teammate requires combining two difficult capabilities: it must perceive and respond to a constantly changing game world under strict latency constraints while interacting naturally with players, keeping its speech synchronized with its actions. Ally therefore combines agentic tool use with real-time game control. A language-model agent uses a controlled interface to inspect game information, interpret player speech, maintain context, decide what to say, and issue high-level action choices that steer a faster control layer for movement, combat, and recovery. Because the player's and Ally's speech and actions continually shape each other and the course of the match, training requires data from actual gameplay. We therefore collect data across nearly 39k sessions in which real players play alongside Ally, recording gameplay, player speech, agent decisions, tool use, actions, and player feedback, and use these records for iterative training. To evaluate teammate quality, we use player feedback and preference comparisons to identify gaps between offline evaluations and player preferences, and iteratively refine the evaluation criteria. Deploying Ally in live service further requires low-latency on-device execution and safeguards for player-facing communication, which we address through model compression, context compaction, targeted safety training, runtime guardrails, and memory redaction. During the live service, we surveyed players in 141 countries. Among respondents whose play with Ally was confirmed in game records, positive responses exceeded negative responses by 25.1 percentage points when asked whether they would recommend Ally, with players describing Ally not only as a tool but also as a teammate or companion.
著者のコメント
55 pages, 19 figures, 16 tables
arXiv ID: 2609.29837 / 要約の誤りについて