arXiv論文メモ
新着一覧
cs.AI · 査読状況未確認

対話役割エージェントの弱点に合わせて訓練場面を更新

Adversarial Closed-Loop Curriculum for Evolving Role-Playing Agents

Zheng Zhang, Liu Liu, Qi Chai, Deheng Ye, Peilin Zhao, Mao Zheng, Hao Wang

この論文をやさしく読む

ひとことで言うと

役割演技エージェントの苦手な場面を自動で作り直し、訓練内容を学習の進み具合に合わせて変える方法。

何に役立つ?

対話エージェントの訓練場面が固定される問題を改善し、多様な人物像や文脈への対応を評価する材料になる。

この研究の面白いところ

Actorの現在の評価点を下げる書き換えほどRewriterに報酬を与え、苦手な領域を継続的に探す。

どこまで分かった?

三つの既存ベンチマークと新しい多言語ベンチマークでの比較結果であり、実際の支援業務での性能は要旨では報告されていない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

大規模言語モデルに基づく役割演技エージェントは、個別の支援や社会的シミュレーションなどに広く使われている。最近の強化学習法は通常、学習開始前に集めた固定の場面群で訓練する。しかし、エージェントが上達すると苦手な場面も変わるのに、訓練分布はそのままであり、学習上の制約となる。本研究は、役割演技の強化学習を閉じた循環のカリキュラムへ変える、敵対的な文脈書き換え法AdvRoleを提案する。役を演じるActorと、人物像および対話文脈をそのActorにとって難しい場面へ編集するRewriterを交互に訓練する。Rewriterには、元の場面と比べて現在のActorの評価点を下げる書き換えを優遇する、性能差に基づく報酬を使う。 この結果、場面群はActorとともに変化し、人物像と文脈の空間のうち、まだ十分に習得していない領域を継続して狙う。英語と中国語を含む三つの役割演技ベンチマークと、著者らが新たに公開した多言語ベンチマークでの実験では、AdvRoleは比較手法を一貫して上回った。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-23(UTC)
最新改訂
2026-09-23 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Role-playing agents based on large language models have been widely applied in areas such as personalized assistance and social simulation. Recent RL methods typically train on a fixed scenario pool collected before learning begins. This creates a distributional bottleneck: as the agent improves, the scenarios where it performs poorly also change, while the training distribution remains static. Therefore, we propose AdvRole, an adversarial context rewriting framework that turns role-playing RL into a closed-loop curriculum. AdvRole alternates between an Actor that learns to role-play and a Rewriter that edits character profiles and dialogue contexts into actor-specific hard scenarios. The Rewriter is trained with a performance-gap reward, which favors rewrites that reduce the current Actor's score relative to the original scenario. As a result, the scenario pool evolves with the Actor and continuously targets under-mastered regions of the character-context space. Experiments on three role-playing benchmarks covering English and Chinese, as well as a new multilingual benchmark we release, show that AdvRole consistently outperforms baselines.

arXiv ID: 2609.28609 / 要約の誤りについて