arXiv論文メモ
新着一覧
cs.AI · 査読状況未確認

少数の例示が言語モデル内部の世界表現の利用を促す

Few-Shot Demonstrations Elicit the Use of In-Context World Representations in LLMs

Kohsei Matsutani, Gouki Minegishi, Core Francisco Park, Takeshi Kojima, Yusuke Iwasawa, Yutaka Matsuo

この論文をやさしく読む

ひとことで言うと

言語モデルに別々の環境の例を少数見せると、内部に作った環境の表現を予測に使いやすくなるかを調べています。正解率だけでなく、内部表現への介入も行っています。

何に役立つ?

新しい環境を観測しながら動くエージェントに、どのような例を文脈に含めるかを考える材料になります。グラフ追跡だけでなく、ウェブ操作やオセロなどでも改善を調べています。

この研究の面白いところ

例示によって情報が増えるだけでなく、世界表現の内部での位置が変わっています。その表現を乱すと他の部分空間を乱すより性能が下がることから、予測への利用を検証しています。

どこまで分かった?

検証は指定したモデルと課題の範囲です。要旨には改善幅の具体値がなく、少数例を追加すれば任意の環境で必ず改善することを示したわけではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

大規模言語モデル(LLM)がエージェントとして動作するときには、文脈中の観測データを受け取り、世界の背後にある潜在的な状態空間を推測し、それを後続の予測に利用することが期待される。しかし先行研究は、データ生成過程を支配するグラフの表現を構築し、その後の予測へ使う必要があるグラフ追跡課題で、LLMが文脈内で学習した表現を使うことに苦労すると示した。 本論文では、各例示が、同じまたは異なるグラフの接続構造を持つ別の世界から生成される少数例設定へ拡張すると、4モデル系列の6モデルで予測が改善することを示す。この改善を理解するため、隠れ状態にグラフ情報を符号化した低次元の世界表現を、線形プローブで調べる。特に、少数の例示は世界表現の位置を移し、予測での利用を増やすことが分かった。具体的には、各モデルで世界表現は元の部分空間にほぼ直交する方向へ移り、この表現への介入は、他の部分空間への介入よりも選択的に性能を損なう。 この知見と整合的に、異なる世界からの観測を含む少数例の提示によって、ARC-AGI-1および2、ウェブエージェント課題、オセロで性能が改善することを示す。結果は、文脈内の世界モデリングにおける少数例提示の役割と内部機構を明らかにする。より広くは、LLMエージェントが文脈内観測から学ぶ仕組みの理解を進め、さらなる改善に示唆を与える。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-21(UTC)
最新改訂
2026-09-21 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Large language models (LLMs), when acting as agents, are expected to take observed data in context, infer the latent state space underlying the world, and leverage it for downstream prediction. However, prior work demonstrated that LLMs struggle to use representations learned in context on a graph tracking task, where the model needs to construct a representation of the graph governing data generation process and use it for subsequent predictions. In this paper, we show that extending this to few-shot settings, where each demonstration is generated from a different world with either the same or different graph topologies, enhances its prediction on 6 models from 4 model families. To understand this improvement, we linearly probe a low-dimensional world representation that encodes graph information in the hidden states. Notably, we find that few-shot demonstrations relocate the world representation and increase its predictive use. Specifically, for each model, these world representations shift in directions nearly orthogonal to their original subspace, and interventions on these representations selectively impair performance more than interventions on other subspaces. Consistent with this insight, we show that few-shot demonstrations with observations from different worlds improve performance on ARC-AGI-1&2, web agent tasks, and Othello. Our findings elucidate the role and internal mechanisms of few-shot demonstrations in in-context world modeling. More broadly, our work advances our understanding of how LLM agents learn from in-context observations and provides implications for their further improvement.

arXiv ID: 2609.24352 / 要約の誤りについて