arXiv論文メモ
新着一覧
cs.CL · 査読状況未確認

言語モデルが物語を説明的な要約に変える偏りの検証計画

Summarization Bias: The Directional Collapse of Objective Projection into Told-Mode Labels in Large Language Models --- A Conceptual Framework and Registered Test Protocol

Levent Bulut

この論文をやさしく読む

ひとことで言うと

言語モデルが物語の暗示を読み取るより、感情を明示する短いラベルへ寄せがちではないか、という仮説を整理します。

何に役立つ?

AIを文章の評価者や報酬モデルとして使う際、どの表現を好むかを検証するための試験設計になります。

この研究の面白いところ

生成と評価の二つの場面を分け、仮説を棄却する判断規則まで含めた試験を事前登録しています。

どこまで分かった?

要旨はバイアスが実証済みとは主張していません。既存研究を整合的な方向性の証拠として再解釈した概念枠組みと検証計画です。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

本論文は「要約バイアス」を導入し、測定可能な形で定義する。これは、大規模言語モデルが物語の意味を、その意味を生み出す再構成可能な推論構造としてではなく、抽象的な要約ラベルとして表す、という系統的傾向の仮説である。Bulut Doctrineでは、物語の効果を「説明する」から「見せる」への軸で理論化する。説明する様式では、感情や情報が明示され、読者による再構成はほとんど要らない。見せる様式では、内容が表層では伏せられ、物理的な手がかりや間接表現から再構成しなければならない。これをObjective Projectionと呼ぶ。見せる様式は、この理論が測定するために設計された、より高い負荷の条件である。 主張するのは、言語モデルがこの軸に沿って特定の方向に失敗するということである。要約バイアスは、二つの状況で働くと仮定する。(i)生成では、Objective Projectionを通して感情を表すよう求められたモデルが、代わりに感情を明言してしまう。(ii)評価では、物語の質を判断するモデルが、説明する様式の明示性を高く評価し、見せる様式での抑制を十分に検出しない。 評価の状況の方が影響は大きい。言語モデルは判定者や報酬モデルとして使われることが増えており、説明する様式への方向性を持つ偏りがあるなら、文章を平板な明言へ劣化させる選択圧が生まれるからである。 本報告は、この偏りが実証済みだとは主張しない。概念を定義し、言語モデルを判定者に使う際の偏りとの関係に位置づけ、完了済みの独立した信頼性研究を、この仮説と整合する方向性の証拠として再解釈する。そして、どの条件ならこの概念を放棄するかという判断規則を含む、二つの状況の検証を事前登録する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-17(UTC)
最新改訂
2026-09-17 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

This paper introduces and operationalizes summarization bias: a proposed systematic tendency of large language models (LLMs) to represent narrative meaning as an abstract summary label rather than as the reconstructable inferential structure that produces it. Within the Bulut Doctrine, narrative effect is theorized along a told-shown axis: in told mode, emotional and informational content is declared explicitly and requires little reader reconstruction; in shown mode, that content is suppressed at the surface and must be reconstructed from physical cues and indirection (Objective Projection). Shown mode is the higher-load condition the doctrine is designed to measure. The claim is that LLMs fail along this axis in a specific direction. Summarization bias is hypothesized to operate in two regimes: (i) a generative regime, in which a model asked to render an emotion through Objective Projection defaults to declaring it instead; and (ii) an evaluative regime, in which a model judging narrative quality rewards told-mode explicitness and under-detects shown-mode suppression. The evaluative regime is the more consequential, since LLMs increasingly serve as judges and reward models, and a directional bias toward told mode would impose a selection pressure degrading prose toward flat declaration. This report does not claim the bias is validated. It defines the construct, situates it against LLM-as-judge biases, rereads a completed independent reliability study as directional evidence consistent with it, and pre-registers a two-regime test with decision rules under which the construct would be abandoned.

著者のコメント

v1.1. 8 pages. Also archived at Zenodo: https://doi.org/10.5281/zenodo.22817289

arXiv ID: 2609.20712 / 要約の誤りについて