AIへの意思決定委託が所得課税の厚生効果を変える
Welfare-Opaque Income: Taxation under AI-Agent Delegation
この論文をやさしく読む
ひとことで言うと
AIに経済的な選択を任せると、同じ課税所得の変化でも本人の厚生への影響が違い得ると論じます。
何に役立つ?
課税の効果を所得変化だけから評価する際、AIが希望を行動に変換する仕組みも必要になることを整理します。税制度の理論分析に向けた知見です。
この研究の面白いところ
生産能力の見えなさに加えて、AIの実行ルールの見えなさを組み込んでいます。五つのAIエンジン、4500回のモデル実行による統制された比較も行っています。
どこまで分かった?
税率と目的の相互作用の向きはエンジン間で共通せず、集計結果もQwenを含めるかに依存します。モデル実行の結果であり、現実の納税者全体の行動を実証したものではありません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
政府から見えない規則を通じてAIエージェントが経済的に重要な選択を実行する場合の、所得課税を研究する。観測できない生産能力に加え、この隠れた選好から実行への対応関係が「二重の観測不能性」を生む。つまり、同じ観測可能な課税ベースの反応でも、厚生に異なる影響を与え得る。その結果生じる所得を、本研究では「厚生が不透明な所得」と呼ぶ。構成例により、機械的な厚生ウェイトが同一であっても、課税ベースの統計量は一致する一方、税制改革の厚生効果は異なり得ることを示す。 よく知られた十分統計量に、反応で重み付けした実行の歪みを加える最適課税条件を導出する。限界税率を高めると、局所的な過剰実行のもとでは是正の便益が得られ、局所的な過少実行のもとでは追加の費用が生じる。この歪みを観測すれば、現行税率表のもとでの限界的改革の厚生効果を識別でき、歪みの上下限が分かれば、その効果の上下限も得られる。 統制された実験では、五つのAIエンジンにわたる4,500回のモデル実行を比較する。忠実な委託では、ほぼすべての実行でスコアを最大化する選択肢が選ばれる。目標が競合すると反応は多様になる。Claudeはおおむねスコア最大化を維持し、GLMは主に下方へ動き、GPT-miniとQwenは分布の下側の裾に集中した増加を示す。Qwenは大幅な下方調整も行う。エンジンごとに、設計された分布の中で逸脱が起きる位置と方向が異なる。明示的なスコアを与えるとモデルの順位付けは一致し、数式で目的を指示すると一致の程度はより不均一になる。Qwenでは税と目的の明確な正の交互作用が見られるが、その方向はエンジン間で一般化できず、統合した結果の符号はQwenを含めるかどうかに依存する。この分析は、実行に関する情報が、従来の課税ベースの統計量を補完することを示す。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-17(UTC)
- 最新改訂
- 2026-09-17 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-17 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
We study income taxation when an AI agent implements economically relevant choices through a rule hidden from the government. Alongside unobserved productive ability, this hidden preference-to-execution mapping creates \emph{double unobservability}: the same observable tax-base response can carry different welfare consequences. We call the resulting income \emph{welfare-opaque}. Our constructions show that tax-base statistics can coincide while reform welfare effects differ, even when mechanical welfare weights are identical. We derive an optimal-tax condition that adds a response-weighted execution wedge to the familiar sufficient statistics. A higher marginal rate gains a corrective benefit under local over-execution and an additional cost under local under-execution. Observing the wedge identifies the welfare effect of a marginal reform at the prevailing schedule; bounds on it deliver bounds on that effect. A controlled laboratory compares 4,500 model runs across five AI engines. Faithful delegation selects the score maximizer in essentially all runs. Conflicted objectives produce heterogeneous responses: Claude largely preserves the score maximizer, GLM moves predominantly downward, and GPT-mini and Qwen show concentrated lower-tail increases. Qwen also makes substantial downward adjustments. Different engines locate their departures at different points and in different directions of the designed distribution. Explicit scores align model rankings; formula-based objective instructions yield more uneven agreement. Qwen shows a clear positive tax-by-objective interaction, but its direction does not generalize across engines and the pooled sign depends on its inclusion. The analysis identifies execution information as a complement to conventional tax-base statistics.
arXiv ID: 2609.20425 / 要約の誤りについて