arXiv論文メモ
新着一覧
cs.HC · 査読状況未確認

教師が設定した教育用AIチャットボットは意図どおりに振る舞うか

Will It Teach as Intended? How Teachers Configure Educational AI Chatbots

Bahare Riahi, Deniz Ozturk, Alice Guth, Jiayu Li, Daksh Pratap Singh, Xiaoyi Tian, Jennifer Chiu, Nicholas Lytle, Tiffany Barnes, Veronica Catete

この論文をやさしく読む

ひとことで言うと

教師が教育用AIチャットボットに設定した指導方針が、実際の応答にどれだけ反映されるかを調べた研究。

何に役立つ?

教師向けチャットボット作成ツールで、設定項目の説明や応答の試用・改善機能を設計する際の参考になる。

この研究の面白いところ

応答性やペルソナの一致度は比較的高い一方、授業の目的を表す設定の一致度は59.3%で、設定可能であることと意図どおり動くことの差を示した。

どこまで分かった?

対象はワークショップに参加した中学校教師27人と、その設定・対話記録である。生徒の学習成果への効果は要旨に記載されていない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

教師は授業を支えるために生成AIを使うことが増えているが、教育上の意図がチャットボットの設定へどう変換され、実際の振る舞いにどう表れるかは明確でない。本研究は、中学校教師27人が参加する専門能力開発ワークショップで、教師向けのチャットボット作成ツールを調べた。フォーカスグループの面接に加え、設定記録と対話記録を分析した。教師はチャットボットを、学習者に応じた支援を提供し、支援へのアクセスを広げ、教師が定めた範囲内で生徒自身の思考を保つ教育上の足場として捉えていた。 設定の分析では、「Purpose」は主に授業上の目標と内容の焦点を記し、「Rules」は教育的な振る舞い、制約、学習者に応じた調整を指定する傾向があった。記録に基づく評価では、意図との一致度は応答性が88.9%、ペルソナが81.5%で、「Rules」の70.4%や「Purpose」の59.3%より高かった。この結果は、設定を変えられる制御項目があるだけでは教育上の意図への忠実さは保証されないことを示す。教師が意図する振る舞いを表現し、試し、改善できる作成ツールが必要である。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-24(UTC)
最新改訂
2026-09-24 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Teachers are increasingly using generative AI to support instruction, yet it remains unclear how pedagogical intentions are translated into chatbot configurations and reflected in chatbot behavior. We studied a teacher-facing chatbot authoring tool in professional development workshops with 27 middle school teachers, analyzing focus-group interviews alongside configuration and interaction logs. Teachers envisioned chatbots as instructional scaffolds that could provide differentiated support, extend access to assistance, and preserve student thinking within teacher-defined boundaries. Configuration analysis showed that Purpose primarily captured instructional goals and content focus, whereas Rules more often specified pedagogical behavior, guardrails, and learner-specific adaptations. Log-based evaluation showed stronger alignment for responsiveness (88.9%) and persona (81.5%) than for rules (70.4%) and purpose (59.3%). These findings show that configurable controls alone do not ensure pedagogical fidelity and highlight the need for authoring tools that help teachers express, test, and refine intended chatbot behavior.

著者のコメント

17 pages, 5 figures

arXiv ID: 2609.29993 / 要約の誤りについて