arXiv論文メモ
新着一覧
cs.HC · 査読状況未確認

AI家庭教師の学習効果を指導設計と音声・文字に分けて検証

When AI Tutors Speak: Evidence from a Randomized Field Experiment

Shihao Yang, Marshall Van Alstyne, Chrysanthos Dellarocas

この論文をやさしく読む

ひとことで言うと

AIに声で話せることと、教材に沿って教える仕組みがあることを分け、大学院の授業で学習への効果を調べています。

何に役立つ?

AI学習支援を設計するとき、学習効果、対話の好み、運用費用を別々に判断する材料になります。音声への好みや会話量の増加は、学習の改善と同じではありません。

この研究の面白いところ

指導の有無は学生間で無作為化し、音声と文字は同じ学生の中で週ごとに切り替えています。教育設計と入力方法を切り分けた実験です。

どこまで分かった?

86人のオンラインMBA企業財務科目での結果です。6.63点の改善と教員の期末試験の2.6点差は別の評価です。週ごとの習熟度は音声と文字で統計的に同等とされますが、あらゆる教育場面へ一般化した結果ではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

学生が生成AIとともに学ぶ機会は増えている。しかし、流暢な回答へ指導なしでアクセスできると認知的な作業を外部へ委ねやすく、どのようなAI指導の構成が学習につながるかについての証拠は少ない。通常は、教育上の構造、すなわち教師がどう教えるかと、対話のモダリティ、すなわち学生がどう話しかけるかという二つの設計要素がひとまとめにされる。本研究では両者を分離する。 オンラインMBAの大学院企業財務科目で、事前登録した無作為化フィールド実験を行った。86人の学生を、授業教材に基づく構造化された教師を利用する群と、一般向けAIを自由に利用できる状態を保つ対照群に無作為に割り付けた。指導群では各学生の対話経路を音声と文字で毎週交替させ、同じ学生の中でモダリティの効果を識別した。指導の構造は重要だった。指導群は能力をそろえた学生に比べて得点の伸びが6.63点大きかった(p=0.007)。改善は文章による推論に集中し、要素間の関係を結びつける質に達した回答の割合は、指導群では8%から49%、対照群では8%から27%へ上がった。 モダリティは学習に影響しなかった。無作為化した全86人について保存されている担当教員自身の期末試験でも、同じ方向の差があった。差は100点中2.6点で、介入前の中間試験には差がなかった。音声では対話量がほぼ倍になり、提供費用は2.8倍になったが、週ごとの習熟度は文字と統計的に同等だった。それでも学生は音声を好むようになった。教育上の構造は学生が何を練習するかを形づくり、モダリティは教師とのやりとりの仕方を形づくる。AIを人間らしくするだけでは、教育上の効果が高まるわけではない。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-21(UTC)
最新改訂
2026-09-21 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Students increasingly study alongside generative artificial intelligence (AI), yet unguided access to fluent answers invites cognitive offloading, and there is little evidence on which configurations of AI tutoring produce learning. Two design margins are usually bundled together: pedagogical structure (how the tutor teaches) and interaction modality (how students talk to it). We separate them. In a preregistered randomized field experiment in a graduate corporate-finance course of an online MBA, we randomized 86 students between a structured tutor grounded in the course materials and a holdout in which consumer AI remained freely available. Within the tutored arm, each student's channel alternated weekly between voice and text, so the modality effect is identified within student. Structure mattered: tutored students gained 6.63 points more than ability-matched peers (p=.007), and the gain was concentrated in written reasoning, where the share of answers reaching relational quality rose from 8% to 49% in the tutored arm against 8% to 27% in the holdout. Modality did not matter for learning. The instructor's own final, on file for all 86 randomized students, shows the same direction (2.6 points of 100, with no difference on a pre-treatment midterm). Voice nearly doubled conversational interaction and cost 2.8 times as much to deliver, yet it produced weekly mastery statistically equivalent to text, even as students came to prefer it. Pedagogical structure shapes what students practice, and modality shapes how they interact with the tutor. Making an AI more humanlike does not by itself make it more educational.

arXiv ID: 2609.23958 / 要約の誤りについて