arXiv論文メモ
新着一覧
cs.CY · 査読状況未確認

ChatGPTを使うプログラミング学習で成績と記憶に生じる差

Your Programming Students' Cognition with ChatGPT: Higher Performance, Lower Retention, and Reduced Ownership

Christian Bergh, Benjamin Tag, Alexandra Vassar, Jake Renzella

この論文をやさしく読む

ひとことで言うと

ChatGPTを使った学生はコード課題で高得点を取った一方、後で内容を思い出す得点と、自分で作ったという感覚は低かった、という実験です。

何に役立つ?

生成AIを使う授業で、提出コードの完成度と学習内容の保持を別々に評価する根拠になります。説明や思い出す活動を組み込むことは著者の提案で、その教育効果をこの実験で検証したわけではありません。

この研究の面白いところ

直後から再生得点に差がある一方、その後48時間の忘却量の群間差は有意ではありません。「AIを使うと忘れる速度が速くなる」とは異なる結果です。

どこまで分かった?

解析対象は学部生55人、C言語の入門課題3つ、ChatGPT-4.5を使う特定の実験設定です。生理指標は大幅な欠損があり、有意差がないことの解釈には制約があります。長期的な学習成果までは示されていません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

生成AIは学生のプログラミング成績を改善できるが、課題をうまく終えられたことが、その内容を保持していることを示すとは限らない。コンピュータサイエンス専攻の学部生59人を対象とする統制された参加者間実験で、成績、保持、認知負荷、自分の成果だという感覚を調べ、55人を解析対象とした。参加者は、ChatGPT-4.5を利用するか、生成AIを使わず従来型のウェブ検索を利用して、C言語の入門課題を3つ実施した。課題成績、自己申告の精神的努力と難しさ、瞳孔反応、心拍変動、自分の成果だという感覚を測定し、直後と48時間後に手がかり再生を評価した。 ChatGPTの支援を受けた学生は、コーディング得点が高かった(89%対69%)一方、直後の再生得点(41%対53%)と48時間後の再生得点(39%対52%)は低かった。48時間の間に失われた再生情報の量について、群間に有意差はなかった。自己申告の精神的努力は、ChatGPT条件のほうが課題を通じた増加が小さく(Holm補正後p=0.047)、学生が提出コードのうち自分によるものとした割合も低かった(45%対81%)。確認的な生理指標の検定では、条件間の推移に有意差は検出されなかったが、大幅なデータ欠損が解釈を制約する。 これらの結果は、この設定で、支援下の課題成績と、その後の再生および自分の成果だという感覚の間に隔たりがあることを示す。学生が教育の能動的な参加者として、提出物を説明し、思い出し、その作成に貢献することを求める評価方法やAI学習ツールの必要性を示唆している。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-18(UTC)
最新改訂
2026-09-18 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Generative AI can improve students' programming performance, but successful task completion may not reflect what they retain. We examined performance, retention, cognitive load, and ownership in a controlled between-subjects experiment with 59 undergraduate computer science students, 55 were retained for analysis. Participants completed three introductory C programming tasks with access to ChatGPT-4.5 or conventional web search without generative AI. We measured task performance, self-reported mental effort and difficulty, pupillary responses, heart rate variability, and ownership, and assessed cued recall immediately and 48 hours later. ChatGPT-assisted students achieved higher coding scores (89% vs. 69%) but lower recall scores immediately (41% vs. 53%) and after 48 hours (39% vs. 52%). There was no significant difference in the loss of recall information over 48 hours between the groups. Self-reported mental effort increased less across tasks in the ChatGPT condition (Holm-adjusted p = .047), and students attributed less of the submitted code to themselves (45% vs. 81%). Confirmatory physiological tests did not detect significant differences in trajectories between conditions; substantial data loss limits their interpretation. These findings reveal a gap between assisted task performance and subsequent recall and sense of ownership in this setting. They motivate the need for assessment practices and AI learning tools that require students to explain, retrieve, and contribute to the work they submit as active participants in their education.

著者のコメント

17 pages, 9 figures, 7 tables

arXiv ID: 2609.21194 / 要約の誤りについて