arXiv論文メモ
新着一覧
stat.ME / stat.AP · 査読状況未確認

回答と所要時間の相互情報量から性急な回答を見分ける

Mutual Information as a Tool for Optimal Classification: Application to Identifying Rapid-Responding Behaviour

Santeri Holopainen, Jari Metsämuuronen, Mikko-Jussi Laakso, Janne V. Kujala

この論文をやさしく読む

ひとことで言うと

テストへの回答内容と回答時間の関係から、極端に速い回答を区別するしきい値を決める方法です。

何に役立つ?

大規模学力調査で、回答行動を分析する際の代替手法になります。母集団の分布に特定の形を仮定する必要を減らす狙いがあります。

この研究の面白いところ

正誤だけでなく選択した回答そのものも使い、時間の2分類・3分類との相互情報量を最大化します。情報量に基づいて境界を決める点が特徴です。

どこまで分かった?

PISA 2022数学データへの適用と、一つの方法の条件別検討を報告しています。速い回答が必ず不真面目だと直接証明するものではなく、要旨には分類精度の具体値はありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

大規模な学力評価で性急な回答行動を特定する既存の方法は、母集団についてパラメトリックな仮定を必要とする。本研究では、その代替として、相互情報量に基づく新しいノンパラメトリックな方法群の枠組みを提案する。各方法は、観測された回答と離散化された回答時間の相互情報量を計算し、情報利得を最大化することで、性急な回答と課題に取り組んだ回答を分ける閾値を定める。方法間の違いは、対象となる変数のカテゴリー数だけである。 具体的に3つの方法を示す。第1の方法は回答の正誤と2値化した回答時間を使う。第2の方法は正誤を使い、時間を3群に分ける。第3の方法は回答そのものと2値化した時間を使う。これらを、2022年の国際的な学習到達度調査(PISA)で収集された数学の学力データに適用した。さらに、第1の方法について、いくつかの現実的な条件下での挙動と使いやすさを、母集団レベルと実現したデータのレベルで調べた。 結果は、この枠組みが性急な回答を特定するための有効な代替手段となることを示した。新規性は、ノンパラメトリックであること、および正誤ではなく回答そのものを利用できることにある。最後に、この主題について今後考えられる研究の方向を議論する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-17(UTC)
最新改訂
2026-09-17 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Existing methods for identifying rapid-responding behaviour in large-scale assessments require parametric assumptions about the population. In this study, we propose a novel, non-parametric, mutual information-based framework of methods as an alternative. The methods within this framework compute the mutual information of the observed responses and discretised response times and maximise the information gain to determine a threshold that differentiates rapid responses from engaged responses. The only difference between the methods is the number of categories in the relevant variables. We present three methods explicitly. The first method uses response correctness and binarised response times. The second method uses correctness and categorises time into three groups. The third method uses raw responses and binarised times. We applied these methods to mathematics achievement data collected through the Programme for International Student Assessment in 2022. Furthermore, we examined the behaviour and usability of the first method in certain realistic conditions at the population and realised levels. The results indicated that the proposed framework is a viable alternative for identifying rapid responses. The framework's novelty lies in its non-parametric nature and its ability to utilise raw responses instead of correctness. Finally, we discuss some possible future research directions on this topic.

著者のコメント

Main text: 24 pages, 5 figures. Supplement: 22 pages, 3 figures. The PISA 2022 dataset is available at: https://www.oecd.org/en/data/datasets/pisa-2022-database.html The R codes are available at: https://github.com/sajomaho-uni/Empirical-Results-of-MaxMI-for-RRB-Identification

arXiv ID: 2609.19781 / 要約の誤りについて