[论文解读] A Survey of Knowledge Tracing: Models, Variants, and Applications
本综述全面将知识追踪(KT)模型分为概率、逻辑回归(logistic)和基于深度学习的方法;它回顾了模型变体,发布了开源 KT 库 EduData 和 EduKTM,并讨论 KT 的应用及未来发展方向。
Modern online education has the capacity to provide intelligent educational services by automatically analyzing substantial amounts of student behavioral data. Knowledge Tracing (KT) is one of the fundamental tasks for student behavioral data analysis, aiming to monitor students' evolving knowledge state during their problem-solving process. In recent years, a substantial number of studies have concentrated on this rapidly growing field, significantly contributing to its advancements. In this survey, we will conduct a thorough investigation of these progressions. Firstly, we present three types of fundamental KT models with distinct technical routes. Subsequently, we review extensive variants of the fundamental KT models that consider more stringent learning assumptions. Moreover, the development of KT cannot be separated from its applications, thereby we present typical KT applications in various scenarios. To facilitate the work of researchers and practitioners in this field, we have developed two open-source algorithm libraries: EduData that enables the download and preprocessing of KT-related datasets, and EduKTM that provides an extensible and unified implementation of existing mainstream KT models. Finally, we discuss potential directions for future research in this rapidly growing field. We hope that the current survey will assist both researchers and practitioners in fostering the development of KT, thereby benefiting a broader range of students.
研究动机与目标
- 激励并定义知识追踪(KT)及其在在线教育中的重要性。
- 提供对 KT 模型的系统分类(概率、逻辑回归、基于深度学习)。
- 评审在学习过程的前、中、后阶段建模学习的变体。
- 突出可用的数据集和用于 KT 的开源库(EduData、EduKTM)。
- 讨论 KT 在各种教育情境中的应用并概述未来的研究方向。
提出的方法
- 将 KT 模型分为三类:概率、逻辑回归,以及基于深度学习。
- 总结基础 KT 模型(例如 Bayesian Knowledge Tracing、Dynamic Bayesian Knowledge Tracing、LFA、PFA、KTM)及其技术特征。
- 详细介绍基于深度学习的 KT 家族(DKT、memory-aware KT、exercise-aware KT、attentive KT、graph-based KT)及代表机制(RNN/LSTM、记忆网络、注意力、transformers)。
- 回顾在学习前、学习中、学习后阶段处理更丰富的辅助信息和认知假设的变体。
- 提供开源资源(EduData、EduKTM)并总结 KT 的应用与未来方向。
实验结果
研究问题
- RQ1核心 KT 模型及其在概率、逻辑回归与深度学习范式中的分布是什么?
- RQ2KT 的变体如何捕捉学习前/中/后动态和侧信息?
- RQ3存在哪些数据集和开源工具可用于标准化 KT 的评估与实现?
- RQ4KT 在多样化教育情境中的关键应用有哪些,哪些未来方向最具潜力?
主要发现
- 该领域分为三大模型家族:概率、逻辑回归和基于深度学习的 KT。
- 广泛的 KT 变体扩展了基本模型以捕捉完整的学习过程和附带信息。
- 两个开源库(EduData 和 EduKTM)提供数据集和统一实现,以支持 KT 研究。
- KT 的应用跨越自适应学习、资源推荐和教育游戏等场景。
- 本综述概述了推进 KT 方法及适用性的潜在未来研究方向。
更好的研究,从现在开始
从阅读论文到最终审阅,大幅缩短您的研究时间。
无需绑定信用卡
本解读由 AI 生成,并经人工编辑审核。