Journal of Educational and Behavioral Statistics · 2022 · 48 citations · 30 references
Artificial IntelligenceEngineeringMachine LearningSequential LearningEducationReinforcement Learning (Educational Psychology)Lifelong Reinforcement LearningReinforcement Learning (Computer Engineering)Data ScienceContinual Learning (Lifelong Deep Learning)Human LearningAdaptive Learning ProblemAutonomous LearningSequential Decision MakingComputer ScienceLifelong Deep LearningDeep LearningLearning PolicyDeep Reinforcement LearningDeep Q-learning AlgorithmContinual Learning (Educational Psychology)
The adaptive learning problem concerns how to create an individualized learning plan (also referred to as a learning policy) that chooses the most appropriate learning materials based on a learner’s latent traits. In this article, we study an important yet less-addressed adaptive learning problem—one that assumes continuous latent traits. Specifically, we formulate the adaptive learning problem as a Markov decision process. We assume latent traits to be continuous with an unknown transition model and apply a model-free deep reinforcement learning algorithm—the deep Q-learning algorithm—that can effectively find the optimal learning policy from data on learners’ learning process without knowing the actual transition model of the learners’ continuous latent traits. To efficiently utilize available data, we also develop a transition model estimator that emulates the learner’s learning process using neural networks. The transition model estimator can be used in the deep Q-learning algorithm so that it can more efficiently discover the optimal learning policy for a learner. Numerical simulation studies verify that the proposed algorithm is very efficient in finding a good learning policy. Especially with the aid of a transition model estimator, it can find the optimal learning policy after training using a small number of learners.
30
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver et al. · Nature · 2015 · 28.8K citations
Artificial Intelligence, Engineering, Deep Reinforcement Learning +3
Statistical Theories of Mental Test Scores
William W. Rozeboom, Frederic M. Lord, Melvin R. Novick et al. · American Educational Research Journal · 1969 · 8.1K citations
Bias reduction of maximum likelihood estimates
David Firth · Biometrika · 1993 · 4.2K citations
Parameter Estimation, Engineering, Regular Parametric Problems +10
A Rasch Model for Partial Credit Scoring
Geoff N Masters · Psychometrika · 1982 · 3.7K citations