Baseline Scheduling Policies (Accelerating Human Learning With Deep Reinforcement Learning)
The authors compare TRPO with four baseline scheduling policies: (1) a random policy, (2) the Leitner system, (3) a variant of SuperMemo, and (4) a threshold-based policy. The threshold-based policy directly accesses the student simulator's parameters to calculate recall likelihoods, so it serves as an upper bound in the experiments.
0
1
Contributors are:
Who are from:
Tags
Data Science
Related
Environments (Accelerating Human Learning With Deep Reinforcement Learning)
Analysis (Accelerating Human Learning With Deep Reinforcement Learning)
Research Question (Accelerating Human Learning With Deep Reinforcement Learning)
Implementation Details (Accelerating Human Learning With Deep Reinforcement Learning)
Baseline Scheduling Policies (Accelerating Human Learning With Deep Reinforcement Learning)