Classification

Baseline Scheduling Policies (Accelerating Human Learning With Deep Reinforcement Learning)

The authors compare TRPO with four baseline scheduling policies: (1) a random policy, (2) the Leitner system, (3) a variant of SuperMemo, and (4) a threshold-based policy. The threshold-based policy directly accesses the student simulator's parameters to calculate recall likelihoods, so it serves as an upper bound in the experiments.

0

1

Updated 2026-08-29

Tags

Data Science