Activity (Process)

Statistical Evaluation Procedures (Using deep reinforcement learning for personalizing review sessions on e-learning platforms with spaced repetition)

The study used four statistical evaluation procedures. (1) When varying the number of items, it formed distributions of mean pairwise reward differences across student models, assessed normality with the Shapiro-Wilk test, and, because variance homogeneity did not hold, used Welch's two-sample test followed by the Games-Howell test. (2) It compared TRPO and TNPG using rewards from all episodes and runs and the Kruskal-Wallis test. (3) It compared likelihood-based and average-of-sum-of-outcomes reward functions with the Kruskal-Wallis test, followed by Dunn's post hoc test with Bonferroni correction. (4) It evaluated TRPO with LSTM reward shaping across tutors using the Kruskal-Wallis test, followed by Dunn's test with Bonferroni correction, and also compared the LSTM-based approach with other methods used by the DRL agents.

0

1

Updated 2026-08-11

Tags

Data Science

Related