last modified :2011-10-12release date :2011-10-12date/time of measurement start :2010-04-22date/time of measurement end :2010-11-20collection environment :We recruited students to participate in an ex
This dataset contains the results of experiments comparing the performance of the standard Q-learning based distributional deep reinforcement learning algorithm QL-C51, and a novel variant which uses