论文标题
Hearts Gym:学习强化学习作为团队活动
Hearts Gym: Learning Reinforcement Learning as a Team Event
论文作者
论文摘要
在Covid-19大流行中,本文的作者为数据科学领域的一所研究生院组织了加强学习(RL)课程。我们描述了尽管无处不在的缩放疲劳,但仍在定性地评估课程,以创造令人兴奋的学习体验的策略和材料。关键的组织特征是专注于团队中竞争性的动手设置,并提供了最少的讲座,从而提供了RL基本背景。该课程的实用部分围绕着Hearts Gym,这是我们作为RL的入门级教程开发的RL环境。参与者的任务是培训代理人探索奖励成型和其他RL超参数。为了进行最终评估,参与者的代理人相互竞争。
Amidst the COVID-19 pandemic, the authors of this paper organized a Reinforcement Learning (RL) course for a graduate school in the field of data science. We describe the strategy and materials for creating an exciting learning experience despite the ubiquitous Zoom fatigue and evaluate the course qualitatively. The key organizational features are a focus on a competitive hands-on setting in teams, supported by a minimum of lectures providing the essential background on RL. The practical part of the course revolved around Hearts Gym, an RL environment for the card game Hearts that we developed as an entry-level tutorial to RL. Participants were tasked with training agents to explore reward shaping and other RL hyperparameters. For a final evaluation, the agents of the participants competed against each other.