Search CORE

2 research outputs found

Learning from Monte Carlo Rollouts with Opponent Models for Playing Tron

Author: AL Samuel
CJ Watkins
D Silver
D Silver
G Tesauro
J Baxter
J Schmidhuber
L Kocsis
M Otterlo van
RS Sutton
RS Sutton
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 30/12/2018
Field of study

This paper describes a novel reinforcement learning system for learning to play the game of Tron. The system combines Q-learning, multi-layer perceptrons, vision grids, opponent modelling, and Monte Carlo rollouts in a novel way. By learning an opponent model, Monte Carlo rollouts can be effectively applied to generate state trajectories for all possible actions from which improved action estimates can be computed. This allows to extend experience replay by making it possible to update the state-action values of all actions in a given game state simultaneously. The results show that the use of experience replay that updates the Q-values of all actions simultaneously strongly outperforms the conventional experience replay that only updates the Q-value of the performed action. The results also show that using short or long rollout horizons during training lead to similar good performances against two fixed opponents

Crossref

Proceedings - University of Groningen

University of Groningen

ARTS repository - University of Groningen

Dissertations of the University of Groningen

Learning from Monte Carlo Rollouts with Opponent Models for Playing Tron

Author: Drugan Madalina M.
Knegt Stefan
Rocha A.
van den Herik J.
Wiering Marco
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 30/12/2018
Field of study