Sciweavers

Free Online Productivity Tools i2Speak i2Symbol i2OCR iTex2Img iWeb2Print iWeb2Shot i2Type iPdf2Split iPdf2Merge i2Bopomofo i2Arabic i2Style i2Image i2PDF iLatex2Rtf Sci2ools

11

CIG
2006
IEEE

favoriteEmaildiscussreport

190views Applied Computing» more CIG 2006»

Monte-Carlo Go Reinforcement Learning Experiments

13 years 10 months ago

Monte-Carlo Go Reinforcement Learning Experiments

Download www.math-info.univ-paris5.fr

Abstract— This paper describes experiments using reinforcement learning techniques to compute pattern urgencies used during simulations performed in a Monte-Carlo Go architecture. Currently, Monte-Carlo is a popular technique for computer Go. In a previous study, Monte-Carlo was associated with domain-dependent knowledge in the Go-playing program Indigo. In 2003, a 3x3 pattern database was built manually. This paper explores the possibility of using reinforcement learning to automatically tune the 3x3 pattern urgencies. On 9x9 boards, within the Monte-Carlo architecture of Indigo, the result obtained by our automatic learning experiments is better than the manual method by a 3-point margin on average, which is satisfactory. Although the current results are promising on 19x19 boards, obtaining strictly positive results with such a large size remains to be done.

Bruno Bouzy, Guillaume Chaslot

Real-time Traffic

Applied Computing | CIG 2006 | Monte-Carlo | Monte-Carlo Go Architecture | Reinforcement Learning |

claim paper

Related Content

» MonteCarlo simulation balancing

» Eligibility Traces for OffPolicy Policy Evaluation

» Which landmark is useful Learning selection policies for navigation in unknown environment...

» Evaluation in Go by a Neural Network using Soft Segmentation

» Coevolutionary Temporal Difference Learning for smallboard Go

» Learning to play Tetris applying reinforcement learning methods

» Action Selection in Bayesian Reinforcement Learning

» Using the Interaction Rhythm as a Natural Reinforcement Signal for Social Robots A Matter ...

Post Info
More Details (n/a)

Added	10 Jun 2010
Updated	10 Jun 2010
Type	Conference
Year	2006
Where	CIG
Authors	Bruno Bouzy, Guillaume Chaslot

Comments (0)