Non-Linear Monte-Carlo Search in Civilization II.

descriptionPublicationkeyboard_double_arrow_right Conference object , Article 01 Jul 2011 United States Publisher:AAAI Press/International Joint Conferences on Artificial Intelligence

Authors: Branavan, Satchuthanan R.; Silver, David; Barzilay, Regina;

handle: 1721.1/74248

Non-Linear Monte-Carlo Search in Civilization II.

- Summary
- Metrics

Abstract

This paper presents a new Monte-Carlo search algorithm for very large sequential decision-making problems. Our approach builds on the recent success of Monte-Carlo tree search algorithms, which estimate the value of states and actions from the mean outcome of random simulations. Instead of using a search tree, we apply non-linear regression, online, to estimate a state-action value function from the outcomes of random simulations. This value function generalizes between related states and actions, and can therefore provide more accurate evaluations after fewer simulations. We apply our Monte-Carlo search algorithm to the game of Civilization II, a challenging multi-agent strategy game with an enormous state space and around $10^{21}$ joint actions. We approximate the value function by a neural network, augmented by linguistic knowledge that is extracted automatically from the official game manual. We show that this non-linear value function is significantly more efficient than a linear value function. Our non-linear Monte-Carlo search wins 80\% of games against the handcrafted, built-in AI for Civilization II.

United States. Defense Advanced Research Projects Agency (DARPA Machine Reading Program (FA8750-09-C-0172))

National Science Foundation (U.S.) (CAREER grant IIS-0448168)

National Science Foundation (U.S.) (grant IIS-0835652)

Microsoft Research (New Faculty Fellowship)

Country

United States

Related Organizations

Massachusetts Institute of Technology
United States

Impact byBIP!

	selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	0
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Average
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Average
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Average

Found an issue? Give us feedback

0

Average

Green

Fields of Science (4) View all

natural sciences

Fields of Science

natural sciences

View all