Reinforcement Learning in an Environment Synthetically Augmented with Digital Pheromones

descriptionPublicationkeyboard_double_arrow_right Article 13 Mar 2014 English Publisher:Hindawi LimitedJournal:Advances in Artificial Intelligence, volume 2,014, pages 1-23 (issn: 1687-7470, eissn: 1687-7489,

Copyright policy )

Authors: Salvador E. Barbosa; Mikel D. Petty;

doi: 10.1155/2014/932485

Reinforcement Learning in an Environment Synthetically Augmented with Digital Pheromones

- Summary
- Metrics

Abstract

Reinforcement learning requires information about states, actions, and outcomes as the basis for learning. For many applications, it can be difficult to construct a representative model of the environment, either due to lack of required information or because of that the model's state space may become too large to allow a solution in a reasonable amount of time, using the experience of prior actions. An environment consisting solely of the occurrence or nonoccurrence of specific events attributable to a human actor may appear to lack the necessary structure for the positioning of responding agents in time and space using reinforcement learning. Digital pheromones can be used to synthetically augment such an environment with event sequence information to create a more persistent and measurable imprint on the environment that supports reinforcement learning. We implemented this method and combined it with the ability of agents to learn from actions not taken, a concept known as fictive learning. This approach was tested against the historical sequence of Somali maritime pirate attacks from 2005 to mid-2012, enabling a set of autonomous agents representing naval vessels to successfully respond to an average of 333 of the 899 pirate attacks, outperforming the historical record of 139 successes.

Related Organizations

University of Alabama in Huntsville
United States

Impact byBIP!

	selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	2
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Average
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Average
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Average