Stochastic Game Theoretic trajectory optimization in continuous time

descriptionPublicationkeyboard_double_arrow_right Article , Conference object 01 Dec 2016Publisher:IEEEJournal:2016 IEEE 55th Conference on Decision and Control (CDC)

Authors: Wei Sun 0032; Evangelos A. Theodorou; Panagiotis Tsiotras;

doi: 10.1109/cdc.2016.7799217

Stochastic Game Theoretic trajectory optimization in continuous time

- Summary
- Metrics

Abstract

A Stochastic Game Theoretic Differential Dynamic Programming (SGT-DDP) algorithm is derived to solve a differential game under stochastic dynamics. We present the update law for the minimizing and maximizing controls for both players and provide a set of backward differential equations for the second order value function approximation. We compute the extra terms in the backward propagation equations that arise from the stochastic assumption compared with the original GTDDP. We present the SGT-DDP algorithm and analyze how the design of the cost function affects the feed-forward and feedback parts of the control policies under the game theoretic formulation. The performance of SGT-DDP is then investigated through simulations on two examples, namely, a first order nonlinear system, the inverted pendulum and the cart pole problems with conflicting controls. We conclude with some possible future extensions.

Related Organizations

Georgia Institute of Technology
United States

Impact byBIP!

	selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	6
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Top 10%
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Average
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Average