Chance-Constrained Control With Lexicographic Deep Reinforcement Learning

descriptionPublicationkeyboard_double_arrow_right Article , Preprint , Other literature type 01 Jul 2020Embargo end date: 01 Jan 2020 Italy Publisher:Institute of Electrical and Electronics Engineers (IEEE)Journal:IEEE Control Systems Letters, volume 4, pages 755-760 (eissn: 2475-1456,

Copyright policy )Funded by:EC | 5G-ALLSTAR

Authors: Giuseppi A.; Pietrabissa A.;

doi: 10.1109/lcsys.2020.2979635 , 10.48550/arxiv.2010.09468

arXiv: 2010.09468

handle: 11573/1382636

Chance-Constrained Control With Lexicographic Deep Reinforcement Learning

- Summary
- Subjects
- Metrics

Abstract

This paper proposes a lexicographic Deep Reinforcement Learning (DeepRL)-based approach to chance-constrained Markov Decision Processes, in which the controller seeks to ensure that the probability of satisfying the constraint is above a given threshold. Standard DeepRL approaches require i) the constraints to be included as additional weighted terms in the cost function, in a multi-objective fashion, and ii) the tuning of the introduced weights during the training phase of the Deep Neural Network (DNN) according to the probability thresholds. The proposed approach, instead, requires to separately train one constraint-free DNN and one DNN associated to each constraint and then, at each time-step, to select which DNN to use depending on the system observed state. The presented solution does not require any hyper-parameter tuning besides the standard DNN ones, even if the probability thresholds changes. A lexicographic version of the well-known DeepRL algorithm DQN is also proposed and validated via simulations.

published version at: https://doi.org/10.1109/LCSYS.2020.2979635 in this version we fixed a typo in (9)

Country

Italy

Related Organizations

Roma Tre University
Italy
Sapienza University of Rome
Italy

Keywords

FOS: Computer and information sciences, Computer Science - Machine Learning, constrained control.; deep reinforcement learning; Markov decision processes, Artificial Intelligence (cs.AI), Computer Science - Artificial Intelligence, FOS: Electrical engineering, electronic engineering, information engineering, Systems and Control (eess.SY), Electrical Engineering and Systems Science - Systems and Control, Machine Learning (cs.LG)

Impact byBIP!

	selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	6
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Top 10%
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Average
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Top 10%