Paper: Safety-Aware Apprenticeship Learning

FLOC 2018: FEDERATED LOGIC CONFERENCE 2018

Safety-Aware Apprenticeship Learning

Authors: Weichao Zhou and Wenchao Li

Title:	Safety-Aware Apprenticeship Learning
Authors:	Weichao Zhou and Wenchao Li
Proceedings:	CAV All Papers
Editors:	Georg Weissenbacher, Hana Chockler and Igor Konnov
Keywords:	AI Safety, Apprenticeship Learning, Inverse Reinforcement Learning, Reinforcement Learning, Probabilistic Model Checking, Counterexample-Guided Inductive Synthesis
Abstract:	ABSTRACT. Apprenticeship learning (AL) is a class of “learning from demonstrations” techniques where the reward function of a Markov Decision Process (MDP) is unknown to the learning agent and the agent has to derive a good policy by observing an expert’s demonstrations. In this paper, we study the problem of how to make AL algorithms inherently safe while still meeting its learning objective. We consider a setting where the unknown reward function is assumed to be a linear combination of a set of state features, and the safety property is specified in Probabilistic Computation Tree Logic (PCTL). By embedding probabilistic model checking inside AL, we propose a novel counterexample-guided approach that can ensure both safety and performance of the learnt policy. We demonstrate the effectiveness of our approach on several challenging AL scenarios where safety is essential.
Pages:	18
Talk:	Jul 15 16:30 (Session 107A: Probabilistic Systems)
Paper: