Dynamic Safe Interruptibility for Decentralized Multi-Agent Reinforcement Learning

El Mhamdi, El Mahdi; Hendrikx, Hadrien; Guerraoui, Rachid; Maurer, Alexandre David Olivier

El Mhamdi, El Mahdi; Hendrikx, Hadrien; Guerraoui, Rachid; Maurer, Alexandre David Olivier

2017

Formats

Format
BibTeX
MARC
MARCXML
DublinCore
EndNote
NLM
RefWorks
RIS

Files

Abstract

In reinforcement learning, agents learn by performing actions and observing their outcomes. Sometimes, it is desirable for a human operator to \textit{interrupt} an agent in order to prevent dangerous situations from happening. Yet, as part of their learning process, agents may link these interruptions, that impact their reward, to specific states and deliberately avoid them. The situation is particularly challenging in a multi-agent context because agents might not only learn from their own past interruptions, but also from those of other agents. Orseau and Armstrong~\cite{orseau2016safely} defined \emph{safe interruptibility} for one learner, but their work does not naturally extend to multi-agent systems. This paper introduces \textit{dynamic safe interruptibility}, an alternative definition more suited to decentralized learning problems, and studies this notion in two learning frameworks: \textit{joint action learners} and \textit{independent learners}. We give realistic sufficient conditions on the learning algorithm to enable dynamic safe interruptibility in the case of joint action learners, yet show that these conditions are not sufficient for independent learners. We show however that if agents can detect interruptions, it is possible to prune the observations to ensure dynamic safe interruptibility even for independent learners.

Details

Title Dynamic Safe Interruptibility for Decentralized Multi-Agent Reinforcement Learning

Author(s) El Mhamdi, El Mahdi ; Hendrikx, Hadrien ; Guerraoui, Rachid ; Maurer, Alexandre David Olivier

Date 2017

Publisher EPFL

Keywords

ml-ai

Laboratories DCL

Record Appears in Scientific production and competences > I&C - School of Computer and Communication Sciences > IINFCOM > DCL - Distributed Computing Laboratory
Peer-reviewed publications
Working papers
Work produced at EPFL
Published

Record creation date 2017-06-29

Actions

Preview

Select file: