Probabilistic Guarantees for Safe Deep Reinforcement Learning

Jul-8-2020–arXiv.org Artificial Intelligence

Deep reinforcement learning has been successfully applied to many control tasks, but the application of such controllers in safetycritical scenarios has been limited due to safety concerns. Rigorous testing of these controllers is challenging, particularly when they operate in probabilistic environments due to, for example, hardware faults or noisy sensors. We propose MOSAIC, an algorithm for measuring the safety of deep reinforcement learning controllers in stochastic settings. Our approach is based on the iterative construction of a formal abstraction of a controller's execution in an environment, and leverages probabilistic model checking of Markov decision processes to produce probabilistic guarantees on safe behaviour over a finite time horizon. It produces bounds on the probability of safe operation of the controller for different initial configurations and identifies regions where correct behaviour can be guaranteed. We implement and evaluate our approach on controllers trained for several benchmark control problems.

controller, machine learning, reinforcement learning, (20 more...)

arXiv.org Artificial Intelligence

Jul-8-2020

arXiv.org PDF

Add feedback

Country:
- Europe > United Kingdom > England > West Midlands > Birmingham (0.04)

Genre:
- Research Report (0.64)

Industry:
- Information Technology (0.68)

Technology:
- Information Technology > Artificial Intelligence > Machine Learning
  - Reinforcement Learning (1.00)
  - Neural Networks > Deep Learning (0.94)
  - Learning Graphical Models > Undirected Networks
    - Markov Models (0.90)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found