Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Amirhossein Zolfagharian

SMARLA: A Safety Monitoring Approach for Deep Reinforcement Learning Agents

Aug 03, 2023

Amirhossein Zolfagharian, Manel Abdellatif, Lionel C. Briand, Ramesh S

Figure 1 for SMARLA: A Safety Monitoring Approach for Deep Reinforcement Learning Agents

Figure 2 for SMARLA: A Safety Monitoring Approach for Deep Reinforcement Learning Agents

Figure 3 for SMARLA: A Safety Monitoring Approach for Deep Reinforcement Learning Agents

Figure 4 for SMARLA: A Safety Monitoring Approach for Deep Reinforcement Learning Agents

Abstract:Deep reinforcement learning algorithms (DRL) are increasingly being used in safety-critical systems. Ensuring the safety of DRL agents is a critical concern in such contexts. However, relying solely on testing is not sufficient to ensure safety as it does not offer guarantees. Building safety monitors is one solution to alleviate this challenge. This paper proposes SMARLA, a machine learning-based safety monitoring approach designed for DRL agents. For practical reasons, SMARLA is designed to be black-box (as it does not require access to the internals of the agent) and leverages state abstraction to reduce the state space and thus facilitate the learning of safety violation prediction models from agent's states. We validated SMARLA on two well-known RL case studies. Empirical analysis reveals that SMARLA achieves accurate violation prediction with a low false positive rate, and can predict safety violations at an early stage, approximately halfway through the agent's execution before violations occur.

Via

Access Paper or Ask Questions

Search-Based Testing Approach for Deep Reinforcement Learning Agents

Jun 15, 2022

Amirhossein Zolfagharian, Manel Abdellatif, Lionel Briand, Mojtaba Bagherzadeh, Ramesh S

Figure 1 for Search-Based Testing Approach for Deep Reinforcement Learning Agents

Figure 2 for Search-Based Testing Approach for Deep Reinforcement Learning Agents

Figure 3 for Search-Based Testing Approach for Deep Reinforcement Learning Agents

Figure 4 for Search-Based Testing Approach for Deep Reinforcement Learning Agents

Abstract:Deep Reinforcement Learning (DRL) algorithms have been increasingly employed during the last decade to solve various decision-making problems such as autonomous driving and robotics. However, these algorithms have faced great challenges when deployed in safety-critical environments since they often exhibit erroneous behaviors that can lead to potentially critical errors. One way to assess the safety of DRL agents is to test them to detect possible faults leading to critical failures during their execution. This raises the question of how we can efficiently test DRL policies to ensure their correctness and adherence to safety requirements. Most existing works on testing DRL agents use adversarial attacks that perturb states or actions of the agent. However, such attacks often lead to unrealistic states of the environment. Their main goal is to test the robustness of DRL agents rather than testing the compliance of agents' policies with respect to requirements. Due to the huge state space of DRL environments, the high cost of test execution, and the black-box nature of DRL algorithms, the exhaustive testing of DRL agents is impossible. In this paper, we propose a Search-based Testing Approach of Reinforcement Learning Agents (STARLA) to test the policy of a DRL agent by effectively searching for failing executions of the agent within a limited testing budget. We use machine learning models and a dedicated genetic algorithm to narrow the search towards faulty episodes. We apply STARLA on a Deep-Q-Learning agent which is widely used as a benchmark and show that it significantly outperforms Random Testing by detecting more faults related to the agent's policy. We also investigate how to extract rules that characterize faulty episodes of the DRL agent using our search results. Such rules can be used to understand the conditions under which the agent fails and thus assess its deployment risks.

Via

Access Paper or Ask Questions