Search-Based Testing of Reinforcement Learning

Martin Tappler*, Filip Cano Cordoba, Bernhard Aichernig, Bettina Könighofer

*Korrespondierende/r Autor/-in für diese Arbeit

Publikation: Beitrag in Buch/Bericht/KonferenzbandBeitrag in einem KonferenzbandBegutachtung

Abstract

Evaluation of deep reinforcement learning (RL) is inherently challenging. Especially the opaqueness of learned policies and the stochastic nature of both agents and environments make testing the behavior of deep RL agents difficult. We present a search-based testing framework that enables a wide range of novel analysis capabilities for evaluating the safety and performance of deep RL agents. For safety testing, our framework utilizes a search algorithm that searches for a reference trace that solves the RL task. The backtracking states of the search, called boundary states, pose safety-critical situations. We create safety test-suites that evaluate how well the RL agent escapes safety-critical situations near these boundary states. For robust performance testing, we create a diverse set of traces via fuzz testing. These fuzz traces are used to bring the agent into a wide variety of potentially unknown states from which the average performance of the agent is compared to the average performance of the fuzz traces. We apply our search-based testing approach on RL for Nintendo's Super Mario Bros
Originalspracheenglisch
TitelThirty-First International Joint Conference on Artificial Intelligence (IJCAI 2022)
Redakteure/-innenLuc De Raedt
Herausgeber (Verlag)ijcai.org
Seiten503-510
ISBN (elektronisch) 978-1-956792-00-3
DOIs
PublikationsstatusVeröffentlicht - 2022
Veranstaltung31st International Joint Conference on Artificial Intelligence and the 25th European Conference on Artificial Intelligence: IJCAI-ECAI 2022 - Vienna, Österreich
Dauer: 23 Juli 202229 Juli 2022
Konferenznummer: 31
https://ijcai-22.org/

Konferenz

Konferenz31st International Joint Conference on Artificial Intelligence and the 25th European Conference on Artificial Intelligence
KurztitelIJCAI 2022
Land/GebietÖsterreich
OrtVienna
Zeitraum23/07/2229/07/22
Internetadresse

Fingerprint

Untersuchen Sie die Forschungsthemen von „Search-Based Testing of Reinforcement Learning“. Zusammen bilden sie einen einzigartigen Fingerprint.

Dieses zitieren