cs.LGJul 8, 2026

Gimitest: A Comprehensive Tool for Testing Reinforcement Learning Policies

Authors: Dennis GrossQuentin MazouniHelge SpiekerArnaud Gotlieb

Abstract

Reinforcement learning (RL) policies can be unsafe and vulnerable to attacks. Ensuring their reliability is often a pain point as existing automated testing methods target only selected environments, testing scenarios, and RL algorithms. To address this, we propose a comprehensive framework for testing single- and multi-agent RL policies under varying conditions. Our implementation of this framework, Gimitest, is an open-source tool that supports various gym frameworks and allows for modifications of their integrated components. This article describes the framework and details Gimitest's functionality and architecture. It showcases its effectiveness in testing multiple RL policies in environments such as the official Farama Gymnasium and PettingZoo.

Explore similar work

CardsList
  1. Evaluating Fuzz Testing for Reinforcement Learning Agents

    Jul 27, 2026Zhibin Kang, Hanmo You, Dong Wang +2FuzzingCrashes