Skip to content
Book Open access

EvoDriver: Novelty-Search Driven Evolution of Behavioral Test Suites for Autonomous Vehicles

Apr 2026 · SEAMS@ICSE · pp. 225-238 · 0 citations · 97 references
Computer Science

TL;DR

This paper introduces a bottom-up, exploratory approach to automatic test case generation for AVs, where test cases are evolved organically to reveal uncertainty posed by diverse and unexpected maneuvers from an external vehicle in the operating environment.

Abstract

In addition to environmental-based uncertainty conditions, self-adaptive Autonomous Vehicles (AVs) must be able to handle uncertainty posed by external human drivers sharing the operating environment. Existing search-based AV testing approaches have explored objective-oriented, top-down approaches to develop a range of test cases to assess whether the system satisfies functional and safety properties with respect to specific operational contexts. However, these techniques lack search pressure to identify test cases that may be obscured in unintuitive regions of the search space (i.e., edge cases). This paper introduces a bottom-up, exploratory approach to automatic test case generation for AVs, where test cases are evolved organically (i.e., in a more open-ended fashion) to reveal uncertainty posed by diverse and unexpected maneuvers from an external vehicle in the operating environment. Specifically, novelty search-based genetic programming is used to evolve controllers for an external vehicle that interacts with the AV under study in order to reveal the most diverse behavioral responses from the AV under study. By leveraging diversity as a metric through novelty search, our approach automatically discovers behaviors for an external vehicle that result in increasingly different interactions with the AV under study. Our approach overcomes challenges typically associated with traditional search-based AV testing, such as developer bias and high development costs of manually tuning search objectives. We demonstrate the efficacy of our proof-of-concept framework in several traffic scenarios, discovering diverse test cases that enable us to reveal undesirable behaviors in the AV under study, including many edge cases that might otherwise not be detected.

Read PDF

Similar papers

Open access Oct 2026

MG-Fuzz: Model-Guided Fuzzing for Unsafe Scenario Discovery in Autonomous Driving Systems

As autonomous driving systems (ADS) are increasingly deployed in real-world environments, discovering diverse unsafe driving scenarios remains a fundamental yet difficult problem. Existing scenario generation and testing approaches often rely on black-box exploration or externally-observed heuristic feedback, which str...

Yulong Lyu, Rui-Qi Hong, Jia-Wan Wang et al. · 0 citations
Open access Oct 2026

Are We Stuck? Modeling and Detecting Deadlocks in Multi-autonomous Vehicle Systems

Autonomous driving system (ADS) testing is essential to ensure the safety and reliability of autonomous vehicles (AVs) prior to deployment. As ADSs are increasingly deployed in multi-AV traffic environments, it becomes crucial to assess their cooperative performance, particularly with respect to deadlock, a fundamental...

Ming-Fei Cheng, Xiao-Fei Xie, Li-Li Quan et al. · 0 citations
#small language model Open access Aug 2026

GSMultiAgent: A Multi-Agent Loop Framework atop Hermes Agent for Intelligent Design of Guidance Systems

The strong coupling among guidance laws, control loops, aerodynamics, and mission constraints poses growing challenges to the design of modern tactical missile guidance systems for autonomous flight. Conventional manual tuning and simulation-based trial-and-error result in long iteration cycles, limited reuse of design...

Ji-Song Xiao, Chengwei Yang, Xiao Xu et al. · 0 citations
#machine learning Preprint Sep 2026

Probabilistic Modelling of Operational Design Domains, A New Approach for Testing AI Systems

The conventional testing process quickly fails when applied to ML-based systems such as obstacle detection in vehicles: if an obstacle is not detected in a test, classical bug fixing is impossible and an AI system will always retain shortcomings. Test results can therefore only be interpreted statistically, which in tu...

H. Wiesbrock · 0 citations
#artificial intelligence Preprint Sep 2026

When Successful Strategies Fail: Adaptation to Environmental Novelty in Terminal Agents

This work introduces AGNI, an automated pipeline that extracts trajectory-relevant assumptions, injects targeted environmental changes, and validates that the resulting novel tasks remain solvable and highlights a gap between task competence and adaptive capability and motivate environmental variation as a core dimensi...

Janvijay Singh, Vaishnavi Shrivastava, Dilek Hakkani-Tur et al. · 0 citations
#artificial intelligence Preprint Sep 2026

PlannerForge: LLM Agents for Scenario-Based Testing of Motion Planners in Autonomous Driving

Ensuring the safety of autonomous driving is a critical challenge. Scenario-based testing is a systematic process used to validate Autonomous Driving Systems (ADSs), but it remains a fragmented modular pipeline in which scenario generation, retrieval, modification, ADS execution, and results analysis are performed by s...

Yuan Gao, Sebastian Müller, Mattia Piccinini et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.