To Kill a Mutant: An Empirical Study of Mutation Testing Kills
Hang Du, Vijay Krishna Palepu, James A. Jones
Abstract
Mutation testing has been used and studied for over four decades as a method to assess the strength of a test suite. This technique adds an artificial bug (i.e., a mutation) to a program to produce a mutant, and the test suite is run to determine if any of its test cases are sufficient to detect this mutation (i.e., kill the mutant). In this situation, a test case that fails is the one that kills the mutant. However, little is known about the nature of these kills. In this paper, we present an empirical study that investigates the nature of these kills. We seek to answer questions, such as: How are test cases failing so that they contribute to mutant kills? How many test cases fail for each killed mutant, given that only a single failure is required to kill that mutant? How do program crashes contribute to kills, and what are the origins and nature of these crashes? We found several revealing results across all subjects, including the substantial contribution of "crashes" to test failures leading to mutant kills, the existence of diverse causes for test failures even for a single mutation, and the specific types of exceptions that commonly instigate crashes. We posit that this study and its results should likely be taken into account for practitioners in their use of mutation testing and interpretation of its mutation score, and for researchers who study and leverage mutation testing in their future work.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get b88b425c-6db2-4f5c-b864-4c6b36c8e1ddCited by top-tier papers3
- Ripples of a Mutation - An Empirical Study of Propagation Effects in Mutation TestingHang Du, Vijay Krishna Palepu, James A. JonesICSE 2024 · 6 citations
- Leveraging Propagated Infection to Crossfire MutantsHang Du, Vijay Krishna Palepu, James A. JonesICSE 2025 · 1 citation
- State Field Coverage: A Metric for Oracle QualityFacundo Molina, Nazareno Aguirre, Alessandra GorlaASE 2025 · 1 citation
Related papers
- LLMutantKiller: Using Large Language Models to Generate Tests That Kill MutantsFarideh Khalili, Aidan Domondon, Harshit Garg, Frank TipISSTA 2026
- How Does Killing Surviving Mutants Help Detect Real Bugs with Assertion Generation? A Controlled ExperimentHang Du, Vijay Krishna Palepu, James A. JonesISSTA 2026
- FlakiMe: Laboratory-Controlled Test Flakiness Impact AssessmentMaxime Cordy, Renaud Rwemalika, Adriano Franci, Mike Papadakis et al.ICSE 2022 · 14 citations
- Cost measures matter for mutation testing study validityGiovani Guizzo, Federica Sarro, Mark HarmanFSE 2020 · 11 citations
- Systematic Assessment of Fuzzers using Mutation AnalysisPhilipp Görz, Björn Mathis, Keno Hassler, Emre Güler et al.USENIX Security 2023
