Trainify: A CEGAR-Driven Training and Verification Framework for Safe Deep Reinforcement Learning
Peng Jin, Jiaxu Tian, Dapeng Zhi, Xuejun Wen, Min Zhang
Abstract
Abstract Deep Reinforcement Learning (DRL) has demonstrated its strength in developing intelligent systems. These systems shall be formally guaranteed to be trustworthy when applied to safety-critical domains, which is typically achieved by formal verification performed after training. This train-then-verify process has two limits: (i) trained systems are difficult to formally verify due to their continuous and infinite state space and inexplicable AI components (i.e., deep neural networks), and (ii) the ex post facto detection of bugs increases both the time- and money-wise cost of training and deployment. In this paper, we propose a novel verification-in-the-loop training framework called Trainify for developing safe DRL systems driven by counterexample-guided abstraction and refinement. Specifically, Trainify trains a DRL system on a finite set of coarsely abstracted but efficiently verifiable state spaces. When verification fails, we refine the abstraction based on returned counterexamples and train again on the finer abstract states. The process is iterated until all predefined properties are verified against the trained system. We demonstrate the effectiveness of our framework on six classic control systems. The experimental results show that our framework yields more reliable DRL systems with provable guarantees without sacrificing system performance such as cumulative reward and robustness than conventional DRL approaches.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 9da97385-323b-4db6-92f3-6baabb680db6Cited by top-tier papers6
- Unifying Qualitative and Quantitative Safety Verification of DNN-Controlled SystemsDapeng Zhi, Peixin Wang, Si Liu, C.-H. Luke Ong et al.CAV 2024 · 11 citations
- Robustness Verification of Deep Reinforcement Learning Based Control Systems Using Reward MartingalesDapeng Zhi, Peixin Wang, Cheng Chen, Min ZhangAAAI 2024 · 4 citations
- Boosting Verification of Deep Reinforcement Learning via Piece-Wise Linear Decision Neural NetworksJiaxu Tian, Dapeng Zhi, Si Liu, Peixin Wang et al.NeurIPS 2023 · 4 citations
- Deductive Synthesis of Reinforcement Learning Agents for Infinite Horizon TasksYuning Wang, He ZhuCAV 2025
- Automating the Refinement of Reinforcement Learning SpecificationsTanmay Ambadkar, Djordje Zikelic, Abhinav VermaICLR 2026
Related papers
- An Iterative Scheme of Safe Reinforcement Learning for Nonlinear Systems via Barrier Certificate GenerationZhengfeng Yang, Yidan Zhang, Wang Lin, Xia Zeng et al.CAV 2021 · 15 citations
- An Abstraction-Based Framework for Neural Network VerificationYizhak Yisrael Elboher, Justin Gottschlich, Guy KatzCAV 2020 · 97 citations
- Learning Contract Invariants Using Reinforcement LearningJunrui Liu, Yanju Chen, Bryan Tan, Isil Dillig et al.ASE 2022 · 17 citations
- Safe DNN-type Controller Synthesis for Nonlinear Systems via Meta Reinforcement LearningHanrui Zhao, Xia Zeng, Niuniu Qi, Zhengfeng Yang et al.DAC 2023 · 4 citations
- Verifying learning-augmented systemsTomer Eliyahu, Yafim Kazak, Guy Katz, Michael SchapiraSIGCOMM 2021 · 45 citations
