Trainify: A CEGAR-Driven Training and Verification Framework for Safe Deep Reinforcement Learning
Peng Jin, Jiaxu Tian, Dapeng Zhi, Xuejun Wen, Min Zhang
摘要
Abstract Deep Reinforcement Learning (DRL) has demonstrated its strength in developing intelligent systems. These systems shall be formally guaranteed to be trustworthy when applied to safety-critical domains, which is typically achieved by formal verification performed after training. This train-then-verify process has two limits: (i) trained systems are difficult to formally verify due to their continuous and infinite state space and inexplicable AI components (i.e., deep neural networks), and (ii) the ex post facto detection of bugs increases both the time- and money-wise cost of training and deployment. In this paper, we propose a novel verification-in-the-loop training framework called Trainify for developing safe DRL systems driven by counterexample-guided abstraction and refinement. Specifically, Trainify trains a DRL system on a finite set of coarsely abstracted but efficiently verifiable state spaces. When verification fails, we refine the abstraction based on returned counterexamples and train again on the finer abstract states. The process is iterated until all predefined properties are verified against the trained system. We demonstrate the effectiveness of our framework on six classic control systems. The experimental results show that our framework yields more reliable DRL systems with provable guarantees without sacrificing system performance such as cumulative reward and robustness than conventional DRL approaches.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper6
- Unifying Qualitative and Quantitative Safety Verification of DNN-Controlled SystemsDapeng Zhi, Peixin Wang, Si Liu, C.-H. Luke Ong 等CAV 2024 · 被引用 11 次
- Robustness Verification of Deep Reinforcement Learning Based Control Systems Using Reward MartingalesDapeng Zhi, Peixin Wang, Cheng Chen, Min ZhangAAAI 2024 · 被引用 4 次
- Boosting Verification of Deep Reinforcement Learning via Piece-Wise Linear Decision Neural NetworksJiaxu Tian, Dapeng Zhi, Si Liu, Peixin Wang 等NeurIPS 2023 · 被引用 4 次
- Deductive Synthesis of Reinforcement Learning Agents for Infinite Horizon TasksYuning Wang, He ZhuCAV 2025
- Automating the Refinement of Reinforcement Learning SpecificationsTanmay Ambadkar, Djordje Zikelic, Abhinav VermaICLR 2026
相关 Paper
- An Iterative Scheme of Safe Reinforcement Learning for Nonlinear Systems via Barrier Certificate GenerationZhengfeng Yang, Yidan Zhang, Wang Lin, Xia Zeng 等CAV 2021 · 被引用 15 次
- An Abstraction-Based Framework for Neural Network VerificationYizhak Yisrael Elboher, Justin Gottschlich, Guy KatzCAV 2020 · 被引用 97 次
- Learning Contract Invariants Using Reinforcement LearningJunrui Liu, Yanju Chen, Bryan Tan, Isil Dillig 等ASE 2022 · 被引用 17 次
- Safe DNN-type Controller Synthesis for Nonlinear Systems via Meta Reinforcement LearningHanrui Zhao, Xia Zeng, Niuniu Qi, Zhengfeng Yang 等DAC 2023 · 被引用 4 次
- Verifying learning-augmented systemsTomer Eliyahu, Yafim Kazak, Guy Katz, Michael SchapiraSIGCOMM 2021 · 被引用 45 次
