RLFTSim: Realistic and Controllable Multi-Agent Traffic Simulation via Reinforcement Learning Fine-Tuning
Ehsan Ahmadi, Hunter Schofield, Behzad Khamidehi, Fazel Arasteh, Jinjun Shan, Lili Mou, Dongfeng Bai, Kasra Rezaee
Abstract
Supervised open-loop training has been widely adopted for training traffic simulation models; however, it fails to capture the inherently dynamic, multi-agent interactions common in complex driving scenarios. We introduce RLFTSim, a reinforcement-learning-based fine-tuning framework that enhances scenario realism by aligning simulator rollouts with real-world data distributions and provides a method for distilling goal-conditioned controllability in scenario generation. We instantiate RLFTSim on top of a pre-trained simulation model, design a reward that balances fidelity and controllability, and perform comprehensive experiments on the Waymo Open Motion Dataset. Our results show improvements in realism, achieving state-of-the-art performance. Compared with other heuristic search-based fine-tuning methods, RLFTSim requires significantly fewer samples due to a proposed low-variance and dense reward signal, and it directly addresses the realism alignment issue by design. We also demonstrate the effectiveness of our approach for distilling traffic simulation controllability through goal conditioning. The project page is available at https://ehsan-ami.github.io/rlftsim.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5e92105e-33ea-4b48-9a3a-564cae0bc11fBuilds on7
- Large Scale Interactive Motion Forecasting for Autonomous Driving : The Waymo Open Motion DatasetScott Ettinger, Shuyang Cheng, Benjamin Caine, Chenxi Liu et al.ICCV 2021 · 817 citations
- SMART: Scalable Multi-agent Real-time Motion Generation via Next-token PredictionWei Wu, Xiaoxin Feng, Ziyan Gao, Yuheng KanNeurIPS 2024 · 104 citations
- BehaviorGPT: Smart Agent Simulation for Autonomous Driving with Next-Patch PredictionZikang Zhou, Haibo Hu, Xinhong Chen, Jianping Wang et al.NeurIPS 2024 · 73 citations
- Trajeglish: Traffic Modeling as Next-Token PredictionJonah Philion, Xue Bin Peng, Sanja FidlerICLR 2024 · 61 citations
- TrafficSim: Learning To Simulate Realistic Multi-Agent BehaviorsSimon Suo, Sebastian Regalado, Sergio Casas, Raquel UrtasunCVPR 2021
Related papers
- Advancing Multi-agent Traffic Simulation via R1-Style Reinforcement Fine-TuningMuleilan Pei, Shaoshuai Shi, Shaojie ShenICLR 2026 · 21 citations
- SceneDiffuser: Efficient and Controllable Driving Simulation Initialization and RolloutChiyu Max Jiang, Yijing Bai, Andre Cornman, Christopher Davis et al.NeurIPS 2024 · 76 citations
- LangTraj: Diffusion Model and Dataset for Language-Conditioned Trajectory SimulationWei-Jer Chang, Wei Zhan, Masayoshi Tomizuka, Manmohan Chandraker et al.ICCV 2025 · 5 citations
- Closed-Loop Supervised Fine-Tuning of Tokenized Traffic ModelsZhejun Zhang, Péter Karkus, Maximilian Igl, Wenhao Ding et al.CVPR 2025
- Transferring Causal Driving Patterns for Generalizable Traffic Simulation with Diffusion-Based DistillationYuhang Chen, Jie Sun, Jialin Fan, Jian SunAAAI 2026
