A comprehensive review of approximately 100 peer-reviewed studies that apply RL in the CARLA simulator shows that model-free RL overwhelmingly dominates the field, accounting for more than 80% of existing studies, with DQN, PPO, and SAC being the most frequently adopted algorithms.
Abstract
Reinforcement learning (RL) has become an increasingly important framework for autonomous driving, while the CARLA simulator has emerged as a leading benchmark environment for training and evaluating RL-based driving agents. Despite rapid growth in this area, the literature remains fragmented, making it difficult to identify prevailing methods, experimental practices, and open challenges. This paper presents a comprehensive review of approximately 100 peer-reviewed studies that apply RL in the CARLA simulator. The surveyed works are organized into major methodological categories, including model-free, model-based, hierarchical, hybrid, and other specialized RL approaches. Our analysis shows that model-free RL overwhelmingly dominates the field, accounting for more than 80% of existing studies, with DQN, PPO, and SAC being the most frequently adopted algorithms. We also examine how these studies formulate driving problems through different state representations, action spaces, and reward designs, and we summarize the evaluation landscape in terms of metrics, towns, scenarios, and traffic configurations. Finally, we highlight persistent research challenges such as sparse reward design, generalization, sim-to-real transfer, safety, and limited behavioral diversity, and we discuss emerging directions that may help address these limitations. This review provides a structured reference for researchers entering the field and offers a foundation for future advances in RL-based autonomous driving in CARLA.
This survey examines RL-based AD in modular and end-to-end pipelines and relates reported methods to task formulation and deployment evidence and examines deployment barriers, including safety, Sim2Real generalization, data efficiency, computation, embodied alignment, and evaluation readiness.
B. Shuai, Min Hua, Le-Tian Tao et al.· Communications in Transporta...· 0 citations
A coherent map of the rapidly expanding landscape of visual RL is provided to provide researchers and practitioners with a coherent map of the rapidly expanding landscape of visual RL and to highlight promising directions for future inquiry.
CL4AD is presented, the first integration of curriculum learning into batched autonomous driving simulators by framing scenario selection as an unsupervised environment design problem, and utility functions that shape curricula based on success rates and the realism of the agent's behavior are introduced, in addition t...
Cevahir Koprulu, D. Paz, Feng Tao et al.· 1 citation
Reinforcement learning (RL) is increasingly used for real-time control of complex dynamical systems, but its practical performance must be evaluated under hardware constraints that are often simplified in simulation. This paper presents a comprehensive review of published RL-based control studies using the Quanser Aero...
Ghulam E. Mustafa Abro, S. Memon, Jawad Tanveer· Electronics· 0 citations
Imitation learning (IL) teaches autonomous driving models what an expert does, but not why that action is safe or optimal. This critical gap arises because a single expert trajectory cannot illuminate the broader solution space: a complex performance landscape with multiple, distinct solutions and sharp “performance cl...
A Self-Evolutional single-agent/multi-agent Reinforcement Learning (SE-RL) framework that utilizes a Large Language Model (LLM) to design various RL algorithm modules, such as agent model design, reward function, profiling, communication, and state imagination, by leveraging the LLM generating module output or code.
Vincent Fu, Xin-Xin Xu, Weichen Xu et al.· Proceedings of the 32nd ACM...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.