Skip to content
Review Open access

A Comprehensive Review of Reinforcement Learning for Autonomous Driving in the CARLA Simulator

Sep 2025 · IEEE Access · Vol 14, pp. 132298-132330 · 12 citations · 129 references
Computer Science

TL;DR

A comprehensive review of approximately 100 peer-reviewed studies that apply RL in the CARLA simulator shows that model-free RL overwhelmingly dominates the field, accounting for more than 80% of existing studies, with DQN, PPO, and SAC being the most frequently adopted algorithms.

Abstract

Reinforcement learning (RL) has become an increasingly important framework for autonomous driving, while the CARLA simulator has emerged as a leading benchmark environment for training and evaluating RL-based driving agents. Despite rapid growth in this area, the literature remains fragmented, making it difficult to identify prevailing methods, experimental practices, and open challenges. This paper presents a comprehensive review of approximately 100 peer-reviewed studies that apply RL in the CARLA simulator. The surveyed works are organized into major methodological categories, including model-free, model-based, hierarchical, hybrid, and other specialized RL approaches. Our analysis shows that model-free RL overwhelmingly dominates the field, accounting for more than 80% of existing studies, with DQN, PPO, and SAC being the most frequently adopted algorithms. We also examine how these studies formulate driving problems through different state representations, action spaces, and reward designs, and we summarize the evaluation landscape in terms of metrics, towns, scenarios, and traffic configurations. Finally, we highlight persistent research challenges such as sparse reward design, generalization, sim-to-real transfer, safety, and limited behavioral diversity, and we discuss emerging directions that may help address these limitations. This review provides a structured reference for researchers entering the field and offers a foundation for future advances in RL-based autonomous driving in CARLA.

Read PDF

Similar papers

#reinforcement learning Review Open access Sep 2026

Recent Advances of Reinforcement Learning Algorithms for Autonomous Driving System

This survey examines RL-based AD in modular and end-to-end pipelines and relates reported methods to task formulation and deployment evidence and examines deployment barriers, including safety, Sim2Real generalization, data efficiency, computation, embodied alignment, and evaluation readiness.

B. Shuai, Min Hua, Le-Tian Tao et al. · 0 citations
Review Open access Aug 2026

Reinforcement Learning for Multimodal Foundation Models: A Survey

A coherent map of the rapidly expanding landscape of visual RL is provided to provide researchers and practitioners with a coherent map of the rapidly expanding landscape of visual RL and to highlight promising directions for future inquiry.

Weijia Wu, Chen Gao, Joya Chen et al. · 0 citations
Preprint Aug 2026

Scaling Curriculum Learning For Autonomous Driving

CL4AD is presented, the first integration of curriculum learning into batched autonomous driving simulators by framing scenario selection as an unsupervised environment design problem, and utility functions that shape curricula based on success rates and the realism of the agent's behavior are introduced, in addition t...

Cevahir Koprulu, D. Paz, Feng Tao et al. · 1 citation
Review Open access Sep 2026

Reinforcement Learning for Real-Time Control Using Quanser Platforms: A Structured Narrative Review

Reinforcement learning (RL) is increasingly used for real-time control of complex dynamical systems, but its practical performance must be evaluated under hardware constraints that are often simplified in simulation. This paper presents a comprehensive review of published RL-based control studies using the Quanser Aero...

Ghulam E. Mustafa Abro, S. Memon, Jawad Tanveer · 0 citations
Nov 2026

Consequence Learning for Trajectory Planning in Autonomous Driving

Imitation learning (IL) teaches autonomous driving models what an expert does, but not why that action is safe or optimal. This critical gap arises because a single expert trajectory cannot illuminate the broader solution space: a complex performance landscape with multiple, distinct solutions and sharp “performance cl...

Yi-Xuan Fan, Yali Li, Shengjin Wang · 0 citations
Book Open access Aug 2026

Large Language Model (LLM) as an Excellent Reinforcement Learning Researcher in both Single-Agent and Multi-Agent Scenarios

A Self-Evolutional single-agent/multi-agent Reinforcement Learning (SE-RL) framework that utilizes a Large Language Model (LLM) to design various RL algorithm modules, such as agent model design, reward function, profiling, communication, and state imagination, by leveraging the LLM generating module output or code.

Vincent Fu, Xin-Xin Xu, Weichen Xu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.