Skip to content
Preprint

Knowledge-Data-Dual-Driven Reinforcement Learning for Autonomous Vehicle Control in Mixed Traffic

Aug 2026 · 0 citations · 35 references
Computer Science Engineering

TL;DR

This work proposes Knowledge-Data Dual-driven Reinforcement Learning (KDDRL), a conditional deep generative model that effectively handles intention uncertainty, accelerates training convergence, and outperforms conventional baseline methods in terms of safety, efficiency, and comfort.

Abstract

In mixed traffic, decision-making for autonomous vehicles (AVs) confronts three interrelated challenges. First, physics-based priors incorporated into reinforcement learning (RL) models fail to capture latent interactive vehicle intentions and diverse driver behaviors, limiting the proactive reasoning capabilities. Second, abrupt maneuvers by surrounding vehicles cause non-stationarity, leaving long-tail safety events under-explored. Third, hybrid action spaces destabilize unified RL training due to the different temporal scales of continuous car-following and discrete lane-changing maneuvers. To address these issues, we propose Knowledge-Data Dual-driven Reinforcement Learning (KDDRL). First, a conditional deep generative model synthesizes intention-aware future trajectories, converting passive perception into proactive predictive states. Second, a knowledge-data dual-driven paradigm operates on these predictive states, fusing probabilistic data-driven insights with physical constraints to guide safe exploration through safety-critical scenarios. Third, a coupling module compresses both intention-aware trajectories and physical constraints into compact shared embeddings. This unified representation enables asynchronous multi-timescale optimization of continuous car-following and discrete lane-changing while preserving mutual information. Evaluations on dataset-calibrated simulations demonstrate that KDDRL effectively handles intention uncertainty, accelerates training convergence, and outperforms conventional baseline methods in terms of safety, efficiency, and comfort.

View source

Similar papers

Open access Aug 2026

Deep Reinforcement Learning for Communication-Free Distributed Control of Autonomous Vehicles in Unstructured Intersection.

A novel deep reinforcement learning framework that enables safe and efficient navigation in such communication-free, signal-free, and lane-free intersection environments while meeting stringent safety requirements for practical deployment is proposed.

Ruihao Zeng, M. Ramezani · 0 citations
Preprint Aug 2026

Offline Multi-Agent Reinforcement Learning with a Physics-Informed World Model for Cooperative Mixed Traffic Control

This study investigates cooperative control of connected and automated vehicles (CAVs) at partially observable highway bottlenecks in mixed traffic, aiming to mitigate congestion without relying on complete global traffic states or online trial-and-error. We propose a physics-informed world model-based offline multi-ag...

Lu Liu, Chi Xie, Xi Xiong · 0 citations
Open access Sep 2026

DKD-MARL: a data-knowledge dual-driven multi-agent reinforcement learning framework for traffic crash severity prediction.

Accurate prediction of traffic crash severity is critical for post-crash emergency response and proactive safety interventions. However, existing methods either rely on structured data-driven statistical learning and overlook domain semantic knowledge, or use single large language models (LLMs) suffering from unstable...

Jia-Zhao Zhang, Shuai Dai, Dan Zhao et al. · 0 citations
Open access Aug 2026

Rain-Aware Lane Change Decision Model for Autonomous Vehicle Using Deep Reinforcement Learning

Autonomous vehicles (AVs) have shown significant potential in recent years, with increasing interest from the public, industry, and academia in their adoption on roads. To achieve reliable autonomous driving, AVs must be able to operate safely under adverse weather conditions such as rain-induced wet roads which po...

A. Alzubaidi, R. Babu, Young-Ji Byon et al. · 1 citation
Open access Aug 2026

Interaction-Field Residual Learning for Calibrated Lane-Change Intent under Mixed Traffic

The prediction of lane-change intent in mixed traffic environments, characterized by the coexistence of human-driven vehicles and autonomous vehicles, presents a formidable challenge due to the stochastic nature of human driving behaviors and complex vehicular interactions. Conventional trajectory prediction and intent...

Anele Mthembu, Thabo Pretorius · 0 citations
Conference Aug 2026

A Survey of Motion Planning Methods for Autonomous Driving in Mixed Traffic

Autonomous driving systems inevitably operate in mixed traffic where autonomous vehicles coexist with humandriven vehicles (HDVs). Motion planning must therefore handle uncertain intent, heterogeneous driving styles, asymmetric responsibility, and behavioral adaptation to the ego vehicle’s actions. This survey reviews...

Bai Li, Jin-Di Hao, Xiao-Han Yang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.