Skip to content

TAWJEEH: An Integrated Deep Reinforcement Learning and Heuristic Framework for Cellular Vehicle-to-Everything Enabled Real-Time Multidepot Vehicle Routing Optimization in Urban Logistics

Aug 2026 · Transportation Research Record · 0 citations · 24 references

TL;DR

TAWJEEH is introduced, a novel hybrid framework that integrates deep reinforcement learning, classical heuristics, and cellular vehicle-to-everything communications for time-dependent MDCVRP optimization and proves to be a robust, scalable, and computationally efficient solution.

Abstract

Urban logistics are increasingly strained by dynamic traffic conditions and complex operational constraints, rendering traditional optimization methods for the multidepot capacitated vehicle routing problem inadequate. This paper introduces TAWJEEH, a novel hybrid framework that integrates deep reinforcement learning, classical heuristics, and cellular vehicle-to-everything (C-V2X) communications for time-dependent MDCVRP optimization. The framework employs a deep Q network to learn adaptive policies for customer-to-vehicle assignment, complemented by clustering algorithms for initial customer grouping and heuristics for route refinement. Leveraging real-time data streams from C-V2X messages, TAWJEEH dynamically adjusts routes in response to live traffic conditions. Extensive and realistic simulations using SUMO on Hamburg and Luxembourg road networks validate our approach. The performance of TAWJEEH is benchmarked against the Clarke-Wright savings (CWS) heuristic and ant colony optimization (ACO). Results show significant and consistent reductions in key performance metrics; for instance, in the large-scale Luxembourg scenario, TAWJEEH reduces total travel distance by up to 55.4% compared with CWS and 8.9% against ACO. These improvements translate to substantial reductions in cumulative travel time, fuel consumption, CO 2 emissions, and overall operational costs. TAWJEEH proves to be a robust, scalable, and computationally efficient solution, highlighting the potential of combining advanced artificial intelligence techniques with vehicular communication technologies to address complex urban logistics challenges.

View source

Similar papers

Conference Aug 2026

A real-time dynamic vehicle path optimization framework for urban logistics based on deep reinforcement learning

The study describes the urban vehicle routing problem as a Markov decision process, integrating fleet operations, dynamic traffic conditions, and constantly arriving customer orders from heterogeneous realtime data streams, and proposes a graph-based neural network architecture for capturing complex spatio-temporal dep...

Jinyan Wang, Hong-Juan Cong · 0 citations
#machine learning Preprint Aug 2026

JAMPR+/L2D: scalable neural heuristic for constrained vehicle routing problems in dynamic environment

It is shown that the JAMPR+/L2D model, proposed in to solve large CPDPTW problems can be adopted in the case of substantial changes of graph distance matrix, and generalizes well for tasks with simpler constraints (CVRP, VRPTW), for different problem sizes and for moderate changes in distance matrixes.

A. Soroka, A. Meshcheryakov · 0 citations
Open access Aug 2026

Application of Reinforcement Learning for Optimizing the Capacitated Vehicle Routing Problem

These findings demonstrate that reinforcement learning is a promising and scalable alternative to conventional heuristic and metaheuristic approaches for capacitated routing problems, particularly in dynamic logistics environments that require rapid and adaptive decision making.

Audrey Ariij Sya'imaa.HS, Hilda Azkiyah, Khandker Farid Uddin Ahmed · 0 citations
Open access Aug 2026

Quantum Federated Reinforcement Learning‐Based Traffic Offloading and Resource Allocation for RSMA‐Enabled Space–Air–Ground Integrated Networks

A Quantum Federated Reinforcement Learning (QFRL)‐based traffic offloading framework for RSMA‐enabled SAGINs is proposed, allowing distributed small cells to jointly optimize traffic offloading ratios, bandwidth allocation, RSMA power distribution, and UAV trajectory planning while satisfying stringent delay and reliab...

Ishan Budhiraja, Abhay Bansal, B. Unhelkar et al. · 0 citations
#edge computing Open access Sep 2026

Vehicle as a Service: Fuzzy Reward-Based Multi-Agent Deep Reinforcement Learning for Task Scheduling in Vehicular Edge Computing

A reinforcement learning-based VEC task scheduling approach that integrates a fuzzy reward mechanism with multi-agent proximal policy optimization (FRMPPO) that satisfies the real-time processing demands of perception tasks in VaaS scenarios is proposed.

Qiang-Qiang Jiang, Jia-Mei Jin, Xu Xin et al. · 0 citations
#federated learning Open access Sep 2026

MAQDRL: QuadTree-Based Multi-Agent Federated Deep Q-Learning for Collision Avoidance and Routing Optimization with MEC-Aware Communication Analysis

Experimental results demonstrate that the proposed reinforcement learning-based Multi-agent Federated Deep Q-Network model, which integrates federated learning with QuadTree spatial partitioning and indexing framework for collision detection, is a promising solution for advancing autonomous vehicle routing systems.

A. Raj, Anurag Sharma, K. Naik · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.