Sep 2026· Transactions of the Institute of Measurement and Control· 0 citations· 20 references
TL;DR
This paper proposes a trajectory-predictive and attention-enhanced deep recurrent State-Action-Reward-State-Action (SARSA) algorithm (PA-DR-SARSA), which incorporates future motion information into on-policy reinforcement learning through a Gated Recurrent Unit–based trajectory prediction module that learns temporal motion patterns from historical observation sequences.
Abstract
Autonomous unmanned aerial vehicle path planning in rescue missions must cope with dynamically moving targets and obstacles, where decision-making based solely on instantaneous observations often becomes myopic and fails to anticipate future motion behaviors. To address this issue, this paper proposes a trajectory-predictive and attention-enhanced deep recurrent State-Action-Reward-State-Action (SARSA) algorithm (PA-DR-SARSA). This algorithm incorporates future motion information into on-policy reinforcement learning through a Gated Recurrent Unit–based trajectory prediction module that learns temporal motion patterns from historical observation sequences. Short-horizon trajectory forecasts are fused with current observations and attention-enhanced features to construct a prediction-enhanced decision representation for SARSA action-value evaluation. Furthermore, an attention mechanism is introduced to adaptively weight predictive and instantaneous features, enabling the agent to prioritize decision-relevant motion information. Moreover, a risk-aware reward shaping strategy leverages these predicted trajectories to guide proactive action evaluation under dynamic uncertainty. Simulation results in dynamic grid-based rescue environments demonstrate that, compared with representative planning and reinforcement learning baselines, the proposed algorithm reduces average path length by up to 13.2% and turning points by up to 39.3%, while improving dynamic obstacle avoidance and task success rates by up to 17.2% and 23.7%, respectively, without compromising on-policy learning stability.
The rapid growth of wireless devices and the emergence of dynamic traffic hotspots have increased the need for intelligent trajectory planning in unmanned aerial vehicle base stations (UAV-BSs) operating in complex urban environments, where conventional reactive methods relying only on current system states cannot anti...
Tariq, Zhuo-Xiu Wei, K. Shaukat et al.· Scientific Reports· 0 citations
These findings support maneuver phase as an interpretable context for prediction-horizon adaptation within the evaluated architecture, while limiting the conclusions to the investigated scenario and experimental setting.
George Protogeros, M. Roumeliotis· Electronics· 0 citations
A dynamic path planning method for low-altitude Unmanned Aerial Vehicles (UAVs) tailored for urban inspection missions and constrains the average response latency for high-priority emergency tasks to within 40 s even under 50 concurrent dynamic tasks is proposed.
Changqi Yang, Hongjie Hu, Yi Ai· Drones· 0 citations
A framework that fuses the output of Trajectron++, a neural network-based trajectory predictor, with extended Kalman filter (EKF)-based multiple trajectory candidates at a late stage indicates that EKF-based trajectory candidates can effectively complement neural trajectory prediction through learned fusion.
Seong-Jun Kim, Seung-Hyun Kong· Journal of Institute of Cont...· 0 citations
Results indicate that combining physical structure with learned residual correction provides a more accurate, physically consistent, and operationally interpretable approach for UAV trajectory forecasting.
Md Ashraful Islam, Stanley Förster, Tianxiong Zhang et al.· Scientific Reports· 0 citations
It is argued that progress will depend less on further algorithmic proliferation than on integrated, verifiable architectures that combine data-driven adaptation with model-based structure, standardized evaluation, and staged real-world assurance.
Weijun Wang, Ming-Jie Li, Bushuo Wang et al.· Journal of Marine Science an...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.