Skip to content

IRS-Assisted UAV Secure Communication via Agentic AI

2026 · IEEE Transactions on Cognitive Communications and Networking · Vol 12, pp. 10311-10326 · 0 citations · 48 references

Abstract

Low-altitude uncrewed aerial vehicle (UAV) communication offers notable advantages over terrestrial base stations in terms of flexibility and deployment efficiency. However, the high likelihood of line-of-sight (LoS) propagation renders the communication links between UAVs and ground users (GUs) particularly susceptible to eavesdropping. To address this issue, we consider an intelligent reflecting surface (IRS)-assisted low-altitude UAV secure communication system, in which communication security is strengthened through adaptive control of the wireless propagation environment, even when eavesdroppers are present. We aim to maximize the secrecy rate of GUs while minimizing the UAV energy consumption by jointly optimizing the continuous UAV trajectory, power allocation, and discrete IRS phase shifts. Considering the dynamic, non-convex, and NP-hard nature of the optimization problem, we propose an agentic artificial intelligence (AI) approach, namely alternating optimization (AO) and generative diffusion model-based deep deterministic policy gradient (AO-GDMDDPG) approach. The proposed agentic AI approach is composed of two cooperative agents that operate over a hybrid and high-dimensional decision space, in which the UAV agent adopts a generative AI (GenAI)-enhanced deep reinforcement learning (DRL) method to optimize continuous decision variables, whereas the IRS agent relies on the AO method to determine discrete IRS phase shifts. Simulation results demonstrate the superiority of the AO-GDMDDPG approach over benchmark algorithms with respect to secrecy rate improvement and UAV energy consumption reduction.

View source

Similar papers

2026

Sensing-Then-ISAC: A Distance-Constrained Safe Reinforcement Learning for UAV Secure Communications

Integrated sensing and communication (ISAC) technology, when deployed on unmanned aerial vehicles (UAVs), enables aerial base stations to simultaneously provide wireless connectivity to ground users and perform environmental sensing through echo signal analysis. However, the broadcast nature of wireless transmission, combined with the line-of-sight (LoS) propagation characteristics of UAVs, increases the risk of passive eavesdropping on transmitted signals during ISAC missions. This paper investigates the joint trajectory design and power allocation (JTDPA) problem for UAV-enabled ISAC systems in environments with multiple mobile ground users and potential eavesdroppers. The proposed approach formulates the optimization problem as a constrained Markov decision process (CMDP), aiming to balance communication rate, secrecy rate, and energy consumption. To address the limitations of existing secure trajectory designs, such as unnecessary energy expenditure and overly conservative avoidance actions, we propose a two-stage (TS) strategy that incorporates the safe twin delayed deep deterministic policy gradient (Safe-TD3) algorithm, referred to as TS-SafeTD3. In the first stage (sensing stage), the UAV navigates toward a user-centric location without communication to enhance initial coverage efficiency, while satisfying the minimum-distance safety constraints with respect to potential eavesdroppers.In the second stage (ISAC stage), Safe-TD3 is employed to jointly optimize both trajectory and power allocation under the same safety constraints to maximize the weighted secrecy rate. Simulation results indicate that the proposed algorithm improves the weighted secrecy rate and energy efficiency under various operational conditions, while maintaining a low violation probability of the safety constraints.

Yu-Jia Chen, Hai-Yan Huang, Ting-Wei Chen et al. · 0 citations
Preprint Aug 2026

Toward Secure Communications for a UAV Swarm with Movable Antennas in SAGIN: CKM-Enabled Multi-Agent Reinforcement Learning Framework

Space-air-ground integrated networks (SAGINs) can provide ubiquitous and reliable connectivity for unmanned aerial vehicles (UAVs). However, air-to-ground links, which are typically dominated by line-of-sight (LoS) propagation, are vulnerable to passive eavesdropping due to the broadcast nature of wireless channels. To enhance physical-layer security, we investigate a SAGIN-enabled secure downlink communication system in which UAVs select service links among satellite, aerial, and terrestrial networks while adjusting the positions of the movable antenna (MA) array to fully exploit connectivity and spatial degrees of freedom for improved secrecy communication performance. Specifically, we maximize the secrecy energy efficiency (SEE) of a UAV swarm by jointly optimizing the MA positions, UAV trajectories, and link selections, subject to UAV mobility, MA movement, and link connectivity constraints. To reduce the real-time channel state information (CSI) acquisition overhead, we propose a channel knowledge map (CKM)-assisted multi-agent reinforcement learning framework. Specifically, the CKM is first constructed from sparse channel measurements via Kriging interpolation and is then leveraged together with satellite ephemeris information to enable efficient storage and retrieval of CSI. To reduce the action-space dimensionality and computational complexity, we model the MA array using rigid-body kinematics and adjust its position through global rigid-body translation, thereby constructing a low-dimensional hybrid action space for the joint optimization decisions. To align local decisions with system-wide performance under system constraints, we design an individual-team collaborative reward mechanism and introduce action masks to enforce constraints on UAV mobility, collision avoidance, MA regions, and connectivity capacity.

Jiayang Wan, Yafei Wang, Jiawei Zhuang et al. · 0 citations
2026

ISAC Enabled Anti-UAV: Joint Beamforming and Trajectory Design for Multi-UAVs

The rapid proliferation of Uncrewed Aerial Vehicles (UAVs) introduces significant challenges to low-altitude airspace security, particularly from unauthorized intrusions. To address these vulnerabilities, Integrated Sensing and Communication (ISAC) has emerged as a key enabler for anti-UAV systems. However, existing studies focusing on cellular networks with fixed base stations are ill-suited for the continuous movement of target UAVs, thus failing to meet the dual demands of flexible sensing and reliable positioning. To address this, we propose an ISAC-enabled anti-UAV scheme solely based on cooperative UAVs. Specifically, we first derive the optimal transmit power under the constraint of space-air transmission outage probability tolerance. Subsequently, we deduce the sensing Fisher information matrix and Cramér-Rao Bound (CRB) by incorporating the position uncertainty of the target UAV. Then, we formulate a long-term CRB minimization problem to enhance cooperative sensing performance. To tackle this NP-hard problem, we design a robust optimization algorithm that jointly optimizes transmit-receive beamforming, association scheduling, and UAV trajectory, by transforming the structurally complex CRB matrix into a set of semi-definite constraints, and resolving the inherent position uncertainty. Numerical results demonstrate that our proposed algorithm outperforms representative algorithms in terms of sensing accuracy and robustness.

Xiaojie Wang, Lingfei Li, Zhaolong Ning et al. · 1 citation
Preprint Aug 2026

Resource Allocation for Secure Dual-UAV-Assisted ISAC System

This work investigates the secrecy performance of a dual-uncrewed aerial vehicle (UAV)-assisted secure ISAC system, and maximizes the average secrecy rate by optimizing user scheduling strategies, time allocation, transmit power, and UAV trajectories.

Hongjiang Lei, Jianshuo Geng, Ki-Hong Park et al. · 1 citation
2026

AoI-Aware UAV-Assisted Secure Status Updating: An Agentic AI-Enabled DRL Approach

The rapid expansion of real-time Internet of Things (IoT) applications has positioned uncrewed aerial vehicles (UAVs) as a promising solution for flexible and timely data collection in areas lacking robust infrastructure. This paper investigates a UAV-assisted secure status updating system, where a UAV serves as a mobile relay to forward status updating packets from ground devices (GDs) under the threat of a potential eavesdropper. To ensure information freshness and operational sustainability, we formulate a long-term stochastic optimization problem to minimize the cumulative average age-of-information (AoI) of all GDs and energy consumption of the UAV. The formulated optimization problem is an online mixed-integer non-linear programming problem, which involves the joint optimization of the flight speed, direction, and transmission power of the UAV as well as the binary scheduling indicator of GDs. To tackle the inherent non-convexity and complex spatial-temporal coupling, we propose an agentic artificial intelligence (AI)-enabled deep reinforcement learning (DRL) approach, named adaptive truncated quantile critics with large language models (LLM)-enabled state representation and reward function design (ATQC-L). Specifically, an adaptive truncated quantile mechanism is incorporated to mitigate distributional overestimation in dynamic environments. Furthermore, we leverage the reasoning capability of LLMs as an offline design-time agent to generate task-aware state representation and intrinsic reward functions. Simulation results demonstrate that the proposed ATQC-L algorithm outperforms representative DRL baselines in balancing information freshness and energy consumption of the UAV, while maintaining stable performance under different network scales, LLM backbones, truncation-parameter settings, imperfect eavesdropping channel state information, and mobile eavesdropping scenarios.

Chuang Zhang, Geng Sun, Jiahui Li et al. · 0 citations

Multi-Objective Optimization for Secure UAV-Assisted Data Collection via Intermittent Jamming

This paper forms a multi-objective optimization problem aimed at minimizing AoI and energy consumption while maximizing the eavesdropper’s Bit Error Rate by jointly optimizing UAV trajectories, time scheduling, and jamming parameters and develops an efficient iterative algorithm.

Xiujuan Zhang, Yujiao Han, Shiyu Wang et al. · 0 citations