Skip to content

Multi-Critic Reinforcement Learning for Frame- and Tick-Rate Aware Satellite–Ground Integrated Heterogeneous Edge Networks

2026 · IEEE Transactions on Network Science and Engineering · Vol 13, pp. 11239-11254 · 0 citations · 40 references

Abstract

Satellite–ground integrated networks (SAGIN) enable wide-area support for latency-sensitive interactive applications such as UAV teleoperation and remote robotic control. Unlike conventional data services that focus on average latency or throughput, these applications operate in closed-loop feedback cycles in which perception frames and control actions must be synchronized and satisfy strict deadlines. Late-arriving data often become obsolete rather than recoverable. Moreover, practical edge servers execute tasks in discrete scheduling cycles, introducing tick-quantized completion times that directly affect deadline violations. This paper proposes a heterogeneous satellite–ground mobile edge computing framework in which terrestrial base stations (BSs) and satellites (SATs) both provide computing services with distinct capacity characteristics. We develop a server-side queuing and scheduling model that captures continuous frame generation, stochastic control inputs, buffer constraints, and discrete tick-based task completion. Based on this model, we formulate a multi-objective optimization problem that jointly determines user association, transmission power, and bandwidth allocation to minimize deadline violations, dropped tasks, and worst-case delay. To solve the resulting mixed-integer nonlinear program under dynamic satellite topology, we design a multi-critic reinforcement learning (RL) algorithm that decomposes synchronization, latency, and capacity constraints. Simulation results demonstrate substantial reductions in deadline violations and maximum delay compared with existing approaches, including proximal policy optimization (PPO) deep deterministic policy gradient (DDPG), greedy algorithm, proportional fairness scheme, random access strategy, and equal resource allocation method.

View source

Similar papers

Conference Aug 2026

Priority-Aware Hybrid Actor-Critic for Task Scheduling of UAV-Assisted Edge Computing

Unmanned Aerial Vehicle (UAV)-assisted Mobile Edge Computing (MEC) reduces network latency by leveraging the high mobility and flexibility of UAVs to provide on-demand computing services for heterogeneous user terminals (UTs). However, scheduling tasks across UAVs is challenging due to the diverse priority requirements...

Yun-Fei Chen, Quan-Xi Zhou, Wen-Can Mao et al. · 0 citations
Preprint Sep 2026

Dynamic Task and Resource Scheduling Towards Space-Air-Ground-Sea Integrated Network

In the context of 6G ubiquitous connectivity, the space-air-ground-sea integrated network (SAGSIN) emerges as a new paradigm for pervasive service provisioning. To support expanding maritime activities in infrastructure-scarce ocean areas, we propose an innovative dynamic task and resource scheduling approach for SAGSI...

Yu-Fei Ye, Shi-Jian Gao, Xin-Hu Zheng et al. · 0 citations
#edge computing Book Open access Oct 2026

Stability-Oriented Multi-Objective Container Scheduling for UAV-Assisted Edge Networks via TD3

MORA (Multi-Objective Resource Allocation), a stability-oriented deep reinforcement learning scheduler based on Twin Delayed Deep Deterministic Policy Gradient (TD3), with a normalized multi-objective reward that jointly optimizes energy, latency, SLA adherence, and proactively minimizes container migrations is propose...

Shabir Ahmad, F. Khan, Ibrar Ali Shah et al. · 0 citations
2026

Leveraging Lyapunov-Guided Contextual Bandits for Scalable Distributed Vehicular Access in Satellite–Terrestrial Integrated Networks

The integration of Low Earth Orbit (LEO) satellite with terrestrial Road Side Units (RSUs) offers a promising architecture for 6G-enabled Intelligent Transportation Systems by combining wide-area coverage with low-latency access for Connected Autonomous Vehicles. However, realizing efficient multi-vehicle cooperative a...

Zhan-Xi Ma, Jian-Zhe Xue, Zi-Da Zhang et al. · 0 citations
Open access Sep 2026

Energy-Aware Persistent Multi-UAV Coverage via Reinforcement Learning Guided by User Priority and Outage

In disaster response and other infrastructure-limited settings, UAV-mounted access points can rapidly restore service availability for mobile ground users as demand and fleet availability evolve. Existing single-slot coverage formulations, however, can mask prolonged individual outages and do not jointly represent hete...

Hao-Yu Mei, Cheng-Tao Xu, Ruo-Zhe Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.