Skip to content
Preprint

Graph-Based Safe Reinforcement Learning for Multi-Agent Systems with Time-Varying Topology

Sep 2026 · 0 citations · 35 references
Computer Science

TL;DR

A graph-based safe multi-agent reinforcement learning (MARL) framework for cooperative navigation with time-varying topology is presented, integrating a attention-based actor and a Graph Attention Network (GAT) centralized critic, enabling scale-insensitive policy learning under time-varying communication topologies.

Abstract

This paper presents a graph-based safe multi-agent reinforcement learning (MARL) framework for cooperative navigation with time-varying topology. To address the critical challenge of ensuring safety in environments with sensing constraints, a safety-decoupled mechanism is introduced through a Control Barrier-Like Function (CBLF) action screening layer. This mechanism bridges the gap between discrete LiDAR perception and continuous safety constraints, ensuring that physical safety constraints are strictly satisfied regardless of the learning progress. Building upon this safety foundation, a unified structural architecture is proposed, integrating a attention-based actor and a Graph Attention Network (GAT) centralized critic. The actor utilizes a value vector reconstruction mechanism that explicitly encodes relative geometric relations through a collaborative tracking error matrix, enabling scale-insensitive policy learning under time-varying communication topologies. Meanwhile, the GAT-based critic models evolving interaction structures for accurate global value estimation. The proposed framework is validated on real differential-drive robot platforms, and experimental results demonstrate superior stability and safety in dynamic scenarios with limited fields-of-view.

View source

Similar papers

Preprint Sep 2026

Fully Decentralized and Safety-Aware Multi-Agent Reinforcement Learning for Control on Networks

This paper develops a safe and fully decentralized multi-agent reinforcement learning (MARL) algorithm to solve a class of discrete-time control problems on networks, including the persistent monitoring problem. Fully decentralized control of agents, while offering numerous benefits, faces issues such as exponentially...

T. Rogalski, Shirantha Welikala · 0 citations
Conference Aug 2026

Heterogeneous Multi-Agent Autonomous Learning and Safe Cooperative Decision-Making

Unmanned surface and underwater vehicles face challenges in autonomously learning cooperative encirclement for high-value targets under partial observability, intermittent communication, and collision risks. This paper proposes a heterogeneous multi-agent reinforcement learning framework with safe decisionmaking. The f...

Jiang-Li Cao, Chao Liu, Guo-Ping Zhang · 0 citations
Conference Aug 2026

Risk-Informed Multi-Agent Reinforcement Learning for Embedded Systems on Resource-Constrained Hardware

Learning-enabled control systems increasingly rely on multi-agent reinforcement learning to operate in uncertain and interactive environments. While risk-aware decision-making has been shown to improve safety and robustness, deploying such algorithms on resource-constrained embedded platforms remains a significant chal...

Lachlan Talento, Kyle Pham, Bhaskar Ramasubramanian · 0 citations
Open access 2026

SAFE–MA–RRT: Data-Driven Safe Motion Planning for Multi-Agent Systems

This paper proposes a fully data-driven motion-planning framework for homogeneous linear multi-agent systems that operate in shared, obstacle-filled workspaces without access to explicit system models. Each agent independently learns its closed-loop behavior from experimental data by solving convex semidefinite program...

Babak Esmaeili, H. Modares · 0 citations
Open access Aug 2026

Modeling Dynamic Obstacle Avoidance Strategy of Drone Swarms Combined with Multi-Agent Reinforcement Learning

The proposed framework demonstrates robust scalability and real-time coordination capability for dynamic environments, while providing a reliable decision-making paradigm for intelligent multi-agent systems operating in communication-intensive and electromagnetically complex application scenarios.

X.-H. Fang, K. Chen, Cheng-Hao Ren et al. · 0 citations
#graph neural networks Open access Sep 2026

Graph Neural Network‐Based Reinforcement Learning for Decentralized Multi‐Robot Manipulation

A graph neural network (GNN)‐based framework for scalable multiagent reinforcement learning (RL), where each manipulator is represented as a node in a GNN, and message‐passing edges provide a communication mechanism that enables agents to share information effectively.

Tong Chen, Bo Fu, Dawn M. Tilbury et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.