A graph-based safe multi-agent reinforcement learning (MARL) framework for cooperative navigation with time-varying topology is presented, integrating a attention-based actor and a Graph Attention Network (GAT) centralized critic, enabling scale-insensitive policy learning under time-varying communication topologies.
Abstract
This paper presents a graph-based safe multi-agent reinforcement learning (MARL) framework for cooperative navigation with time-varying topology. To address the critical challenge of ensuring safety in environments with sensing constraints, a safety-decoupled mechanism is introduced through a Control Barrier-Like Function (CBLF) action screening layer. This mechanism bridges the gap between discrete LiDAR perception and continuous safety constraints, ensuring that physical safety constraints are strictly satisfied regardless of the learning progress. Building upon this safety foundation, a unified structural architecture is proposed, integrating a attention-based actor and a Graph Attention Network (GAT) centralized critic. The actor utilizes a value vector reconstruction mechanism that explicitly encodes relative geometric relations through a collaborative tracking error matrix, enabling scale-insensitive policy learning under time-varying communication topologies. Meanwhile, the GAT-based critic models evolving interaction structures for accurate global value estimation. The proposed framework is validated on real differential-drive robot platforms, and experimental results demonstrate superior stability and safety in dynamic scenarios with limited fields-of-view.
This paper develops a safe and fully decentralized multi-agent reinforcement learning (MARL) algorithm to solve a class of discrete-time control problems on networks, including the persistent monitoring problem. Fully decentralized control of agents, while offering numerous benefits, faces issues such as exponentially...
Unmanned surface and underwater vehicles face challenges in autonomously learning cooperative encirclement for high-value targets under partial observability, intermittent communication, and collision risks. This paper proposes a heterogeneous multi-agent reinforcement learning framework with safe decisionmaking. The f...
Learning-enabled control systems increasingly rely on multi-agent reinforcement learning to operate in uncertain and interactive environments. While risk-aware decision-making has been shown to improve safety and robustness, deploying such algorithms on resource-constrained embedded platforms remains a significant chal...
Lachlan Talento, Kyle Pham, Bhaskar Ramasubramanian· Conference on Control Techno...· 0 citations
This paper proposes a fully data-driven motion-planning framework for homogeneous linear multi-agent systems that operate in shared, obstacle-filled workspaces without access to explicit system models. Each agent independently learns its closed-loop behavior from experimental data by solving convex semidefinite program...
Babak Esmaeili, H. Modares· IEEE Open Journal of Control...· 0 citations
The proposed framework demonstrates robust scalability and real-time coordination capability for dynamic environments, while providing a reliable decision-making paradigm for intelligent multi-agent systems operating in communication-intensive and electromagnetically complex application scenarios.
X.-H. Fang, K. Chen, Cheng-Hao Ren et al.· Advanced Electromagnetics· 0 citations
A graph neural network (GNN)‐based framework for scalable multiagent reinforcement learning (RL), where each manipulator is represented as a node in a GNN, and message‐passing edges provide a communication mechanism that enables agents to share information effectively.
Tong Chen, Bo Fu, Dawn M. Tilbury et al.· Advanced Intelligent Systems· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.