This work proposes HiGFRL, a Hierarchical Graph Fusion-Driven Reinforcement Learning framework, which designs a fusion-driven dual-network architecture to optimize RL decision-making and incorporates a topology-prior-guided hybrid reward mechanism that distills static topological priors into the learning process to accelerate convergence.
Abstract
Online scheduling of dependency-aware tasks in heterogeneous cloud clusters is a fundamental yet challenging problem due to the complex interplay between DAG topologies and multi-dimensional resource constraints. While DRL has shown promise, existing GNN-based approaches often struggle to efficiently model high-order topological dependencies and suffer from loose coupling between task and resource states, leading to myopic scheduling decisions. To address these limitations, we propose HiGFRL, a Hierarchical Graph Fusion-Driven Reinforcement Learning framework. HiGFRL constructs a novel three-level state representation comprising a Static Hypergraph, a Dynamic Global Graph, and a Local Bipartite Graph to explicitly model the interplay between task dependencies and real-time cluster dynamics. Specifically, we design a fusion-driven dual-network architecture to optimize RL decision-making, where a Context Fusion Allocator integrates local bipartite matching features with fused global context to execute precise task-to-node allocation, and a Global State Evaluator leverages the global dynamic graph representation to accurately estimate expected long-term cumulative reward. Furthermore, we incorporate a topology-prior-guided hybrid reward mechanism that distills static topological priors into the learning process to accelerate convergence. Extensive experiments using real-world Alibaba cluster traces demonstrate that HiGFRL significantly outperforms heuristics and DRL baselines. Specifically, in challenging large-scale high-load scenarios, HiGFRL reduces the Makespan by up to 32.55%, and optimizes the average task flow time and average task wait time by 13.58% and 13.79%, respectively. Experimental results confirm that HiGFRL not only significantly improves cluster throughput but also ensures superior QoS by substantially reducing queuing delays. Code Release:https://github.com/igeng/HiGFRL.
A Graph Attention-Driven Hierarchical Reinforcement Learning framework is developed and model the scheduling process as an event-driven hierarchical semi-Markov decision process (SMDP) that maintains competitive workflow success rate and generally achieves higher container utilization and lower energy consumption.
Zong-Jin Li, Shaohan Feng, Chun-Xi Yang et al.· 0 citations
Cloud computing has emerged as a new paradigm, which entrusts task scheduling to ensure the satisfaction of stringent constraints on latency, energy, and resources for sustainably running real-time applications. State-of-the-art natural DRL-based scheduling solutions mainly rely heavily on DRL techniques and are either...
Krishna Patwari, Raghvendra Kumar, J. Sastry· International Journal of Ele...· 0 citations
ReLA is an RL scheduler built on structured representation learning and aggregation that learns intra-entity representations using self-attention and convolution, captures inter-entity operation–machine interactions using cross-attention, and aggregates multi-scale representations for parallel actor-based scoring of fe...
Zheng-Yi Kwan, Wei Zhang, Aik Beng Ng et al.· Proceedings of the Internati...· 0 citations
Multi-Agent Reinforcement Learning (MARL) has emerged as a pivotal paradigm for complex decision-making in autonomous systems and air combat. While MARL has demonstrated significant potential in air combat, achieving sophisticated tactical coordination remains a non-trivial challenge. This difficulty is largely attribu...
Junlin Liu, Cheng-Wei Li, Yang Gao et al.· 0 citations
The spatial-temporal mismatch between generation and load is exacerbated by high distributed photovoltaic (PV) penetration in distribution service areas, causing power quality degradation and PV accommodation challenges. To tackle this issue, an end-to-end optimal scheduling method based on a heterogeneous graph attent...
Wei Zheng, Han Yan, Jin-Gang Qin et al.· Journal of Renewable and Sus...· 0 citations
Exploring how generative AI could make machine vision more accessible to businesses. The post GenEye in a Box: Making Machine Vision Something You Can Just Ask For appeared first on GPT-Lab.
MIT News · Artificial Intelligence· news.mit.eduOct 7, 2026
Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.
Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve them without rebuilding them. The post Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses appeared first on Microsoft Research.
MIT News · Artificial Intelligence· news.mit.eduOct 6, 2026