2026· IEEE Transactions on Network and Service Management· Vol 23, pp. 7380-7394· 0 citations· 38 references
TL;DR
CARSO (Counterfactual Autoscaling and Resource-efficient Service Orchestration), a proactive and interpretable framework that integrates eXplainable Artificial Intelligence (XAI) into the autoscaling process that outperforms state-of-the-art proactive autoscaling frameworks in both QoS compliance and overall resource utilization.
Abstract
Cloud–edge computing enables scalable and resilient deployment of microservice-based applications, however achieving resource efficiency while ensuring stringent Quality of Service (QoS) remains challenging. The strong interdependencies among microservices and non-linear latency effects near resource saturation render conventional workload-driven autoscaling ineffective in complex distributed environments. This paper introduces CARSO (Counterfactual Autoscaling and Resource-efficient Service Orchestration), a proactive and interpretable framework that integrates eXplainable Artificial Intelligence (XAI) into the autoscaling process. CARSO employs counterfactual reasoning to derive minimal resource adjustments that proactively prevent QoS violations. The framework includes two core components: i) a Counterfactual Vertical Autoscaling (CVA) scheme that anticipates and mitigates performance degradation and ii) a Latency-Aware Resource Orchestration (LARO) policy that coordinates scaling and placement actions to balance resource efficiency and end-to-end latency across the cloud–edge continuum. Extensive experiments demonstrate that CARSO outperforms state-of-the-art proactive autoscaling frameworks in both QoS compliance and overall resource utilization.
Edge computing has become a key paradigm for supporting low-latency and high-concurrency services deployed with microservice architectures. However, limited device resources, highly dynamic workloads, and complex service dependencies make it difficult for existing autoscaling approaches—often relying on static threshol...
Tian-Yang Zheng, Peng-Fei Yang, Zhe Xu et al.· Proceedings of the Internati...· 0 citations
Financial institutions operating complex, mission-critical infrastructure face a persistent modernization dilemma: mainframe systems provide hardware-enforced transactional integrity, deterministic performance, and decades of regulatory validation, while cloud infrastructure offers elastic scalability, managed service...
Narasimha Rao Vanaparthi, M. Dhanekula, Rishabh Srivastava· International journal of com...· 0 citations
In this paper, we examine the Cloud-HPC continuum from the perspective of resource allocation and the execution of indistinguishable HPC and Cloud workloads. After briefly analyzing Kubernetes, an open-source platform for orchestrating containerized applications, we identified inefficiencies in resource usage resulting...
Tarek Menouer, Christophe Cérin, C. Hernández· IEEE Transactions on Paralle...· 0 citations
Managing resources across IoT, edge, and cloud layers calls for continuous, context-aware decisions under constraints that rarely stay fixed. Deep reinforcement learning (DRL) handles this class of problems well, and large language models (LLMs) are increasingly used to augment DRL pipelines, yet the architectural rela...
Antonino Vaccarella, Lan-Pei Li, Vincenzo Lomonaco et al.· 0 citations
The results show that soft SLO limits reduce corrective rescheduling actions by 49% compared to hard-limit approaches while maintaining acceptable performance guarantees, and resource-aware scheduling decreases node-level congestion and further mitigates SLO violations, demonstrating the effectiveness of incorporating...
Oliver Larsson, Thijs Metsch, Cristian Klein et al.· 0 citations
CERLA-SFC is introduced, a hierarchical, multi-objective orchestrator that unifies learning-based placement, topology-aware routing and event-driven resource allocation in a single control loop that maintains near-zero latency violations across all urgency classes while keeping end-to-end delay in the millisecond range...
Yuanfei Xiao, Zhenli He, Xiaolong Zhai et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.