Skip to content

Resource Allocation and Container Scaling for Microservices in Multi-Cluster Edge Computing System

2026 · IEEE Transactions on Network and Service Management · Vol 23, pp. 6244-6260 · 0 citations · 48 references
Computer Science

Abstract

With the advent of the 6G era and the evolution of distributed systems, edge computing has become a pivotal architecture for deploying latency-sensitive, resource-efficient applications. In particular, the microservice architecture, characterized by modular and loosely coupled components, has gained significant traction for building scalable and maintainable applications at the network edge. However, deploying microservice-based applications in heterogeneous and geographically distributed Multi-Cluster Edge Computing (MCEC) environments presents critical challenges, especially in achieving efficient and scalable resource management. Although existing research has explored resource allocation and container scaling for microservice-based systems, most prior works consider container efficiency in isolation or within single-cluster or cloud-centric environments, without jointly addressing container-level efficiency, inter-cluster task offloading, and resource allocation in MCEC scenarios. To address this gap, we propose RACCOON, a request-offloading cascaded resource allocation algorithm tailored for microservice-oriented deployments in MCEC settings. RACCOON aims to minimize user-perceived service latency while optimizing overall resource utilization. Complementing this, we introduce RASCAL, a reinforcement learning (RL)-based container scaling mechanism that dynamically adjusts resource provisioning at the container level to further enhance system performance. Experimental evaluation shows that our approach consistently outperforms methods that address only resource allocation, only task offloading, or only container scaling, by jointly optimizing these dimensions to reduce end-to-end user-perceived latency and computational overhead.

View source

Similar papers

Open access Jul 2026

A Hybrid Framework for Joint Optimization of Resource Allocation and Load Balancing in Cloud Systems

Experiments show that the proposed Hybrid Framework for Joint Optimization of Resource Allocation and Load Balancing that spans two layers in heterogeneous cloud computing systems obtains 25-30% energy savings compared with ordinary methods, significantly reduces p95 latency and also achieves a relatively better Quality Of Service.

Eram Fatma, Nidhi Mishra, Mohammed Abdul Bari · 0 citations
Aug 2026

Multi-application operator placement in cloud-edge infrastructure for big data stream processing

A resource-aware multi-application operator placement method that optimizes both end-to-end latency and network usage, while meeting QoS constraints and application owners’ preferences in heterogeneous cloud-edge environments is proposed.

Simin Ghasemi-Falavarjani, B. S. Ghahfarokhi, M. Nematbakhsh et al. · 1 citation
Jul 2026

Internet of Things-Centric Optimized Service Provisioning in Multi-Cloud Environment

A lightweight, QoS-aware service placement algorithm that evaluates latency, bandwidth, and node load in real time is introduced that yields reduced latency and more consistent wait times relative to heuristic and genetic baselines.

Anshul Atre, K. Singh, Brijesh Kumar Chaurasia et al. · 0 citations
Conference Jul 2026

ELTOS: Energy-Latency Trade-Off Optimization Strategy for Microservice Placement in Edge Environments

The deployment of microservices in edge environments is increasingly critical to support latency-sensitive and data-intensive applications such as IoT analytics, real-time monitoring, and smart city services. Edge infrastructures are highly heterogeneous in terms of compute capacity, communication latency, and energy efficiency. Furthermore, the geographical distance between edge nodes may introduce non-negligible energy consumption for data transfers, which is often overlooked in traditional placement strategies. While prior works focus on minimizing response time, there is a need to adapt and extend such strategies for energy efficiency in edge computing. This paper proposes an Energy-Latency Trade-off Optimization Strategy (ELTOS) for microservice placement in edge environments. ELTOS formulates and solves a cost-based optimization for microservice placement on heterogeneous edge servers. The main goal of this optimization algorithm is to minimize the transmitted data size between microservices on different edge servers, saving energy while ensuring a suitable level of end-to-end latency for each request. Experimental evaluations using real edge computing infrastructure demonstrate the efficiency of the proposed ELTOS compared with previous work. A comparison result of energy consumption shows that ELTOS outperforms the previous method in 70% of all pairwise comparisons; therefore, ELTOS consumes less energy while maintaining QoS performance level. ELTOS enhances the median, 95th, and 99th percentile response times on average by 6.03%, 7.47%, and 3.30%, respectively.

Ida Falco, A. K. Idrees, Carmine Colarusso et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.