Skip to content
Open access

Energy-Efficient Cooperative Data Offloading in Cellular Networks Using Reinforcement Learning

Aug 2026 · Journal of universal computer science (Online) · 0 citations · 30 references

TL;DR

This research is among the first to employ MARL to this extent, and it offers an end-to-end solution that combines cellular, Wi-Fi, and device-to-device (D2D) communications and considers practical network environments like user mobility and channel conditions.

Abstract

To address the growing need for wireless communications energy efficiency, this paper proposes a new multi-agent reinforcement learning (MARL) approach to cooperative data offloading in heterogeneous cellular networks. This research is among the first to employ MARL to this extent, and it offers an end-to-end solution that combines cellular, Wi-Fi, and device-to-device (D2D) communications and considers practical network environments like user mobility and channel conditions. We formulate the offloading problem as a Markov Decision Process (MDP) with correct models of energy consumption and network conditions. The deep Q-network (DQN)-based MARL algorithm allows user equipment (UEs) to learn collaborative strategies for optimizing overall energy consumption and timely offloading of data. Simulations compare MARL against greedy, random, and independent Q-learning baselines in low and high mobility regimes. Experiments show that MARL saves energy by as much as 40% over random offloading and 16.6% over greedy offloading, as well as enhancing average delay, throughput, and fairness. Convergence of the learning rate of the algorithm is within 1000 episodes, and sensitivity analyses confirm its performance across a range of user density and data size settings. Furthermore, the MARL framework accommodates dynamic network conditions and provides an adaptable solution for network operators to maximize performance and sustainability for existing and future wireless networks. The suggested MARL framework still performs better in terms of delay, throughput, and fairness, while at the same time catapulting energy savings over the greedy offloading approach by 9.4% to 12.8% under the very realistic 3GPP-compliant HARQ, adaptive modulation, standardized power control, and urban SLAW mobility models settings. With the extended state/action spaces and realistic 3GPP+SLAW conditions, MARL has an energy savings of 11.7%–14.2% over greedy offloading while still having the leading delay, throughput, and fairness.

Read PDF

Similar papers

Aug 2026

Dual Factors Analysis for Energy Efficient Routing in Cognitive Radio Networks: Lightweight DRL-WsOA

Deep Q-learning with a rapidly converging local search method based on permutation-equivariant neural networks for unseen environments in the given network scenarios is incorporated to ensure faster convergence and a minimal memory footprint.

M. Lakshmi, Arram. Mahesh Babu · 0 citations
#edge computing Sep 2026

M4O: A Novel Task Offloading Framework for High-Density High-Load VEC Networks

Vehicular edge computing (VEC), a key enabler for the Internet of Things (IoT) in intelligent transportation, addresses onboard processing constraints through collaborative task offloading among vehicles, facilitating latency-sensitive applications such as autonomous driving. However, developing efficient offloading strategies remains particularly challenging in high-density vehicular networks, where intensive computational demands coexist with severely constrained intervehicle communication ranges due to signal blockage. To handle this, we propose M4O, a mobility-aware task offloading framework supporting multihop, multiuser, and multitask offloading optimization. M4O intelligently integrates vehicle mobility patterns and enables relay-assisted offloading to enhance system effectiveness and robustness. The framework employs a dual-algorithm approach: the advantage actor–critic (A2C) for indivisible tasks and the hybrid proximal policy optimization (H-PPO) for divisible tasks, both optimized to minimize the temporally coupled composite cost of time and resources. Extensive experiments demonstrate that the deep reinforcement learning (DRL)-based solutions of M4O deliver stable and efficient offloading strategies, outperforming existing benchmarks by significant margins in cost efficiency. Our code is available at https://github.com/Zhouym1028/M4O

Momiao Zhou, Yimin Zhou, Yanshi Sun et al. · 0 citations
Open access 2026

Energy-Efficient Beamforming and Adaptive Computational Task Offloading in ISCC Systems

A nonconvex energy cost minimization problem is introduced by considering a user-specific energy cost ratio coefficient that explicitly balances UE-AP energy consumption according to heterogeneous device energy states and a double-loop framework combining successive convex approximation and alternating direction method of multipliers is developed.

Kai Dong, Lei Wang, S. Vorobyov et al. · 0 citations
Open access Jul 2026

MULTI-AGENT REINFORCEMENT LEARNING FOR TASK OFFLOADING AND RESOURCE ALLOCATION IN MEC SYSTEMS

This paper addresses the joint task offloading and resource allocation problem in multi-user MEC systems and proposes a decentralized control framework based on Multi-Agent Reinforcement Learning (MARL), which achieves lower total system cost and faster convergence than the full-local, full-offload, and heuristic baselines.

Youssef Oukissou, Mohamed Amine Meddaoui, Ayoub Belaidi et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.