Home / Journals / CMC / Online First / doi:10.32604/cmc.2026.084892
Special Issues
Table of Content

Open Access

ARTICLE

Cooperative Task Offloading in Mobile Edge Computing via an Improved MASAC Framework

Zheng Yao1, Jie Liu1, Changjun Deng2,3,*, Wang Lin2,3
1 School of Robotics and Automation, Hubei University of Automotive Technology, Shiyan, China
2 Yunnan International Joint Laboratory of Agricultural Remote Sensing and Digital Technology, Yunnan Hanzhe Technology Co., Ltd., Kunming, China
3 Key Laboratory of Plateau Agricultural Environment Monitoring and Control, Ministry of Agriculture and Rural Affairs, Kunming, China
* Corresponding Author: Changjun Deng. Email: email

Computers, Materials & Continua https://doi.org/10.32604/cmc.2026.084892

Received 06 May 2026; Accepted 30 June 2026; Published online 29 July 2026

Abstract

Mobile edge computing (MEC) is an effective paradigm for supporting latency-sensitive and computation-intensive intelligent applications. However, in dynamic mobile-edge network scenarios, mobile terminals experience time-varying wireless links due to mobility. Tasks may also arrive unpredictably, while multiple terminals compete for limited edge resources. As a result, MEC systems may suffer from service congestion and unbalanced resource utilization, which increases end-to-end latency and energy consumption. This paper investigates cooperative task offloading in dynamic MEC networks. The considered system comprises one macro base station and multiple small base stations equipped with edge-computing resources. In each time slot, each mobile terminal selects a service option, determines the task offloading ratio, and chooses its transmit power for task uploading. This sequential decision process is formulated as a multi-agent problem with continuous action spaces. Under the centralized training and decentralized execution (CTDE) framework, the problem is further modeled as a decentralized partially observable Markov decision process (Dec-POMDP). Standard multi-agent soft actor-critic (MASAC) is not fully suitable for this problem. Its original action model does not handle bounded continuous actions well. Its exploration strength may also be unsuitable at different training stages. Frequent policy updates can further make training unstable when critic estimates are inaccurate. To address these issues, this paper develops an adaptive Beta-policy and delayed-update multi-agent soft actor-critic method, abbreviated as ABDMASAC. This method uses a Beta policy to model bounded actions. It adjusts the entropy coefficient during training and delays policy updates to reduce training oscillations. Experimental results show that, under a unified training budget and a consistent evaluation protocol, the proposed method achieves a better overall trade-off than the selected MASAC-backbone and on-policy MARL baselines under the considered simulation settings in terms of overall reward, average end-to-end latency, and average energy consumption. In the large-scale scenario, compared with MASAC, it improves the overall reward by 17.8%, reduces the average end-to-end latency by 18.0%, and lowers the average energy consumption by 11.4%.

Keywords

Mobile edge computing; multi-agent reinforcement learning; MASAC; cooperative task offloading
  • 130

    View

  • 31

    Download

  • 0

    Like

Share Link