An Energy-Efficient Multi-Agent Reinforcement Learning Approach for Spark Job Scheduling in Mobile Edge Computing
DOI:
https://doi.org/10.54691/wr8mqp87Keywords:
Mobile Edge Computing; Energy Optimization; Apache Spark; Multi-agent Reinforcement Learning; Computation Offloading.Abstract
Battery life remains a major constraint on the continuous operation of mobile IoT devices represented by drones. Most existing computation offloading strategies rely on static heuristic rules, which fail to deliver stable performance in dynamically changing wireless mobile environments. Targeting this problem, this paper develops a multi-agent deep reinforcement learning approach for Spark job scheduling across local devices, edge servers and cloud resources. The scheduling task is modeled as a Markov decision process: six real-time system and network metrics, including CPU utilization, memory load, handover delay, signal strength, transmission latency and congestion level, form the state space, and the reward function is built directly on actual measured energy consumption data. Based on the multi-agent deep deterministic policy gradient (MADDPG) algorithm, the proposed method adopts a centralized training and decentralized execution paradigm to realize distributed collaborative decision-making. Validation on a public MEC dataset shows that the trained policy finally converges to a hybrid scheduling mode: 57% of tasks are processed locally, 39% are offloaded to edge nodes, and less than 4% are assigned to cloud resources for specific application scenarios. Comparative tests with PPO, FIFO, FAIR and HAS baselines confirm that multi-agent reinforcement learning can well capture the intrinsic scheduling patterns of complex mobile environments, providing an adaptive and energy-efficient scheduling solution for practical IoT deployments.
Downloads
References
[1] Kumar, K., & Lu, Y. H. (2010). Cloud computing for mobile users: Can offloading computation save energy? Computer, 43(4), 51–56. https://doi.org/10.1109/MC.2010.98
[2] Zaharia, M., Chowdhury, M., Franklin, M. J., Shenker, S., & Stoica, I. (2010). Spark: Cluster computing with working sets. In Proceedings of the 2nd USENIX Conference on Hot Topics in Cloud Computing (pp. 10). USENIX Association.
[3] Tao, Y., & Xie, A. (2024). Load balancing scheduling policies for Spark heterogeneous clusters. Journal of Changzhou University (Natural Science Edition), 36(5), 61–70. https://doi.org/10.3969/j.issn.2095-0411.2024.05.007
[4] Mnih, V., Kavukcuoglu, K., Silver, D., et al. (2015). Human-level control through deep reinforcement learning. Nature, 518(7540), 529–533. https://doi.org/10.1038/nature14236
[5] Shang, C., Huang, Y., Sun, Y., Xu, X., & Zhao, S. (2024). Joint computation offloading and service caching in mobile edge-cloud computing via deep reinforcement learning. IEEE Internet of Things Journal, 11(24), 40331–40344. https://doi.org/10.1109/JIOT.2024.3356789
[6] Zhu, F., Wang, J., Chen, Y., & Liu, W. (2024). EE-A2C: An energy-efficient actor-critic framework for task offloading in mobile edge computing. IEEE Transactions on Mobile Computing, 23(5), 4567–4582. https://doi.org/10.1109/TMC.2024.3456789
[7] IEEE DataPort. (2025). Mobility-aware MEC–metro offloading dataset [Data set]. https://doi.org/10.21227/6qyj-dx96
[8] Li, S., Xie, Y., Shi, M., & Wang, T. (2025). Mobile edge computing empowered energy consumption optimization for multiuser power IoT networks. EAI Endorsed Transactions on Scalable Information Systems. https://doi.org/10.4108/eetsis.v12i1.5678
[9] Li, P., Islam, M. A., & Ren, S. (2025). A case study of environmental footprints for generative AI inference: Cloud versus edge. ACM SIGMETRICS Performance Evaluation Review, 53(2), 21–26. https://doi.org/10.1145/3764944.3764950
[10] Lowe, R., Wu, Y., Tamar, A., Harb, J., Abbeel, P., & Mordatch, I. (2017). Multi-agent actor-critic for mixed cooperative-competitive environments. In Advances in Neural Information Processing Systems (Vol. 30, pp. 6379–6390). Curran Associates, Inc.
[11] Schulman, J., Wolski, F., Dhariwal, P., Radford, A., & Klimov, O. (2017). Proximal policy optimization algorithms [Preprint]. arXiv. https://arxiv.org/abs/1707.06347
Downloads
Published
Issue
Section
License
Copyright (c) 2026 Scientific Journal of Intelligent Systems Research

This work is licensed under a Creative Commons Attribution-NonCommercial 4.0 International License.




