Hybrid Multi-Agent Q-Learning and Intelligent Water Drops Framework for Adaptive Campus Bandwidth Allocation

Authors

  • O. S. Adeoye Department of Computer Science, Federal University of Technology Minna
  • M. B. Abdullahi Department of Data Science, Federal University of Technology, Minna
  • O. S. Adebayo Department of Cyber Security, National Open University of Nigeria, Abuja
  • O. A. Ojerinde Department of Computer Science, Federal University of Technology Minna
  • M. Danlami Department of Computer Engineering, Federal University of Technology Minna

DOI:

https://doi.org/10.63746/njtd.v23i2.4492

Keywords:

Multi-Agent Quality-Learning (MAQL),, Intelligent Water Drops Algorithm (IWDA),, Network traffic analysis,, Bandwidth allocation, Reinforcement learning

Abstract

Efficient bandwidth allocation in campus networks remains challenging because of dynamic traffic patterns, heterogeneous user demands, congestion, and fairness constraints. This study aims to develop and evaluate a hybrid Intelligent Water Drops Multi-Agent Q-Learning (IWDAMA) framework for adaptive bandwidth allocation in software-defined campus networks. The proposed framework integrates Multi-Agent Q-Learning with the Intelligent Water Drops algorithm to enhance exploration, convergence, and adaptive scheduling. Performance evaluation was conducted using MATLAB simulations in a software-defined campus network under five traffic scenarios involving 200–2000 users with a fixed link capacity of 2000 Mbps. The proposed model was compared with standalone Multi-Agent Q-Learning (MAQL), Q-learning, and First-Come-First-Served (FCFS) scheduling using throughput, delay, bandwidth utilization, fairness index, and packet loss as evaluation metrics. Experimental results showed that IWDAMA achieved throughput ranging from 392 Mbps to 1764 Mbps, bandwidth utilization of 0.99, fairness index of 0.995, packet loss between 0.0159 and 0.0600, and delay between 0.0636 ms and 0.2019 ms, consistently outperforming MAQL, Q-learning, and FCFS across all scenarios. The findings demonstrate that integrating heuristic optimization with cooperative reinforcement learning provides a scalable and robust solution for intelligent bandwidth management in campus and enterprise software-defined networks.

References

Agor, A.D., Asante, M., Hayfron-Acquah, J.B., Peasah, K.O., Agangibi, M., Elliot, M.A.A., Akanferi, A.A. and Asampana, I., 2024. Power-aware intelligent water drops routing algorithm for best path selection in MANETs. International Journal of Communication Networks and Information Security, 16(2), pp.1–13.

Ali, R., Zikria, Y.B., Bashir, A.K., Garg, S. and Kim, H.S., 2021. URLLC for 5G and beyond: Requirements, enabling incumbent technologies and network intelligence. IEEE Access, 9, pp.67064–67095.

Boussaoud, K., En-Nouaary, A. and Ayache, M., 2025. Adaptive congestion detection and traffic control in software-defined networks via data-driven multi-agent reinforcement learning. Computers, 14(6), 236. https://doi.org/10.3390/computers14060236

Cui, T., Lin, X., Li, S., Chen, M., Yin, Q., Li, Q. and Xu, K., 2025. TrafficLLM: Enhancing large language models for network traffic analysis with generic traffic representation. arXiv preprint, arXiv:2504.04222. Available at: https://arxiv.org/abs/2504.04222

Ge, J., Liu, B., Wang, T., Yang, Q., Liu, A. and Li, A., 2020. Q-learning-based flexible task scheduling in a global view for the Internet of Things. Transactions on Emerging Telecommunications Technologies, e4111. https://doi.org/10.1002/ett.4111

Goudarzi, P., Hosseinpour, M., Goudarzi, R. and Lloret, J., 2022. Holistic utility satisfaction in cloud data centre networks using reinforcement learning. Future Internet, 14(12), 368. https://doi.org/10.3390/fi14120368

Guo, Y., Tang, Q., Ma, Y., Tian, H. and Chen, K., 2024. Distributed traffic engineering in hybrid software-defined networks: A multi-agent reinforcement learning framework. IEEE Transactions on Network and Service Management, 21(6), pp.6759–6769. https://doi.org/10.1109/TNSM.2024.3454282

Halder, S., Sharma, H.K., Biswas, A., Prentkovskis, O., Majumder, S. and Ska?kauskas, P., 2023. On enhanced intelligent water drops algorithm for the travelling salesman problem under uncertain paradigm. Transport and Telecommunication, 24(3), pp.228–255. https://doi.org/10.2478/ttj-2023-0019

International Telecommunication Union, 2023. Measuring digital development: Facts and figures 2023. Geneva: ITU. Available at: https://www.itu.int/en/ITU-D/Statistics/Pages/facts/default.aspx

Krishnamoorthy, S., Dua, A. and Gupta, S., 2023. Role of emerging technologies in future IoT-driven Healthcare 4.0: A survey of challenges and directions. Journal of Ambient Intelligence and Humanized Computing, 14(1), pp.361–407.

Koumar, J., Hynek, K., Pešek, J., and ?ejka, T., 2023. NetTiSA: Extended IP flow with time-series features for universal bandwidth-constrained high-speed network traffic classification. arXiv preprint arXiv:2310.05530, 2023. Available at: https://arxiv.org/abs/2310.05530

Mohammed, A.S.B., Muhammad, D. and Gelwasa, Y.G., 2024. Developing a dynamic bandwidth allocation prototype model for campus networks based on network traffic analysis: A case study of KSUSTA. Adeleke University Journal of Engineering and Technology, 7(1), pp.173–183.

Movahed, A.B., Khakbazan, M., Abyaneh, A.G. and Modarres, S., 2025. Adaptive optimization of industrial workforce allocation via the intelligent water drops algorithm. Journal of Future Digital Optimization. Available at: https://iscihub.com/index.php/JFDO

Mushtaq, A., Haq, I.U., Sarwar, M.A., Khan, A., Khalil, W. and Mughal, M.A., 2023. Multi-agent reinforcement learning for traffic flow management of autonomous vehicles. Sensors, 23(5), 2373. https://doi.org/10.3390/s23052373

Sohaib, M., Jeong, J. and Jeon, S.-W., 2021. Dynamic multichannel access via multi-agent reinforcement learning: Throughput and fairness guarantees. arXiv preprint, arXiv:2105.04077. Available at: https://arxiv.org/abs/2105.04077

Tariq, N., Asim, M., Khan, F.A., Baker, T., Khalid, U. and Derhab, A., 2021. A blockchain-based multi-mobile code-driven trust mechanism for detecting internal attacks in the Internet of Things. Sensors, 21(1), 23. https://doi.org/10.3390/s21010023

Wang, H., Liu, Y., Li, W. and Yang, Z., 2024. Multi-agent deep reinforcement learning-based fine-grained traffic scheduling in data center networks. Future Internet, 16(4), 119. https://doi.org/10.3390/fi16040119

Wang, J.-H., He, H., Cha, J., Jeong, I. and Ahn, C.-J., 2025. Multi-agent reinforcement learning for efficient resource allocation in the Internet of Vehicles. Electronics, 14(1), 192. https://doi.org/10.3390/electronics14010192

Wang, P., Zheng, Z., Di, B. and Song, L., 2019. HetMEC: Latency-optimal task assignment and resource allocation for heterogeneous multi-layer mobile edge computing. IEEE Transactions on Wireless Communications, 18(10), pp.4942–4956. https://doi.org/10.1109/TWC.2019.2931315

Yang, X.-S., 2020. Nature-inspired optimization algorithms: Challenges and open problems. Journal of Computational Science, 46, 101104. https://doi.org/10.1016/j.jocs.2020.101104

Yang, Z., Jin, Y., Liu, J., Xu, X., Zhang, Y. and Ji, S., 2025. Research on cloud platform network traffic monitoring and anomaly detection system based on large language models. arXiv preprint, arXiv:2504.17807. Available at: https://arxiv.org/abs/2504.17807

Zhou, Z., Chen, X., Li, E., Zeng, L., Luo, K. and Zhang, J., 2020. Edge intelligence: Paving the last mile of artificial intelligence with edge computing. Proceedings of the IEEE, 107(8), pp.1738–1762. https://doi.org/10.1109/JPROC.2019.2918951

Published

2026-06-30