Please use this identifier to cite or link to this item:
http://hdl.handle.net/20.500.11960/4806| Title: | Optimizing 5G network slicing with DRL: Balancing eMBB, URLLC, and mMTC with OMA, NOMA, and RSMA |
| Authors: | Malta, Silvestre Pinto, Pedro Fernandez-Veiga, Manuel |
| Keywords: | 5G eMBB URLLC mMTC Network slicing OMA NOMA RSMA Deep reinforcement learning Q-learning DQN |
| Issue Date: | 2025 |
| Citation: | Malta, S., Pinto, P., & Fernández-Veiga, M. (2025). Optimizing 5G network slicing with DRL: Balancing eMBB, URLLC, and mMTC with OMA, NOMA, and RSMA. Journal of Network and Computer Applications, 234, Artigo e104068. https://doi.org/10.1016/j.jnca.2024.104068 |
| Abstract: | The advent of 5th Generation (5G) networks has introduced the strategy of network slicing as a paradigm shift, enabling the provision of services with distinct Quality of Service (QoS) requirements. The 5th Generation New Radio (5G NR) standard complies with the use cases Enhanced Mobile Broadband (eMBB), Ultra-Reliable Low Latency Communications (URLLC), and Massive Machine Type Communications (mMTC), which demand a dynamic adaptation of network slicing to meet the diverse traffic needs. This dynamic adaptation presents both a critical challenge and a significant opportunity to improve 5G network efficiency. This paper proposes a Deep Reinforcement Learning (DRL) agent that performs dynamic resource allocation in 5G wireless network slicing according to traffic requirements of the 5G use cases within two scenarios: eMBB with URLLC and eMBB with mMTC. The DRL agent evaluates the performance of different decoding schemes such as Orthogonal Multiple Access (OMA), Non-Orthogonal Multiple Access (NOMA), and Rate Splitting Multiple Access (RSMA) and applies the best decoding scheme in these scenarios under different network conditions. The DRL agent has been tested to maximize the sum rate in scenario eMBB with URLLC and to maximize the number of successfully decoded devices in scenario eMBB with mMTC, both with different combinations of number of devices, power gains and number of allocated frequencies. The results show that the DRL agent dynamically chooses the best decoding scheme and presents an efficiency in maximizing the sum rate and the decoded devices between 84% and 100% for both scenarios evaluated. |
| URI: | http://hdl.handle.net/20.500.11960/4806 |
| ISSN: | 1084-8045 |
| Appears in Collections: | ESTG - Publicações indexadas à WoS/Scopus |
Files in This Item:
| File | Description | Size | Format | |
|---|---|---|---|---|
| 1-s2.0-S1084804524002455-main.pdf | 1.63 MB | Adobe PDF | View/Open |
Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.

