Model-Based Reinforcement Learning Framework of Online Network Resource   Allocation

Bahador Bakhshi; Josep Mangues-Bafalluy

arXiv:2110.09236·cs.NI·October 19, 2021

Model-Based Reinforcement Learning Framework of Online Network Resource Allocation

Bahador Bakhshi, Josep Mangues-Bafalluy

PDF

Open Access

TL;DR

This paper introduces RADAR, a model-based reinforcement learning framework that enhances online network resource allocation by improving sample efficiency and continual learning, demonstrating significant performance gains over traditional methods.

Contribution

The paper presents a novel model-based RL framework, RADAR, specifically designed for online network resource allocation, addressing sample complexity and adaptability in non-stationary environments.

Findings

01

Achieves up to 44% performance improvement over standard model-free RL.

02

Demonstrates continual learning capability in non-stationary ONRA scenarios.

03

Effectively utilizes synthetic samples for policy optimization.

Abstract

Online Network Resource Allocation (ONRA) for service provisioning is a fundamental problem in communication networks. As a sequential decision-making under uncertainty problem, it is promising to approach ONRA via Reinforcement Learning (RL). But, RL solutions suffer from the sample complexity issue; i.e., a large number of interactions with the environment needed to find an efficient policy. This is a barrier to utilize RL for ONRA as on one hand, it is not practical to train the RL agent offline due to lack of information about future requests, and on the other hand, online training in the real network leads to significant performance loss because of the sub-optimal policy during the prolonged learning time. This performance degradation is even higher in non-stationary ONRA where the agent should continually adapt the policy with the changes in service requests. To deal with this…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsEnergy Harvesting in Wireless Networks · Age of Information Optimization · Software-Defined Networks and 5G