Model-Based Reinforcement Learning Framework of Online Network Resource Allocation

10/18/2021
by   Bahador Bakhshi, et al.
0

Online Network Resource Allocation (ONRA) for service provisioning is a fundamental problem in communication networks. As a sequential decision-making under uncertainty problem, it is promising to approach ONRA via Reinforcement Learning (RL). But, RL solutions suffer from the sample complexity issue; i.e., a large number of interactions with the environment needed to find an efficient policy. This is a barrier to utilize RL for ONRA as on one hand, it is not practical to train the RL agent offline due to lack of information about future requests, and on the other hand, online training in the real network leads to significant performance loss because of the sub-optimal policy during the prolonged learning time. This performance degradation is even higher in non-stationary ONRA where the agent should continually adapt the policy with the changes in service requests. To deal with this issue, we develop a general resource allocation framework, named RADAR, using model-based RL for a class of ONRA problems with the known immediate reward of each action. RADAR improves sample efficiency via exploring the state space in the background and exploiting the policy in the decision-time using synthetic samples by the model of the environment, which is trained by real interactions. Applying RADAR on the multi-domain service federation problem, to maximize profit via selecting proper domains for service requests deployment, shows its continual learning capability and up to 44 RL solution.

READ FULL TEXT
research
10/20/2020

Quality of service based radar resource management using deep reinforcement learning

An intelligent radar resource management is an essential milestone in th...
research
11/30/2022

General policy mapping: online continual reinforcement learning inspired on the insect brain

We have developed a model for online continual or lifelong reinforcement...
research
07/13/2022

Hindsight Learning for MDPs with Exogenous Inputs

We develop a reinforcement learning (RL) framework for applications that...
research
03/21/2022

Lean Evolutionary Reinforcement Learning by Multitasking with Importance Sampling

Studies have shown evolution strategies (ES) to be a promising approach ...
research
01/24/2022

Cache Allocation in Multi-Tenant Edge Computing via online Reinforcement Learning

We consider in this work Edge Computing (EC) in a multi-tenant environme...
research
10/04/2021

Reinforcement Learning for Admission Control in Wireless Virtual Network Embedding

Using Service Function Chaining (SFC) in wireless networks became popula...
research
05/22/2023

Scaling Serverless Functions in Edge Networks: A Reinforcement Learning Approach

With rapid advances in containerization techniques, the serverless compu...

Please sign up or login with your details

Forgot password? Click here to reset