Abstract:By leveraging the concept of mobile edge computing (MEC), massive amount of data generated by a large number of Internet of Things (IoT) devices could be offloaded to MEC server at the edge of wireless network for further computational intensive processing. However, due to the resource constraint of IoT devices and wireless network, both the communications and computation resources need to be allocated and scheduled efficiently for better system performance. In this paper, we propose a joint computation offloading and multi-user scheduling algorithm for IoT edge computing system to minimize the long-term average weighted sum of delay and power consumption under stochastic traffic arrival. We formulate the dynamic optimization problem as an infinite-horizon average-reward continuous-time Markov decision process (CTMDP) model. One critical challenge in solving this MDP problem for the multi-user resource control is the curse-of-dimensionality problem, where the state space of the MDP model and the computation complexity increase exponentially with the growing number of users or IoT devices. In order to overcome this challenge, we use the deep reinforcement learning (RL) techniques and propose a neural network architecture to approximate the value functions for the post-decision system states. The designed algorithm to solve the CTMDP problem supports semi-distributed auction-based implementation, where the IoT devices submit bids to the BS to make the resource control decisions centrally. Simulation results show that the proposed algorithm provides significant performance improvement over the baseline algorithms, and also outperforms the RL algorithms based on other neural network architectures.

Delay-Aware Stochastic Resource Management for Mobile Edge Computing Systems Via Constrained Reinforcement Learning

Delay-Aware Power Control for Downlink Multi-User MIMO Via Constrained Deep Reinforcement Learning.

Successive Convex Approximation Based Off-Policy Optimization for Constrained Reinforcement Learning

Stochastic Resource Allocation and Delay Analysis for Mobile Edge Computing Systems

Multi-user Resource Control with Deep Reinforcement Learning in IoT Edge Computing

Energy-Efficient Resource Allocation for Latency-Sensitive Mobile Edge Computing

Delay-Optimal Computation Offloading for Computation-Constrained Mobile Edge Networks.

Energy Efficient Joint Computation Offloading and Service Caching for Mobile Edge Computing: A Deep Reinforcement Learning Approach

Optimal Computation Resource Allocation in Energy-Efficient Edge IoT Systems with Deep Reinforcement Learning

Decentralized Computation Offloading for Multi-User Mobile Edge Computing: A Deep Reinforcement Learning Approach

Robust Offloading for Edge Computing-Assisted Sensing and Communication Systems: A Deep Reinforcement Learning Approach

Deep Reinforcement Learning-based Power Control and Bandwidth Allocation Policy for Weighted Cost Minimization in Wireless Networks

A collaborative optimization strategy for computing offloading and resource allocation based on multi-agent deep reinforcement learning

Decentralized Scheduling for Concurrent Tasks in Mobile Edge Computing Via Deep Reinforcement Learning

A Dynamic Service Placement Based on Deep Reinforcement Learning in Mobile Edge Computing

Priority-Aware Resource Allocation for RIS-assisted Mobile Edge Computing Networks: A Deep Reinforcement Learning Approach

Dynamic Offloading Strategy for Delay-Sensitive Task in Mobile-Edge Computing Networks

Multi-objective Deep Reinforcement Learning for Mobile Edge Computing

Scheduling for Mobile Edge Computing with Random User Arrivals: An Approximate MDP and Reinforcement Learning Approach

Q-greedyUCB: a New Exploration Policy to Learn Resource-Efficient Scheduling

Deep Reinforcement Learning Based Task Offloading and Resource Allocation in Small Cell MEC