Abstract:The growing penetration of renewable energy has brought significant challenges for modern power system operation. Academic research and industrial practice show that adjusting unit commitment (UC) scheduling periodically according to new forecasts of renewable power provides a promising way to improve system stability and economy; however, this greatly increases the computational burden for solution methods. In this paper, a deep reinforcement learning (DRL) method is proposed to obtain timely and reliable solutions for rolling-horizon UC (RHUC). First, based on historical data and day-ahead point forecasting, a data-driven method is designed to construct typical wind power scenarios that are regarded as components of the state space of DRL. Second, a rolling mechanism is proposed to dynamically update the state space based on real-time wind power data. Third, unlike existing reinforcement learning-based UC solution methods that segment the continuous outputs of generators as discrete variables, all the variables in RHUC are regarded as continuous. Additionally, a series of updating regulations are defined to ensure that the model is realistic. Thus, a DRL algorithm, the twin delayed deep deterministic policy gradient (TD3), can be utilized to effectively solve the problem. Finally, several case studies are conducted based on different test systems to demonstrate the efficiency of the proposed method. According to the experimental results, the proposed algorithm can obtain high-quality solutions in a considerably shorter time than traditional methods, which leads to a reduction of at least 1.1% in the power system operation cost.

Research on Unit Commitment Optimization Method Based on Deep Reinforcement Learning

Unit Commitment Based on Improved Discrete Particle Swarm Optimization

Target-Value-Competition-Based Multi-Agent Deep Reinforcement Learning Algorithm for Distributed Nonconvex Economic Dispatch

Multi-agent Deep Reinforcement Learning Algorithm for Distributed Economic Dispatch in Smart Grid.

Combination optimization method of grid sections based on deep reinforcement learning with accelerated convergence speed

Deep Learning-Based Rolling Horizon Unit Commitment under Hybrid Uncertainties

Large-scale deep reinforcement learning method for energy management of power supply units considering regulation mileage payment

Reinforcement Learning and Stochastic Optimization with Deep Learning-Based Forecasting on Power Grid Scheduling

Joint Optimization Dispatching for Hybrid Power System Based on Deep Reinforcement Learning

Optimizing Load Scheduling in Power Grids Using Reinforcement Learning and Markov Decision Processes

Improving the Computational Efficiency of the Unit Commitment Problem in Hydrothermal Systems by Using Multi-Agent Deep Reinforcement Learning

Reactive power optimization via deep transfer reinforcement learning for efficient adaptation to multiple scenarios

Distributed $Q$ -Learning-based Online Optimization Algorithm for Unit Commitment and Dispatch in Smart Grid

An Improved Deep Reinforcement Learning Method for Dispatch Optimization Strategy of Modern Power Systems

An ultra-fast optimization algorithm for unit commitment based on neural branching

Rolling horizon wind-thermal unit commitment optimization based on deep reinforcement learning

A Reinforcement Learning Algorithm Based on Neural Network for Economic Dispatch

A deep reinforcement learning algorithm for the order optimization allocation of total power in the interconnected power grids

Deep Reinforcement Learning Based Wind Farm Cluster Reactive Power Optimization Considering Network Switching

Hybrid heuristic-progressive optimality approach to unit commitment

A Data-driven Method for Fast AC Optimal Power Flow Solutions via Deep Reinforcement Learning