Towards Optimal Pricing of Demand Response – A Nonparametric Constrained Policy Optimization Approach

06/24/2023
by   Jun Song, et al.
0

Demand response (DR) has been demonstrated to be an effective method for reducing peak load and mitigating uncertainties on both the supply and demand sides of the electricity market. One critical question for DR research is how to appropriately adjust electricity prices in order to shift electrical load from peak to off-peak hours. In recent years, reinforcement learning (RL) has been used to address the price-based DR problem because it is a model-free technique that does not necessitate the identification of models for end-use customers. However, the majority of RL methods cannot guarantee the stability and optimality of the learned pricing policy, which is undesirable in safety-critical power systems and may result in high customer bills. In this paper, we propose an innovative nonparametric constrained policy optimization approach that improves optimality while ensuring stability of the policy update, by removing the restrictive assumption on policy representation that the majority of the RL literature adopts: the policy must be parameterized or fall into a certain distribution class. We derive a closed-form expression of optimal policy update for each iteration and develop an efficient on-policy actor-critic algorithm to address the proposed constrained policy optimization problem. The experiments on two DR cases show the superior performance of our proposed nonparametric constrained policy optimization method compared with state-of-the-art RL algorithms.

READ FULL TEXT
research
08/29/2020

Market Model for Demand Response under Block Rate Pricing

Renewable sources are taking center stage in electricity generation. How...
research
09/20/2022

A Deep Reinforcement Learning-Based Charging Scheduling Approach with Augmented Lagrangian for Electric Vehicle

This paper addresses the problem of optimizing charging/discharging sche...
research
09/23/2020

Demand Responsive Dynamic Pricing Framework for Prosumer Dominated Microgrids using Multiagent Reinforcement Learning

Demand Response (DR) has a widely recognized potential for improving gri...
research
03/03/2022

Optimized cost function for demand response coordination of multiple EV charging stations using reinforcement learning

Electric vehicle (EV) charging stations represent a substantial load wit...
research
02/23/2021

Doubly Robust Off-Policy Actor-Critic: Convergence and Optimality

Designing off-policy reinforcement learning algorithms is typically a ve...
research
09/21/2021

Home Energy Management Systems: Operation and Resilience of Heuristics against Cyberattacks

Internet of Things (IoT) and advanced communication technologies have de...
research
11/27/2022

Combined Peak Reduction and Self-Consumption Using Proximal Policy Optimization

Residential demand response programs aim to activate demand flexibility ...

Please sign up or login with your details

Forgot password? Click here to reset