Skip to main navigation Skip to search Skip to main content

A Low-Carbon Economic Dispatch Method for Power Systems with Carbon Capture Plants Based on Safe Reinforcement Learning

  • School of Electrical Engineering and Automation, Harbin Institute of Technology

Research output: Contribution to journalArticlepeer-review

Abstract

To address the high-dimensional and complex scheduling issues in the low-carbon economic dispatch (LCED) with carbon capture plants, in this article, we propose a novel safe reinforcement learning (SRL) based on heterogeneous action space representation, which can make fast decisions for both optimal power flow and carbon capture operation. First, SRL is designed based on the feasible set to ensure that the dispatch results continuously remain within the preset range. Then, to tackle the problem of having a large number of discrete and continuous variables in the LCED, this article employs a parameterized Markov process to represent these discrete-continuous actions and uses a conditional variational autoencoder to depict heterogeneous space. To learn the correlation between discrete and continuous action spaces, a mechanism for approximating action space based on small-sample behavior cloning is proposed, and a method based on dynamic time warping for calculating environment similarity is designed for determining the value of the regularization term. Finally, numerical simulations validate the superiority and scalability of the proposed method in enhancing decision-making efficiency and promoting the low-carbon economic operation of the power system.

Original languageEnglish
Pages (from-to)10542-10553
Number of pages12
JournalIEEE Transactions on Industrial Informatics
Volume20
Issue number8
DOIs
StatePublished - 2024
Externally publishedYes

Keywords

  • Carbon capture plant (CCP)
  • discrete continuous actions
  • low-carbon economic dispatch (LCED)
  • safe reinforcement learning (SRL)

Fingerprint

Dive into the research topics of 'A Low-Carbon Economic Dispatch Method for Power Systems with Carbon Capture Plants Based on Safe Reinforcement Learning'. Together they form a unique fingerprint.

Cite this