Skip to main navigation Skip to search Skip to main content

Safe Learning Control with Optimality and Stability Guarantees

  • Xinyang Wang
  • , Hongwei Zhang*
  • , Shimin Wang
  • , Wei Xiao
  • , Martin Guay
  • *Corresponding author for this work
  • Harbin Institute of Technology
  • Lingnan University
  • Nanyang Technological University
  • Massachusetts Institute of Technology
  • Queen's University Kingston

Research output: Contribution to journalArticlepeer-review

Abstract

Merely pursuing performance may adversely affect safety, while a conservative policy for safe exploration will degrade the performance. How to guarantee both safety and performance in learning-based control problems is an interesting yet challenging issue. This paper aims to enhance system performance with a safety guarantee by solving reinforcement learning (RL)-based optimal control problems for nonlinear systems subject to high-relativedegree state constraints and unknown time-varying disturbance/ actuator faults. A new type of control barrier functions (CBFs), termed high-order reciprocal-based control barrier function, is proposed to handle high-relative-degree constraints, which extends the design of CBFs to enforce robust safety without knowing the disturbance bound. The concept of gradient similarity is proposed to quantify the relationship between safety and performance. Finally, gradient manipulation and adaptive mechanisms are introduced in the model-based safe RL framework to enhance the performance with a safety guarantee. Two simulation examples illustrate the efficacy of the proposed algorithms.

Original languageEnglish
JournalIEEE Transactions on Automatic Control
DOIs
StateAccepted/In press - 2026
Externally publishedYes

Keywords

  • control barrier function
  • reinforcement learning
  • Safety guarantee

Fingerprint

Dive into the research topics of 'Safe Learning Control with Optimality and Stability Guarantees'. Together they form a unique fingerprint.

Cite this