面向复杂非线性系统的数据高效强化学习

Data-Efficient Reinforcement Learning for Complex Nonlinear Systems

IEEE Transactions on Cybernetics · 2023
被引 25
ABS 3

中文导读

提出一种基于Koopman算子的无模型强化学习算法,通过将非线性系统提升为线性模型来减少数据需求,并在电力系统励磁控制中验证了有效性。

Abstract

This article proposes a data-efficient model-free reinforcement learning (RL) algorithm using Koopman operators for complex nonlinear systems. A high-dimensional data-driven optimal control of the nonlinear system is developed by lifting it into the linear system model. We use a data-driven model-based RL framework to derive an off-policy Bellman equation. Building upon this equation, we deduce the data-efficient RL algorithm, which does not need a Koopman-built linear system model. This algorithm preserves dynamic information while reducing the required data for optimal control learning. Numerical and theoretical analyses of the Koopman eigenfunctions for dataset truncation are discussed in the proposed model-free data-efficient RL algorithm. We validate our framework on the excitation control of the power system.

强化学习非线性系统最优控制Koopman算子