跳到主要导航 跳到搜索 跳到主要内容

Disturbance-Rejection Reinforcement Learning for Control of Unknown Nonlinear Systems

  • Beijing Institute of Technology
  • College of Information Engineering
  • Zhejiang University of Technology
  • Zhongyuan University of Technology
  • Beijing Institute of Technology

科研成果: 期刊稿件文章同行评审

摘要

How to realize the ultimate boundedness during training on unknown nonlinear systems under continuous stochastic disturbances is a key challenge in reinforcement learning (RL). To this end, we propose a Lyapunov-based data-driven RL framework for disturbance rejection control and provides a step toward reliable disturbance-rejection learning. First, a stabilizable reference model is constructed via input-convex neural networks (ICNNs) to generate stabilizable, disturbance-aware trajectories from data that reduce sensitivity to nominal model mismatch. Second, we establish a data-driven ultimate-boundedness theorem. It certifies closed-loop stability directly from sampled transitions without requiring prior model knowledge. Building on these foundations, we propose an off-policy algorithm that alternates between reference-model learning and policy optimization. It enforces Lyapunov decrease in expectation and jointly improves performance and robustness to achieve ultimate boundedness. Experiments on a classic control task and MuJoCo benchmarks demonstrate that the proposed method reduces the final-20-episode average cost by up to 75% and achieves average episode lengths over the final 20 episodes up to 3.2 times longer than those of representative Lyapunov-based RL methods. the framework guarantees bounded state trajectories, ensuring safety guarantees while maintaining performance under unknown conditions.

源语言英语
页(从-至)15238-15255
页数18
期刊IEEE Transactions on Automation Science and Engineering
23
DOI
出版状态已出版 - 2026
已对外发布

学术指纹

探究 'Disturbance-Rejection Reinforcement Learning for Control of Unknown Nonlinear Systems' 的科研主题。它们共同构成独一无二的学术指纹。

引用此