摘要
This article investigates the optimal output tracking control problem for a continuous-time linear system with an unknown system model. By integrating adaptive dynamic programming with optimal control theory, a dual policy iteration (PI) learning algorithm composed of two PI schemes is proposed to adaptively learn the optimal tracking controller. The primary advantage of the proposed algorithm lies in that it does not require an initial stabilizing control policy, persistence of excitation, or storage of historical data to guarantee convergence. This feature fundamentally distinguishes it from existing approaches based on the least-squares method, which rely on these conditions. Simulation results demonstrate the effectiveness of the proposed algorithm, and its superiority is further validated through comparisons with existing methods.
| 源语言 | 英语 |
|---|---|
| 期刊 | IEEE Transactions on Cybernetics |
| DOI | |
| 出版状态 | 已接受/待刊 - 2026 |
| 已对外发布 | 是 |
学术指纹
探究 'Online Data-Driven-Based Optimal Output Tracking Control Without Initial Stabilizing Policy' 的科研主题。它们共同构成独一无二的学术指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver