Skip to main navigation Skip to search Skip to main content

Online Data-Driven-Based Optimal Output Tracking Control Without Initial Stabilizing Policy

  • Beijing Institute of Technology

Research output: Contribution to journalArticlepeer-review

Abstract

This article investigates the optimal output tracking control problem for a continuous-time linear system with an unknown system model. By integrating adaptive dynamic programming with optimal control theory, a dual policy iteration (PI) learning algorithm composed of two PI schemes is proposed to adaptively learn the optimal tracking controller. The primary advantage of the proposed algorithm lies in that it does not require an initial stabilizing control policy, persistence of excitation, or storage of historical data to guarantee convergence. This feature fundamentally distinguishes it from existing approaches based on the least-squares method, which rely on these conditions. Simulation results demonstrate the effectiveness of the proposed algorithm, and its superiority is further validated through comparisons with existing methods.

Original languageEnglish
JournalIEEE Transactions on Cybernetics
DOIs
Publication statusAccepted/In press - 2026
Externally publishedYes

Keywords

  • Adaptive dynamic programming
  • optimal control
  • output tracking control
  • policy iteration

Fingerprint

Dive into the research topics of 'Online Data-Driven-Based Optimal Output Tracking Control Without Initial Stabilizing Policy'. Together they form a unique fingerprint.

Cite this