跳到主要导航 跳到搜索 跳到主要内容

Q-Advantage Integrated Human-Guided Reinforcement Learning for Safe End-to-End Autonomous Driving

  • Yong Wang
  • , Pei Wang
  • , Hongwen He*
  • , Jingda Wu
  • , Yingjuan Tang
  • , Zirui Kuang
  • *此作品的通讯作者
  • The University of Hong Kong
  • Seres Automobile
  • Beijing Institute of Technology

科研成果: 期刊稿件文章同行评审

摘要

Reinforcement learning (RL) is a promising approach for end-to-end autonomous driving, but its practical deployment remains challenging due to low sample efficiency and sensitivity to reward design. To address these challenges, this study presents a novel Q-advantage integrated human-guided reinforcement learning (QIHG-RL) framework that effectively combines the strengths of machine learning and human expertise. The QIHG-RL framework features: 1) an ensemble Q-advantage function that aggregates multiple value networks to enhance value estimation, and 2) an integration mechanism that embeds the Q-advantage into both the actor-critic network and the prioritized experience replay. This design allows the agent to leverage sparse and sub-optimal human demonstrations, accelerating policy learning in the early training phase while gradually enhancing exploration as training progresses. The framework is evaluated across three safety-critical driving tasks. Experimental results show a 167% improvement in sample efficiency compared to standard RL methods and a 14% performance gain over a state-of-the-art human-guided RL baseline. Furthermore, a Sim2Real pipeline combining domain randomization and semantic denoised remapping facilitates successful deployment on a real-world autonomous vehicle.

源语言英语
页(从-至)2957-2969
页数13
期刊IEEE Transactions on Intelligent Transportation Systems
27
3
DOI
出版状态已出版 - 2026
已对外发布

学术指纹

探究 'Q-Advantage Integrated Human-Guided Reinforcement Learning for Safe End-to-End Autonomous Driving' 的科研主题。它们共同构成独一无二的学术指纹。

引用此