跳到主要导航 跳到搜索 跳到主要内容

Data-Empowered Trajectory Planning Based on Two-Phase Deep Reinforcement Learning Method

  • Peng Yin
  • , Yiwei Liu
  • , Linye Wang
  • , Yizheng Ge
  • , Weihao Yan
  • , Jihua Lu*
  • , Lihui Feng
  • , Yufan Du
  • *此作品的通讯作者
  • University of Chinese Academy of Sciences
  • Defence Industry Secrecy Examination and Certification Center
  • Beijing Institute of Technology

科研成果: 期刊稿件文章同行评审

摘要

Existing trajectory planning methods for distributed unmanned systems are limited and constrained by data-empowered information from environments. A deep reinforcement learning-based two-phase trajectory planning method is proposed, including the mission transition phase (MTP) and mission maintenance phase (MMP). During MTP, the mobile node transfers from the current to the target position while avoiding obstacles. Meanwhile, assisted communication among nodes in the mission area exists for MMP. The deep learning model is designed for these two phases, respectively, to realize trajectory planning. The optimal model that improves the planning reward is obtained by using experience pools and sampling. Further, it is capable of dealing with complex and high-dimensional optimization as well as adapting to the dynamic environment, making trajectory planning more accurate and efficient. By consuming similar time steps to the optimal path method and 1/3 time steps of coordinate transition methods, the safety of unmanned system is guaranteed and energy consumption is reduced. Moreover, obvious advantages of the method are illustrated in the deployment of large-scale network scenes and auxiliary communication task is fulfilled with simplified processes, resulting in 2/3 and 1/4 computation complexities of particle swarm optimization and scanning methods.

源语言英语
页(从-至)44051-44059
页数9
期刊IEEE Internet of Things Journal
12
21
DOI
出版状态已出版 - 2025
已对外发布

联合国可持续发展目标

此成果有助于实现下列可持续发展目标:

  1. 可持续发展目标 7 - 经济适用的清洁能源
    可持续发展目标 7 经济适用的清洁能源

学术指纹

探究 'Data-Empowered Trajectory Planning Based on Two-Phase Deep Reinforcement Learning Method' 的科研主题。它们共同构成独一无二的学术指纹。

引用此