Skip to main navigation Skip to search Skip to main content

Informative Trajectory Planning Using Reinforcement Learning for Minimum-Time Exploration of Spatiotemporal Fields

  • Tsinghua University

Research output: Contribution to journalArticlepeer-review

Abstract

This article studies the informative trajectory planning problem of an autonomous vehicle for field exploration. In contrast to existing works concerned with maximizing the amount of information about spatial fields, this work considers efficient exploration of spatiotemporal fields with unknown distributions and seeks minimum-time trajectories of the vehicle while respecting a cumulative information constraint. In this work, upon adopting the observability constant as an information measure for expressing the cumulative information constraint, the existence of a minimum-time trajectory is proven under mild conditions. Given the spatiotemporal nature, the problem is modeled as a Markov decision process (MDP), for which a reinforcement learning (RL) algorithm is proposed to learn a continuous planning policy. To accelerate the policy learning, we design a new reward function by leveraging field approximations, which is demonstrated to yield dense rewards. Simulations show that the learned policy can steer the vehicle to achieve an efficient exploration, and it outperforms the commonly-used coverage planning method in terms of exploration time for sufficient cumulative information.

Original languageEnglish
Pages (from-to)17216-17226
Number of pages11
JournalIEEE Transactions on Neural Networks and Learning Systems
Volume35
Issue number12
DOIs
Publication statusPublished - 2024

Keywords

  • Autonomous vehicles
  • optimal control
  • reinforcement learning (RL)
  • trajectory optimization

Fingerprint

Dive into the research topics of 'Informative Trajectory Planning Using Reinforcement Learning for Minimum-Time Exploration of Spatiotemporal Fields'. Together they form a unique fingerprint.

Cite this