跳到主要导航 跳到搜索 跳到主要内容

Dual-Time-Scale Framework for Joint Optimization of Service Caching and UAV Trajectory Based on Self-Attention Deep Reinforcement Learning

  • Xuewei Zhang
  • , Jiantao Li
  • , Yuan Ren*
  • , Fan Jiang
  • , Junxuan Wang
  • , Jie Zeng
  • *此作品的通讯作者
  • Xi'an Institute of Posts and Telecommunications
  • Beijing Institute of Technology

科研成果: 期刊稿件文章同行评审

摘要

Nowadays, unmanned aerial vehicle (UAV)-assisted mobile edge computing (MEC) has been recognized as a promising technique for flexibly handling computation tasks in 5G advanced and 6G networks. This article investigates the joint optimization of service caching and computation offloading within a dual time-scale framework. We maximize the caching utility and minimize the task processing delay by jointly optimizing service caching policies, UAV flight trajectory, and computation offloading decisions. Specifically, for the long-term problem, we use the latent Dirichlet allocation (LDA) model to predict user preferences and propose a Lagrangian dual decomposition-based algorithm. For the short-term problem, a self-attention-based multiagent proximal policy optimization (MAPPO) algorithm is designed. Under the centralized training with decentralized execution (CTDE) framework, this algorithm integrates a multihead self-attention mechanism with curriculum learning. Each UAV is regarded as an agent, and a self-attention encoder (SAE) is integrated at the front-end of each actor network. This enables the agent to dynamically capture the relative importance between itself and all users, and context-aware features are extracted to make more intelligent and trajectory designs. Through extensive simulation experiments, the long-term algorithm yields the performance improvements of 76.5% in cache hit rate and 66% in caching utility, as compared to the second-best baseline. In dynamic scenarios, the short-term algorithm achieves a 16% reduction in total processing latency with respect to the proximal policy optimization (PPO) policy.

源语言英语
页(从-至)34629-34645
页数17
期刊IEEE Internet of Things Journal
13
15
DOI
出版状态已出版 - 1 8月 2026
已对外发布

学术指纹

探究 'Dual-Time-Scale Framework for Joint Optimization of Service Caching and UAV Trajectory Based on Self-Attention Deep Reinforcement Learning' 的科研主题。它们共同构成独一无二的学术指纹。

引用此