Skip to main navigation Skip to search Skip to main content

Adaptive Mode Switching in AoI-Aware Multi-UAV Hybrid MEC-DC Networks: A Multi-Agent Reinforcement Learning Approach

  • Beijing Institute of Technology

Research output: Contribution to journalArticlepeer-review

Abstract

Multi-UAV networks are promising for supporting time-sensitive IoT applications, yet most existing studies focus on a single service type and fail to address the coexistence of heterogeneous tasks with fundamentally different timeliness and resource characteristics. In hybrid MEC–DC systems, data collection (DC) and mobile edge computing (MEC) tasks exhibit distinct age of information (AoI) evolution rules and computation–communication couplings, which makes the AoI-aware joint optimization of trajectory planning, task scheduling, and service mode selection under energy, mobility, and communication constraints highly challenging. To tackle these challenges, we propose an adaptive mode-switching multi-agent reinforcement learning framework (AMS-MARL) based on heterogeneous-agent proximal policy optimization (HAPPO). Specifically, a randomized agent update order is employed to decompose the joint advantage into sequential individual advantages, enabling stable and decentralized learning. In addition, a rank-based adaptive reward shaping mechanism is designed to balance information freshness across heterogeneous sensor nodes (SNs) by adjusting reward weights based on AoI deviation from the global average. Extensive simulations under diverse spatial distributions, task ratios, and packet sizes show that AMS-MARL consistently outperforms state-of-the-art baselines in reducing AoI and exhibits strong robustness across varying system settings.

Original languageEnglish
JournalIEEE Transactions on Mobile Computing
DOIs
Publication statusAccepted/In press - 2026

Keywords

  • Adaptive mode switching
  • age of information (AoI)
  • hybrid MEC–DC systems
  • multi -agent reinforcement learning (MARL)
  • multi -UAV networks

Fingerprint

Dive into the research topics of 'Adaptive Mode Switching in AoI-Aware Multi-UAV Hybrid MEC-DC Networks: A Multi-Agent Reinforcement Learning Approach'. Together they form a unique fingerprint.

Cite this