Scholay

学术搜索 · AI 审稿 · LaTeX 协作

Deep Reinforcement Learning Enabled UAV Trajectory Optimization for A2G Communication Systems

作者:Hao Jiang, Xuting Pan, Wangqi Shi, Lin-Zhou Zeng, Zhen Chen, Feng Shu, Jiangzhou Wang · 发表于:IEEE Transactions on Cognitive Communications and Networking · 年份:2026 · DOI:10.1109/tccn.2025.3633751 · 被引用次数:27 · 研究领域:Computer Science

This paper proposes a geometry-based air-to-ground (A2G) channel model that captures time-varying velocities of uncrewed aerial vehicles (UAVs), mobile receiver (MR), and scatterers in realistic propagation environments. To enhance the UAV energy efficiency during flight, the model integrates deep reinforcement learning (DRL) to enable real-time decision-making within each discrete time interval. Specifically, the approach utilizes an enhanced twin-delayed deep deterministic policy gradient (TD3) algorithm with a dual-layer actor network and twin critic networks, which further improves its ability to efficiently handle complex decision-making tasks in dynamic environments. Additionally, key statistical characteristics are derived rigorously and confirmed through numerical experiments, including spatial, temporal, and frequency domain correlations. The numerical results demonstrate that the proposed approach achieves superior energy efficiency and stability over the soft actor-critic (SAC), proximal policy optimization (PPO), and deep deterministic policy gradient (DDPG) methods, ensuring optimal communication performance and UAV trajectories. The study underscores the effectiveness of DRL in optimizing UAV trajectories, offering valuable insights for the design of advanced A2G wireless communication systems.