Dynamic Trajectory Control and User Association for Unmanned-Aerial-Vehicle-Assisted Mobile Edge Computing: A Deep Reinforcement Learning Approach
Libo Wang, Xiangyin Zhang, Kaiyu Qin, Zhuwei Wang, Hang Yin, Jiayi Zhou, Deyu Song
University of Electronic Science and Technology of China Institute of Optics and Electronics, Chinese Academy of Sciences Beijing University of Technology Chinese Academy of Sciences
阅读操作
确认中在文库中上传 PDF 后可生成中文音频讲解。
摘要与影响
Mobile edge computing (MEC) has become an effective framework for latency-sensitive and computation-intensive applications by deploying computing resources at network edge. The unmanned aerial vehicle (UAV)-assisted MEC leverages UAV mobility and communication advantages to enable services in dynamic environments, where frequent adjustments to flight trajectories and user association are required due to dynamic factors such as time-varying task requirements, user mobility, and communication environment variation. This paper addresses the joint optimization problem of UAV flight trajectory control and user association in dynamic environments, which explicitly incorporates the constraints governed by UAV flight dynamics. The joint problem is formulated as a non-convex optimization formulation that involves continuous–discrete hybrid decision variables. To overcome the inherent complexity of this problem, a novel proximal policy optimization-based dynamic control (PPO-DC) algorithm is developed. This algorithm aims to reduce the weighted combination of delay and energy consumption by dynamically controlling the UAV trajectory and user association. The numerical results validate that the PPO-DC algorithm successfully enables real-time UAV trajectory control under flight dynamics constraints, ensuring feasible and efficient flight trajectory. Compared to the state-of-the-art hybrid-action deep reinforcement learning (DRL) algorithms or metaheuristics, the PPO-DC achieves notable improvements in system performance by simultaneously lowering system delay and energy consumption.
逐年被引趋势
关键指标
同类平均 = 1
同领域 · 同年份 · 同类型
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:摘要
可就本文提问;依据不足时会说明。
学术脉络
学科主题
工程Autonomous Vehicle Technology and Safety
Evacuation and Crowd Dynamics
参考文献 42
此处列出前 3 条
引用本文 2
按被引量排序,此处列出前 3 条