UAV Trajectory Optimization for Directional THz Links Using Deep Reinforcement Learning
As an alternative solution for quick disaster recovery of backhaul/fronthaul links, in this paper, a dynamic unmanned aerial vehicles (UAV)-assisted heterogeneous (HetNet) network equipped with directional terahertz (THz) antennas is studied to solve the problem of transferring traffic of distributed small cells. To this end, we first characterize a detailed three-dimensional modeling of the dynamic UAV-assisted HetNet, and then, we formulate the problem for UAV trajectory to minimize the maximum outage probability of directional THz links. Then, using deep reinforcement learning (DRL) method, we propose an efficient algorithm to learn the optimal trajectory. Finally, using simulations, we investigate the performance of the proposed DRL-based trajectory method.
Code (0)
등록된 구현이 없습니다.
Tasks
Deep Reinforcement Learningreinforcement-learningSimilar Papers 제목 키워드 기반
Directional Alignment Mitigates Reward Hacking in Reinforcement Learning for Language Models
Reward hacking arises when a model improves a proxy reward by exploiting shortcuts rather than solving the intended task. We study this failure mode through the geometry of reinforcement learning updates in language mode…
Reinforcement LearningMathematical ReasoningBiC-MPPI: Goal-Pursuing, Sampling-Based Bidirectional Rollout Clustering Path Integral for Trajectory Optimization
This paper introduces the Bidirectional Clustered MPPI (BiC-MPPI) algorithm, a novel trajectory optimization method aimed at enhancing goal-directed guidance within the Model Predictive Path Integral (MPPI) framework. Bi…
Autonomous NavigationTrajectory PlanningA Generalized Pointing Error Model for FSO Links with Fixed-Wing UAVs for 6G: Analysis and Trajectory Optimization
Free-space optical (FSO) communication is a promising solution to support wireless backhaul links in emerging 6G non-terrestrial networks. At the link level, pointing errors in FSO links can significantly impact capacity…
AcroRL: Learning Aggressive Quadrotor Inversion using Bidirectional Thrust
Bidirectional thrust grants quadrotors a second equilibrium condition and increased control authority, expanding the envelope of possible aggressive maneuvers and enabling inverted flight, perching, and sensing. Prior ge…
Reinforcement LearningReinforcement Learning for Mitigating Intermittent Interference in Terahertz Communication Networks
Emerging wireless services with extremely high data rate requirements, such as real-time extended reality applications, mandate novel solutions to further increase the capacity of future wireless networks. In this regard…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)