避障
趋同(经济学)
计算机科学
障碍物
强化学习
人工智能
增强学习
实时计算
控制理论(社会学)
控制(管理)
移动机器人
机器人
政治学
经济增长
经济
法学
作者
Songyue Yang,Guizhen Yu,Zhijun Meng,Zhangyu Wang,Han Li
摘要
In the intelligent unmanned systems, unmanned aerial vehicle (UAV) obstacle avoidance technology is the core and primary condition. Traditional algorithms are not suitable for obstacle avoidance in complex and changeable environments based on the limited sensors on UAVs. In this article, we use an end-to-end deep reinforcement learning (DRL) algorithm to achieve the UAV autonomously avoid obstacles. For the problem of slow convergence in DRL, a Multi-Branch (MB) network structure is proposed to ensure that the algorithm can get good performance in the early stage; for non-optimal decision-making problems caused by overestimation, the Revise Q-value (RQ) algorithm is proposed to ensure that the agent can choose the optimal strategy for obstacle avoidance. According to the flying characteristics of the rotor UAV, we build a V-Rep 3D physical simulation environment to test the obstacle avoidance performance. And experiments show that the improved algorithm can accelerate the convergence speed of agent and the average return of the round is increased by 25%.
科研通智能强力驱动
Strongly Powered by AbleSci AI