植绒(纹理)
强化学习
航路点
计算机科学
规划师
人工智能
机器人
风速
模拟
搜救
控制工程
实时计算
工程类
材料科学
复合材料
物理
气象学
作者
Pramod Abichandani,Christian Speck,Donald J. Bucci,William A. McIntyre,Deepan Lobo
出处
期刊:IEEE Access
[Institute of Electrical and Electronics Engineers]
日期:2021-01-01
卷期号:9: 132491-132507
被引量:11
标识
DOI:10.1109/access.2021.3115711
摘要
Enabling coordinated motion of multiple quadrotors is an active area of research in the field of small unmanned aerial vehicles (sUAVs). While there are many techniques found in the literature that address the problem, these studies are limited to simulation results and seldom account for wind disturbances. This paper presents the experimental validation of a decentralized planner based on multi-objective reinforcement learning (RL) that achieves waypoint-based flocking (separation, velocity alignment, and cohesion) for multiple quadrotors in the presence of wind gusts. The planner is learned using an object-focused, greatest mass, state-action-reward-state-action (OF-GM-SARSA) approach. The Dryden wind gust model is used to simulate wind gusts during hardware-in-the-loop (HWIL) tests. The hardware and software architecture developed for the multi-quadrotor flocking controller is described in detail. HWIL and outdoor flight tests results show that the trained RL planner can generalize the flocking behaviors learned in training to the real-world flight dynamics of the DJI M100 quadrotor in windy conditions.
科研通智能强力驱动
Strongly Powered by AbleSci AI