General multi-agent reinforcement learning integrating adaptive manoeuvre strategy for real-time multi-aircraft conflict resolution

强化学习 计算机科学 冲突解决 钢筋 航空学 人工智能 多智能体系统 适应性学习 实时计算 工程类 运筹学 政治学 法学 结构工程
作者
Yutong Chen,Minghua Hu,Lei Yang,Yan Xu,Hua Xie
出处
期刊:Transportation Research Part C-emerging Technologies [Elsevier BV]
卷期号:151: 104125-104125 被引量:26
标识
DOI:10.1016/j.trc.2023.104125
摘要

Reinforcement learning (RL) techniques are under investigation for resolving conflict in air traffic management (ATM), exploiting their computational capabilities and ability to cope with flight uncertainty. However, the limitations of generalisation make it difficult for existing RL-based conflict resolution (CR) methods to be effective in practice. This paper proposes a general multi-agent reinforcement learning (MARL) method that integrates an adaptive manoeuvre strategy to enhance both the solution’s efficiency and the model’s generalisation in multi-aircraft conflict resolution (MACR). A partial observation approach based on the imminent threat detection sectors is used to gather critical environmental information, enabling the model to be applied in arbitrary scenarios. Agents are trained to provide the correct flight intention (such as increasing speed and yawing to the left), while an adaptive manoeuvre strategy generates the specific manoeuvre (speed and heading parameters) based on the flight intention. To address flight uncertainty and performance challenges caused by the intrinsic non-stationarity in MARL, a warning area for each aircraft is introduced. We employ a state-of-the-art Deep Q-learning Network (DQN) method, Rainbow DQN, to improve the efficiency of the RL algorithm. The multi-agent system is trained and deployed in a distributed manner to adapt to real-world scenarios. A sensitivity analysis of uncertainty levels and warning area sizes is conducted to explore their impact on the proposed method. Simulation experiments confirm the effectiveness of the training and generalisation of the proposed method.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
吴宣京完成签到,获得积分10
1秒前
lf-leo完成签到,获得积分10
1秒前
FBI完成签到,获得积分10
2秒前
Ge发布了新的文献求助10
3秒前
print完成签到,获得积分20
4秒前
优秀念柏完成签到,获得积分10
5秒前
yuchuncheng完成签到,获得积分10
5秒前
畅快呼吸完成签到 ,获得积分10
5秒前
酷波er的应助被烟火会翻滚采纳,获得10
6秒前
科研通AI6.2的应助被kkkay采纳,获得10
6秒前
SciGPT的应助被六六采纳,获得30
6秒前
realtimes完成签到,获得积分10
8秒前
Lucy关注了科研通微信公众号
9秒前
AAA完成签到,获得积分10
9秒前
dong完成签到 ,获得积分10
9秒前
9秒前
可耐的天菱完成签到,获得积分10
10秒前
琪梦李完成签到 ,获得积分10
11秒前
11秒前
世界尽头完成签到,获得积分10
11秒前
11秒前
cdercder的应助被陆上飞采纳,获得10
12秒前
科研通AI6.2的应助被kkkay采纳,获得10
12秒前
iitj发布了新的文献求助10
12秒前
我是老大的应助被222采纳,获得10
13秒前
meshia的应助被奋斗的秋凌采纳,获得10
13秒前
梅花易数完成签到,获得积分10
14秒前
12306完成签到,获得积分10
14秒前
Yw_M完成签到,获得积分10
15秒前
鱼鱼鱼的阁楼主子完成签到,获得积分10
15秒前
wangyinghao发布了新的文献求助10
16秒前
Lucy发布了新的文献求助10
17秒前
innocence完成签到,获得积分10
17秒前
豆丁小猫完成签到,获得积分10
18秒前
南山无梅落完成签到,获得积分10
18秒前
科研通AI6.4的应助被kkkay采纳,获得10
18秒前
zg完成签到 ,获得积分10
19秒前
21秒前
jianhua完成签到,获得积分10
21秒前
uus13发布了新的文献求助10
22秒前
高分求助中
(应助此贴封号)通过应助OA文献获取积分 10000
Rosenblum, Global Change Biology 800
Organizational Behavior 510
Management and the Arts 510
Convergent and bidirectional strategies towards the total synthesis of hemibrevetoxin B 300
Geschichtliche Grundbegriffe (GGB), Band 5: Pro–Soz 300
Die Religion in Geschichte und Gegenwart (RGG), 4. Auflage, Band 7: R–S 300
热门求助领域 (近24小时)
化学 材料科学 医学 生物 计算机科学 工程类 纳米技术 内科学 物理 有机化学 化学工程 生物化学 复合材料 光电子学 细胞生物学 心理学 量子力学 催化作用 物理化学 电极
热门帖子
关注 科研通微信公众号,转发送积分 7797841
求助须知:如何正确求助?哪些是违规求助? 9333151
关于积分的说明 20458083
捐赠科研通 7388507
什么是DOI,文献DOI怎么找? 3325487
关于科研通互助平台的介绍 2472847
邀请新用户注册赠送积分活动 2342846