亲爱的研友该休息了!由于当前在线用户较少,发布求助请尽量完整地填写文献信息,科研通机器人24小时在线,伴您度过漫漫科研夜!身体可是革命的本钱,早点休息,好梦!

Multi-Agent Deep Reinforcement Learning for Urban Traffic Light Control in Vehicular Networks

强化学习 交叉口(航空) 计算机科学 过程(计算) 交通拥挤 分布式计算 人工智能 工程类 运输工程 操作系统
作者
Tong Wu,Pan Zhou,Kai Liu,Yali Yuan,Xiumin Wang,Huawei Huang,Dapeng Wu
出处
期刊:IEEE Transactions on Vehicular Technology [Institute of Electrical and Electronics Engineers]
卷期号:69 (8): 8243-8256 被引量:179
标识
DOI:10.1109/tvt.2020.2997896
摘要

As urban traffic condition is diverse and complicated, applying reinforcement learning to reduce traffic congestion becomes one of the hot and promising topics. Especially, how to coordinate the traffic light controllers of multiple intersections is a key challenge for multi-agent reinforcement learning (MARL). Most existing MARL studies are based on traditional Q-learning, but unstable environment leads to poor learning in the complicated and dynamic traffic scenarios. In this paper, we propose a novel multi-agent recurrent deep deterministic policy gradient (MARDDPG) algorithm based on deep deterministic policy gradient (DDPG) algorithm for traffic light control (TLC) in vehiclar networks. Specifically, the centralized learning in each critic network enables each agent to estimate the policies of other agents in the decision-making process and each agent can coordinate with each other, alleviating the problem of poor learning performance caused by environmental instability. The decentralized execution enables each agent to make decisions independently. We share parameters in actor networks to speed up the training process and reduce the memory footprint. The addition of LSTM is beneficial to alleviate the instability of the environment caused by partial observable state. We utilize surveillance cameras and vehicular networks to collect status information for each intersection. Unlike previous work, we have not only considered the vehicle but also considered the pedestrians waiting to pass through the intersection. Moreover, we also set different priorities for buses and ordinary vehicles. The experimental results in a vehicular network show that our method can run stably in various scenarios and coordinate multiple intersections, which significantly reduces vehicle congestion and pedestrian congestion.

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
丘比特应助科研通管家采纳,获得10
15秒前
情怀应助科研通管家采纳,获得10
16秒前
酷波er应助科研通管家采纳,获得10
16秒前
优美的烙完成签到,获得积分10
21秒前
39秒前
冷艳的萝莉完成签到,获得积分10
43秒前
46秒前
53秒前
受伤的爆米花完成签到,获得积分10
1分钟前
1分钟前
光亮靳完成签到,获得积分10
1分钟前
深情安青应助顺心面包采纳,获得10
1分钟前
高兴中心完成签到,获得积分10
1分钟前
sbt完成签到 ,获得积分10
1分钟前
漂亮孤风完成签到,获得积分10
1分钟前
鲁成危完成签到,获得积分10
2分钟前
碧蓝山灵完成签到,获得积分10
2分钟前
大模型应助科研通管家采纳,获得10
2分钟前
2分钟前
2分钟前
Axel发布了新的文献求助10
2分钟前
顺心面包发布了新的文献求助10
2分钟前
深情的从丹完成签到,获得积分10
2分钟前
呆萌尔风完成签到,获得积分10
3分钟前
柔弱的妙旋完成签到,获得积分10
3分钟前
Axel完成签到,获得积分10
3分钟前
思源应助科研通管家采纳,获得10
4分钟前
大个应助科研通管家采纳,获得10
4分钟前
4分钟前
感动的易真完成签到,获得积分10
4分钟前
多情敏完成签到,获得积分10
4分钟前
rohiga完成签到,获得积分10
4分钟前
风趣的苑博完成签到,获得积分10
4分钟前
自觉的猕猴桃完成签到,获得积分10
4分钟前
无花果应助ZcLee采纳,获得10
4分钟前
成就的灵槐完成签到,获得积分10
4分钟前
ww完成签到,获得积分10
5分钟前
机灵自中完成签到,获得积分10
5分钟前
野性的幼萱完成签到,获得积分10
5分钟前
务实雪珍完成签到,获得积分10
5分钟前
高分求助中
On lateral buckling of armouring wires in flexible pipes 10000
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Navigating Normative Orders. Interdisciplinary Perspectives 800
Essentials of Carbohydrate Chemistry and Biochemistry, 4th Edition 700
1 Peter and Christ's Descent to the Dead in Its Early Christian Reception 700
Organizational Behavior 510
Management and the Arts 510
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7744638
求助须知:如何正确求助?哪些是违规求助? 9292443
关于积分的说明 20212709
捐赠科研通 7323505
什么是DOI,文献DOI怎么找? 3307639
关于科研通互助平台的介绍 2459546
邀请新用户注册赠送积分活动 2318638