Hybrid Car-Following Strategy Based on Deep Deterministic Policy Gradient and Cooperative Adaptive Cruise Control

巡航控制 自适应控制 计算机科学 控制理论(社会学) 巡航 控制(管理) 控制工程 工程类 航空航天工程 人工智能
作者
Ruidong Yan,Rui Jiang,Bin Jia,Jin Huang,Diange Yang
出处
期刊:IEEE Transactions on Automation Science and Engineering [Institute of Electrical and Electronics Engineers]
卷期号:19 (4): 2816-2824 被引量:46
标识
DOI:10.1109/tase.2021.3100709
摘要

Deep deterministic policy gradient (DDPG)-based car-following strategy can break through the constraints of the differential equation model due to the ability of exploration on complex environments. However, the car-following performance of DDPG is usually degraded by unreasonable reward function design, insufficient training, and low sampling efficiency. In order to solve this kind of problem, a hybrid car-following strategy based on DDPG and cooperative adaptive cruise control (CACC) is proposed. First, the car-following process is modeled as the Markov decision process to calculate CACC and DDPG simultaneously at each frame. Given a current state, two actions are obtained from CACC and DDPG, respectively. Then, an optimal action, corresponding to the one offering a larger reward, is chosen as the output of the hybrid strategy. Meanwhile, a rule is designed to ensure that the change rate of acceleration is smaller than the desired value. Therefore, the proposed strategy not only guarantees the basic performance of car-following through CACC but also makes full use of the advantages of exploration on complex environments via DDPG. Finally, simulation results show that the car-following performance of the proposed strategy is improved compared with that of DDPG and CACC. Note to Practitioners—This article presents a new car-following strategy, which avoids the impact of deep deterministic policy gradient (DDPG) performance degradation on the system. In the proposed strategy, DDPG is replaced with cooperative adaptive cruise control (CACC) when the performance of DDPG is worse than that of CACC. Meanwhile, a switching rule is designed to guarantee that the change rate of acceleration is smaller than the threshold. Simulation results show that the performance of hybrid car-following strategy has been improved compared with that of only using CACC or DDPG. Moreover, the proposed strategy has the advantages of low computational burden, high real-time performance, and good scalability.

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
董雅山发布了新的文献求助10
刚刚
rkai发布了新的文献求助10
1秒前
勤恳易真完成签到,获得积分10
1秒前
cc完成签到,获得积分10
1秒前
小小牛马应助哈哈哈哈哈采纳,获得10
2秒前
化学把我害惨了完成签到,获得积分10
2秒前
3秒前
所所应助搬砖中采纳,获得10
3秒前
俭朴大碗完成签到,获得积分10
3秒前
CipherSage应助人生如梦采纳,获得10
3秒前
3秒前
CodeCraft应助张1采纳,获得10
3秒前
Orange应助qingchidue采纳,获得10
3秒前
3秒前
越努力,越幸运完成签到,获得积分20
3秒前
怡yi发布了新的文献求助20
4秒前
4秒前
刘欣完成签到,获得积分10
4秒前
钟于发布了新的文献求助10
4秒前
4秒前
无奈电灯胆完成签到,获得积分10
4秒前
无心完成签到,获得积分10
5秒前
奶味蓝完成签到,获得积分10
5秒前
Yy00完成签到,获得积分10
5秒前
锦鲤完成签到,获得积分10
6秒前
赘婿应助青青采纳,获得10
6秒前
amu完成签到,获得积分20
6秒前
6秒前
6秒前
文静依萱发布了新的文献求助10
7秒前
萱棚完成签到 ,获得积分10
7秒前
CodeCraft应助jiangnantingyu采纳,获得10
7秒前
7秒前
CC完成签到,获得积分20
7秒前
若菲发布了新的文献求助10
8秒前
cc完成签到,获得积分10
8秒前
墨染樱飞卿清叙完成签到,获得积分10
8秒前
dolabmu完成签到 ,获得积分10
8秒前
vergil完成签到,获得积分10
8秒前
你可真下饭完成签到 ,获得积分10
8秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Introducing the Learning Sciences 1000
2026年中国辛酸癸酸聚乙二醇甘油酯行业市场现状调查及投资机会研判报告 1000
2026年中国辛酸癸酸聚乙二醇甘油酯行业市场规模及竞争格局分析报告 1000
Resiliency Scale for Adolescents--Chinese Version 800
48V Low-voltage Power Distribution Network (PDN) Architecture Industry Report, 2024 800
Fundamentals of Pharmaceutical and Biologics Regulations: A Global Perspective, Second Edition 700
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7324894
求助须知:如何正确求助?哪些是违规求助? 8940274
关于积分的说明 18956752
捐赠科研通 6981684
什么是DOI,文献DOI怎么找? 3215499
关于科研通互助平台的介绍 2382798
邀请新用户注册赠送积分活动 2194821