Towards Robust Decision-Making for Autonomous Driving on Highway

强化学习 计算机科学 人工智能 分布(数学) 可靠性(半导体) 机器学习 运筹学 工程类 数学 数学分析 功率(物理) 物理 量子力学
作者
Kai Yang,Xiaolin Tang,Sen Qiu,Shufeng Jin,Zichun Wei,Hong Wang
出处
期刊:IEEE Transactions on Vehicular Technology [Institute of Electrical and Electronics Engineers]
卷期号:72 (9): 11251-11263 被引量:88
标识
DOI:10.1109/tvt.2023.3268500
摘要

Reinforcement learning (RL) methods are commonly regarded as effective solutions for designing intelligent driving policies. Nonetheless, even if the RL policy is converged after training, it is notoriously difficult to ensure safety. In particular, RL policy is susceptible to insecurity in the presence of long-tail or unseen traffic scenarios, i.e. , out-of-distribution test data. Therefore, the design of the RL-based decision-making method must account for this shift in distribution. This paper proposes a robust decision-making framework for autonomous driving on the highway to improve driving safety. First, a Deep Deterministic Policy Gradient (DDPG)-based RL policy that directly maps observations to actions is constructed. Subsequently, the model uncertainty of the DDPG policy is evaluated at runtime to quantify the policy's reliability and identify unseen scenarios. In addition, a complementary principle-based policy is developed using the intelligent driver model (IDM) and the model for minimizing overall braking induced by lane changes (MOBIL). It will take over the DDPG policy when encountering unseen scenarios to guarantee a lower-bound performance of the decision-making system. Finally, the proposed method is implemented on an embedded system, i.e. , NVIDIA Jetson AGX Xavier, and out-of-training distribution challenging cases are considered in the experiment, i.e. , observation with sensor noise, traffic density increasing significantly, objects falling from the front vehicle, and road construction causing temporal changes in road structure. Results indicate that the proposed framework outperforms state-of-the-art benchmarks. Additionally, the code is provided.

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
花花应助西米露采纳,获得20
刚刚
谢非凡发布了新的文献求助30
1秒前
思源应助清溪采纳,获得10
1秒前
2秒前
Ava应助百川采纳,获得10
2秒前
dagongren完成签到,获得积分10
2秒前
深情安青应助孙朱珠采纳,获得10
3秒前
Lasse发布了新的文献求助10
3秒前
狂野傲珊发布了新的文献求助10
3秒前
CY完成签到,获得积分10
3秒前
顾顾发布了新的文献求助10
4秒前
心灵美盼烟完成签到,获得积分10
5秒前
touch完成签到,获得积分10
6秒前
6秒前
6秒前
lkjhg应助小呆采纳,获得10
7秒前
十年发布了新的文献求助10
8秒前
盛夏如花发布了新的文献求助10
8秒前
8秒前
meng发布了新的文献求助10
9秒前
10秒前
思源应助活力鑫磊采纳,获得10
10秒前
111完成签到,获得积分20
10秒前
彭于晏应助君猪采纳,获得10
11秒前
研友_VZG7GZ应助crash采纳,获得10
11秒前
Lily完成签到,获得积分10
11秒前
anlikek发布了新的文献求助10
11秒前
12秒前
烟花应助自然狗采纳,获得10
12秒前
陈锦雯完成签到,获得积分10
14秒前
谢非凡完成签到,获得积分10
14秒前
奋斗易真发布了新的文献求助10
15秒前
桃桃子发布了新的文献求助10
15秒前
15秒前
鲤鱼笑南完成签到,获得积分10
16秒前
16秒前
星辰大海应助猪头军师采纳,获得10
17秒前
搜集达人应助嘉嘉嘉嘉一采纳,获得10
17秒前
17秒前
凝心发布了新的文献求助10
18秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Single Cell Analysis of the Tumor Microenvironment Landscape Across the Disease Spectrum of Multiple Myeloma 1000
2026年中国辛酸癸酸聚乙二醇甘油酯行业市场现状调查及投资机会研判报告 1000
2026年中国辛酸癸酸聚乙二醇甘油酯行业市场规模及竞争格局分析报告 1000
Fundamentals of Pharmaceutical and Biologics Regulations: A Global Perspective, Second Edition 700
The Cambridge History of China 英文版16册 600
作者名:Kristopher P. Plain,悉尼大学的,目前只能查到其四篇论文,想找到其博士论文 550
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7329737
求助须知:如何正确求助?哪些是违规求助? 8944089
关于积分的说明 18972505
捐赠科研通 6985029
什么是DOI,文献DOI怎么找? 3216528
关于科研通互助平台的介绍 2383224
邀请新用户注册赠送积分活动 2196140