Simulation of Vehicle Interaction Behavior in Merging Scenarios: A Deep Maximum Entropy-Inverse Reinforcement Learning Method Combined With Game Theory

计算机科学 强化学习 马尔可夫决策过程 过程(计算) 熵(时间箭头) 最大熵原理 人工智能 博弈论 马尔可夫过程 机器学习 模拟 数学 数理经济学 统计 物理 量子力学 操作系统
作者
Wenli Li,Fanke Qiu,Lingxi Li,Yinan Zhang,Kan Wang
出处
期刊:IEEE transactions on intelligent vehicles [Institute of Electrical and Electronics Engineers]
卷期号:9 (1): 1079-1093 被引量:15
标识
DOI:10.1109/tiv.2023.3323138
摘要

Simulation testing based on virtual scenarios can improve the efficiency of safety testing for high-level autonomous vehicles (AVs). In most traffic scenarios, such as merging scenarios, the interactions between vehicles are a game process. Therefore, a critical factor is to accurately simulate the game and interaction processes between the background vehicle (BV) and AV in the test environment. With the increasing availability of natural driving data, a data-driven approach can be introduced to identify the underlying driving behavior patterns in actual driving data. Thus, this paper proposes a data-driven method for modeling BV behavior for AV testing in virtual scenarios. The method describes the vehicle decision process in the merging scenario as a standard Markov decision process (MDP). Based on game theory, we considered the BV as a game subject to illustrate the vehicle interaction process. Furthermore, a deep maximum entropy-inverse reinforcement learning combined with the game matrix is proposed to identify the reward function that describes BV behavior. The obtained reward function is used to design a deep Q-network algorithm to simulate the behavior of BV. Finally, the effectiveness and feasibility of the proposed method are verified by comparing it with natural driving data. Moreover, we performed comparative tests with the other two baseline methods; the results show that the proposed method can accurately simulate the interaction behaviors between vehicles in the virtual scenarios.

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
2秒前
QQ完成签到,获得积分10
2秒前
呆一起完成签到 ,获得积分10
3秒前
勋章完成签到 ,获得积分10
5秒前
Literature的应助被and999采纳,获得100
5秒前
Brave发布了新的文献求助10
6秒前
7秒前
离大谱完成签到,获得积分10
10秒前
安详的老五完成签到,获得积分10
11秒前
张小秉完成签到,获得积分10
11秒前
优雅含灵完成签到 ,获得积分10
14秒前
MJN完成签到 ,获得积分10
18秒前
如愿常隐行完成签到 ,获得积分10
18秒前
22秒前
体贴凌寒完成签到 ,获得积分10
23秒前
星先生完成签到 ,获得积分10
24秒前
Zsy完成签到,获得积分10
27秒前
天道酬勤完成签到,获得积分10
28秒前
爱学数学的数学小白完成签到,获得积分10
28秒前
30秒前
wddd333333完成签到,获得积分10
34秒前
王乐乐哈完成签到 ,获得积分10
35秒前
36秒前
靓丽的采白完成签到,获得积分10
38秒前
amonke007完成签到,获得积分10
38秒前
CES_SH完成签到,获得积分10
43秒前
HJJHJH完成签到,获得积分10
44秒前
晓风完成签到,获得积分0
44秒前
46秒前
ding的应助被科研通管家采纳,获得10
46秒前
CipherSage的应助被科研通管家采纳,获得10
46秒前
桐桐的应助被科研通管家采纳,获得10
46秒前
49秒前
111完成签到 ,获得积分10
50秒前
aaa0001984完成签到,获得积分0
51秒前
美满的珠完成签到 ,获得积分10
55秒前
00完成签到 ,获得积分10
55秒前
时尚的访琴完成签到 ,获得积分10
55秒前
牛马完成签到 ,获得积分10
56秒前
茄子完成签到 ,获得积分10
57秒前
高分求助中
(应助此贴封号)通过应助OA文献获取积分 10000
The Student's Guide to Social Neuroscience 800
Rosenblum, Global Change Biology 800
Computational Chemical Reaction Engineering: Modeling, Simulation, and Design with MATLAB 600
Organizational Behavior 510
Management and the Arts 510
Production Logging: Theoretical and Interpretive Elements 400
热门求助领域 (近24小时)
化学 材料科学 医学 生物 计算机科学 工程类 纳米技术 内科学 物理 有机化学 化学工程 生物化学 复合材料 光电子学 细胞生物学 心理学 量子力学 催化作用 物理化学 电极
热门帖子
关注 科研通微信公众号,转发送积分 7813092
求助须知:如何正确求助?哪些是违规求助? 9343854
关于积分的说明 20519899
捐赠科研通 7405895
什么是DOI,文献DOI怎么找? 3330328
关于科研通互助平台的介绍 2476983
邀请新用户注册赠送积分活动 2349879