已入深夜,您辛苦了!由于当前在线用户较少,发布求助请尽量完整地填写文献信息,科研通机器人24小时在线,伴您度过漫漫科研夜!祝你早点完成任务,早点休息,好梦!

PerlAD: Towards Enhanced Closed-loop End-to-end Autonomous Driving with Pseudo-simulation-based Reinforcement Learning

强化学习 计算机科学 人工智能 水准点(测量) 规划师 桥(图论) 运动规划 时差学习 可靠性(半导体) 桥接(联网) 渲染(计算机图形) 机器学习 机器人 路径(计算) 计算 车辆动力学 移动机器人 边距(机器学习) 人群 人工神经网络 仿真
作者
Yinfeng Gao,Qichao Zhang,Deqing Liu,Zhongpu Xia,Guang Li,Kun Ma,Guang Chen,Hangjun Ye,Long Chen,Da‐Wei Ding,Dongbin Zhao
出处
期刊:Cornell University - arXiv [Cornell University]
摘要

End-to-end autonomous driving policies based on Imitation Learning (IL) often struggle in closed-loop execution due to the misalignment between inadequate open-loop training objectives and real driving requirements. While Reinforcement Learning (RL) offers a solution by directly optimizing driving goals via reward signals, the rendering-based training environments introduce the rendering gap and are inefficient due to high computational costs. To overcome these challenges, we present a novel Pseudo-simulation-based RL method for closed-loop end-to-end autonomous driving, PerlAD. Based on offline datasets, PerlAD constructs a pseudo-simulation that operates in vector space, enabling efficient, rendering-free trial-and-error training. To bridge the gap between static datasets and dynamic closed-loop environments, PerlAD introduces a prediction world model that generates reactive agent trajectories conditioned on the ego vehicle's plan. Furthermore, to facilitate efficient planning, PerlAD utilizes a hierarchical decoupled planner that combines IL for lateral path generation and RL for longitudinal speed optimization. Comprehensive experimental results demonstrate that PerlAD achieves state-of-the-art performance on the Bench2Drive benchmark, surpassing the previous E2E RL method by 10.29% in Driving Score without requiring expensive online interactions. Additional evaluations on the DOS benchmark further confirm its reliability in handling safety-critical occlusion scenarios.

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
缥缈不弱发布了新的文献求助10
2秒前
3秒前
VV发布了新的文献求助10
3秒前
尽落完成签到 ,获得积分10
4秒前
科研通AI6.4应助谷安采纳,获得10
5秒前
5秒前
厚脸皮的含羞草完成签到 ,获得积分10
6秒前
wangruiyang完成签到 ,获得积分10
6秒前
7秒前
呐呐呐完成签到,获得积分10
7秒前
Orange应助lxyyyds采纳,获得10
7秒前
ZMO发布了新的文献求助10
8秒前
坦率灵槐发布了新的文献求助10
11秒前
卿莞尔完成签到 ,获得积分0
14秒前
15秒前
充电宝应助maxztz采纳,获得10
18秒前
18秒前
漂亮糖豆完成签到 ,获得积分10
18秒前
19秒前
完美世界应助古德猫宁采纳,获得10
19秒前
19秒前
天天快乐应助mirutio采纳,获得30
20秒前
完美世界应助认真的不评采纳,获得10
21秒前
拾柒完成签到,获得积分10
22秒前
秋秋完成签到,获得积分10
22秒前
23秒前
23秒前
幸福中心完成签到,获得积分10
23秒前
Umar完成签到,获得积分10
26秒前
26秒前
zhaiyiying应助科研通管家采纳,获得10
27秒前
pokexuejiao应助科研通管家采纳,获得10
27秒前
小马甲应助狂野的月光采纳,获得10
27秒前
aajhajkahna应助科研通管家采纳,获得10
27秒前
Copyright应助科研通管家采纳,获得10
27秒前
27秒前
27秒前
27秒前
aajhajkahna应助科研通管家采纳,获得10
27秒前
小蘑菇应助科研通管家采纳,获得10
27秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Organic Chemistry, 5th Edition 1000
Handbook of Social Psychology and Consumer Behavior 900
Nondestructive Testing Handbook: Vol. 4, Thermal and Infrared Testing (IR), 4th ed 800
日本現代怪異事典 副読本 700
Handbook of Social Identity Research 600
作者名:Kristopher P. Plain,悉尼大学的,目前只能查到其四篇论文,想找到其博士论文 590
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7375262
求助须知:如何正确求助?哪些是违规求助? 8982942
关于积分的说明 19099754
捐赠科研通 7016034
什么是DOI,文献DOI怎么找? 3225850
关于科研通互助平台的介绍 2389160
邀请新用户注册赠送积分活动 2206510