Event-Triggered Deep Reinforcement Learning Using Parallel Control: A Case Study in Autonomous Driving

强化学习 马尔可夫决策过程 计算机科学 构造(python库) 事件(粒子物理) 动作(物理) 人工智能 国家(计算机科学) 控制(管理) 实现(概率) 深度学习 价值网络 最优控制 机器学习 马尔可夫过程 数学优化 算法 数学 计算机网络 统计 商业模式 物理 业务 营销 量子力学
作者
Jingwei Lu,Liyuan Han,Qinglai Wei,Xiao Wang,Xingyuan Dai,Fei‐Yue Wang
出处
期刊:IEEE transactions on intelligent vehicles [Institute of Electrical and Electronics Engineers]
卷期号:8 (4): 2821-2831 被引量:98
标识
DOI:10.1109/tiv.2023.3262132
摘要

This paper utilizes parallel control to investigate the problem of event-triggered deep reinforcement learning and develops an event-triggered deep Q-network (ETDQN) for decision-making of autonomous driving, without training an explicit triggering condition . Based on the framework of parallel control, the developed ETDQN incorporates information of actions into the feedback and constructs a dynamic control policy. First, in the realization of the dynamic control policy, we integrate the current state and the previous action to construct the augmented state as well as the augmented Markov decision process. Meanwhile, it is shown theoretically that the goal of the developed dynamic control policy is to learn the variation rate of the action. The augmented state contains information of the current state and the previous action, which enables the developed ETDQN to directly design the immediate reward considering communication loss. Then, based on dueling double deep Q-network (dueling DDQN), we establish the augmented action-value, value, and advantage functions to directly learn the optimal event-triggered decision-making policy of autonomous driving without an explicit triggering condition. It is worth noticing that the developed ETDQN applies to various deep Q-networks (DQNs). Empirical results demonstrate that, in event-triggered control, the developed ETDQN outperforms dueling DDQN and reduces communication loss effectively.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
1秒前
1秒前
倪二妹发布了新的文献求助10
1秒前
小强完成签到,获得积分10
1秒前
乐乐应助TMAC采纳,获得10
1秒前
细腻听白完成签到,获得积分10
2秒前
芒果布丁完成签到 ,获得积分10
2秒前
少女怪物完成签到 ,获得积分10
2秒前
3秒前
咸蛋黄味曲奇完成签到,获得积分10
3秒前
3秒前
科研通AI6.4应助你好采纳,获得10
4秒前
闻妙完成签到,获得积分10
4秒前
秋风举报PDIF-CN2求助涉嫌违规
4秒前
5秒前
yuaasusanaann完成签到,获得积分10
5秒前
派大肘发布了新的文献求助10
7秒前
7秒前
烂漫的惜萱完成签到 ,获得积分10
7秒前
yuaasusanaann发布了新的文献求助10
8秒前
8秒前
8秒前
领导范儿应助hu采纳,获得10
10秒前
10秒前
Lowe发布了新的文献求助10
11秒前
安的沛白发布了新的文献求助10
12秒前
看文献了发布了新的文献求助10
12秒前
科研通AI6.4应助体贴半仙采纳,获得10
12秒前
12秒前
13秒前
13秒前
WLL完成签到,获得积分10
13秒前
科研通AI6.2应助bear采纳,获得10
14秒前
Nole应助arrebol采纳,获得30
15秒前
15秒前
xing_xing应助腼腆的修杰采纳,获得20
16秒前
WLL发布了新的文献求助10
17秒前
17秒前
17秒前
18秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
HYDROLYSE ACIDE DE QUELQUES DIOXASPIROCYCLANES 1314
Navigating Normative Orders. Interdisciplinary Perspectives 800
Essentials of Carbohydrate Chemistry and Biochemistry, 4th Edition 700
1 Peter and Christ's Descent to the Dead in Its Early Christian Reception 700
Organizational Behavior 510
Management and the Arts 510
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7743964
求助须知:如何正确求助?哪些是违规求助? 9292081
关于积分的说明 20210528
捐赠科研通 7322678
什么是DOI,文献DOI怎么找? 3307514
关于科研通互助平台的介绍 2459336
邀请新用户注册赠送积分活动 2318312