Opportunities for reinforcement learning in stochastic dynamic vehicle routing

强化学习 计算机科学 布线(电子设计自动化) 钢筋 数学优化 车辆路径问题 运筹学 人工智能 数学 计算机网络 材料科学 复合材料
作者
Florentin D. Hildebrandt,Barrett W. Thomas,Marlin W. Ulmer
出处
期刊:Computers & Operations Research [Elsevier BV]
卷期号:150: 106071-106071 被引量:101
标识
DOI:10.1016/j.cor.2022.106071
摘要

There has been a paradigm-shift in urban logistic services in the last years; demand for real-time, instant mobility and delivery services grows. This poses new challenges to logistic service providers as the underlying stochastic dynamic vehicle routing problems (SDVRPs) require anticipatory real-time routing actions. The complexity of finding efficient routing actions is multiplied by the challenge of evaluating such actions with respect to their effectiveness given future dynamism and uncertainty. Reinforcement learning (RL) is a promising tool for evaluating actions but it is not designed for searching the complex and combinatorial action space. Thus, past work on RL for SDVRP has either restricted the action space, that is solving only subproblems by RL and everything else by established heuristics, or focused on problems that reduce to resource allocation problems. For solving real-world SDVRPs, new strategies are required that address the combined challenge of combinatorial, constrained action space and future uncertainty, but as our findings suggest, such strategies are essentially non-existing. Our survey paper shows that past work relied either on action-space restriction or avoided routing actions entirely and highlights opportunities for more holistic solutions. • We discuss the challenges and opportunities for reinforcement learning in stochastic dynamic vehicle routing. • We carefully review and classify the existing literature. • We use examples and pseudocode for illustration throughout the papers. • We present two specific and promising means to improve future work in this domain.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
4秒前
wlq完成签到,获得积分10
5秒前
Giaodv发布了新的文献求助10
7秒前
molihuakai应助着急的听枫采纳,获得10
7秒前
7秒前
8秒前
li完成签到,获得积分10
8秒前
10秒前
姜汁树完成签到 ,获得积分10
10秒前
独特的曼柔完成签到,获得积分20
11秒前
li发布了新的文献求助10
11秒前
12秒前
落寞的楼房完成签到,获得积分10
12秒前
AW139发布了新的文献求助10
13秒前
情书完成签到 ,获得积分10
13秒前
晨gegeai发布了新的文献求助10
14秒前
地精术士发布了新的文献求助10
15秒前
16秒前
流莺发布了新的文献求助10
16秒前
16秒前
小小怪下士完成签到,获得积分10
16秒前
17秒前
wei完成签到,获得积分10
17秒前
17秒前
星星完成签到 ,获得积分10
18秒前
狂屌拽霸威震天完成签到,获得积分10
19秒前
20秒前
zoe发布了新的文献求助10
21秒前
21秒前
你说吧完成签到,获得积分10
22秒前
LIUy发布了新的文献求助10
23秒前
23秒前
23秒前
地精术士发布了新的文献求助10
23秒前
123456hhh完成签到,获得积分10
24秒前
林夕完成签到 ,获得积分10
25秒前
上官若男应助猪八戒采纳,获得10
25秒前
yyydmhsj完成签到 ,获得积分10
25秒前
26秒前
liuxinyu发布了新的文献求助10
26秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Essentials of Carbohydrate Chemistry and Biochemistry, 4th Edition 800
Navigating Normative Orders. Interdisciplinary Perspectives 800
A Psychological Understanding of Criticism and Mental Health 600
Organizational Behavior 510
Management and the Arts 510
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7752964
求助须知:如何正确求助?哪些是违规求助? 9299852
关于积分的说明 20254660
捐赠科研通 7335075
什么是DOI,文献DOI怎么找? 3310386
关于科研通互助平台的介绍 2461703
邀请新用户注册赠送积分活动 2323296