清晨好,您是今天最早来到科研通的研友!由于当前在线用户较少,发布求助请尽量完整地填写文献信息,科研通机器人24小时在线,伴您科研之路漫漫前行!

Multi-Agent Mix Hierarchical Deep Reinforcement Learning for Large-Scale Fleet Management

强化学习 计算机科学 比例(比率) 人工智能 运筹学 工程类 地理 地图学
作者
Xiaohui Huang,Jiahao Ling,Xiaofei Yang,Xiong Zhang,Kaiming Yang
出处
期刊:IEEE Transactions on Intelligent Transportation Systems [Institute of Electrical and Electronics Engineers]
卷期号:24 (12): 14294-14305 被引量:17
标识
DOI:10.1109/tits.2023.3302014
摘要

In recent years, ride-sharing has gained popularity as a daily means of transportation. The primary challenge for large-scale online ride-sharing platforms is to design an efficient fleet management policy that reallocates vehicles to appropriate regions to receive orders, thereby improving the platform’s cumulative revenue and order response rate. Combinatorial optimization algorithms and reinforcement learning methods are commonly employed for this task, but they typically learn a unified repositioning policy for all regions. However, different regions, such as hot and cold zones, may require different repositioning policies due to varying travel patterns. In this paper, we propose a multi-agent mixed hierarchical reinforcement learning approach, called MIX-H, for efficient large-scale fleet management by formulating it as a Markov decision process. MIX-H adopts multi-level controllers, including a leader controller and follower controller, for multi-level action learning. The leader controller plans the goal to be executed by the follower controller. Additionally, to improve the algorithm’s stability, we introduce a MIX module to compute the total value of joint action. Finally, experiments on real-world datasets demonstrate that the proposed method outperforms the state-of-the-art methods.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
小松挂六万八完成签到,获得积分10
7秒前
温柔山槐完成签到 ,获得积分10
8秒前
hj完成签到 ,获得积分10
10秒前
坚强的文博完成签到 ,获得积分10
14秒前
久晓完成签到 ,获得积分10
18秒前
Gloria完成签到 ,获得积分10
21秒前
安菲完成签到 ,获得积分10
21秒前
欣慰怀梦完成签到,获得积分10
22秒前
my完成签到 ,获得积分10
22秒前
上岸应助十三采纳,获得188
29秒前
上岸应助十三采纳,获得177
29秒前
晨雾锁阳完成签到 ,获得积分10
30秒前
40秒前
蔡从安发布了新的文献求助10
43秒前
王志新完成签到 ,获得积分10
52秒前
眯眯眼的安雁完成签到 ,获得积分10
52秒前
大胆的夜白完成签到,获得积分10
56秒前
一二完成签到 ,获得积分10
1分钟前
十三发布了新的文献求助177
1分钟前
田小甜完成签到 ,获得积分10
1分钟前
helen李完成签到 ,获得积分10
1分钟前
不想起床完成签到 ,获得积分10
1分钟前
执意完成签到 ,获得积分10
1分钟前
甜美千山完成签到 ,获得积分10
1分钟前
自然亦凝完成签到,获得积分10
1分钟前
热心的氯化钾完成签到 ,获得积分10
1分钟前
1分钟前
汉堡包应助科研通管家采纳,获得10
1分钟前
1分钟前
1分钟前
琳llin完成签到 ,获得积分10
1分钟前
义气柜子完成签到 ,获得积分10
1分钟前
高大星月完成签到,获得积分10
1分钟前
keyanxinshou完成签到 ,获得积分10
1分钟前
zwd完成签到 ,获得积分10
2分钟前
笨笨的乘风完成签到 ,获得积分10
2分钟前
修仙中应助lly2025采纳,获得10
2分钟前
2分钟前
mochalv123完成签到 ,获得积分10
2分钟前
欢喜完成签到 ,获得积分10
2分钟前
高分求助中
Markov Chain Monte Carlo 10000
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Common Foundations of American and East Asian Modernisation: From Alexander Hamilton to Junichero Koizumi 5000
Matrix Methods in Data Mining and Pattern Recognition Second Edition 610
政治传播过程中的外交与说服——以中苏友好协会为例的历史考察 566
Discerning Saints: Moralization of Intrinsic Motivation and Selective Prosociality at Work 500
Handbuch Trainingswissenschaft – Trainingslehre 500
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7579227
求助须知:如何正确求助?哪些是违规求助? 9158766
关于积分的说明 19592952
捐赠科研通 7162115
什么是DOI,文献DOI怎么找? 3265649
关于科研通互助平台的介绍 2430663
邀请新用户注册赠送积分活动 2256443