清晨好,您是今天最早来到科研通的研友!由于当前在线用户较少,发布求助请尽量完整地填写文献信息,科研通机器人24小时在线,伴您科研之路漫漫前行!

Deep Reinforcement Learning for Online Assortment Customization: A Data-Driven Approach

个性化 强化学习 计算机科学 钢筋 人工智能 万维网 心理学 社会心理学
作者
Tao Li,Chenhao Wang,Yao Wang,Shaojie Tang,Ningyuan Chen
出处
期刊:Production and Operations Management [Wiley]
卷期号:35 (2): 665-684 被引量:1
标识
DOI:10.1177/10591478251351737
摘要

When a platform has limited inventory, it is important to have a variety of products available for each customer while managing the remaining stock. To maximize revenue over the long term, the assortment policy needs to take into account the complex purchasing behavior of customers whose arrival orders and preferences may be unknown. We propose a data-driven approach for dynamic assortment planning that utilizes historical customer arrivals and transaction data. To address the challenge of online assortment customization, we use a Markov decision process framework and employ a model-free deep reinforcement learning (DRL) approach to solve the online assortment policy because of the computational challenge. Our method uses a specially designed deep neural network (DNN) model to create assortments while observing the inventory constraints, and an advantage actor-critic algorithm to update the parameters of the DNN model, with the help of a simulator built from the historical transaction data. To evaluate the effectiveness of our approach, we conduct simulations using both a synthetic data set generated with a pre-determined customer type distribution and ground-truth choice model, as well as a real-world data set. Our extensive experiments demonstrate that our approach produces significantly higher long-term revenue compared to some existing methods and remains robust under various practical conditions. We also demonstrate that our approach can be easily adapted to a more general problem that includes reusable products, where customers might return purchased items. In this setting, we find that our approach performs well under various usage time distributions.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
笑然完成签到,获得积分20
1秒前
2秒前
超帅连虎发布了新的文献求助10
8秒前
修哥完成签到,获得积分20
9秒前
举个栗子8完成签到 ,获得积分10
14秒前
goodsheperd完成签到 ,获得积分10
16秒前
博修完成签到,获得积分20
19秒前
南宫士晋完成签到 ,获得积分0
26秒前
巨型肥猫完成签到 ,获得积分10
36秒前
斯文的炳完成签到 ,获得积分10
40秒前
KKwang完成签到 ,获得积分10
52秒前
cat应助科研通管家采纳,获得10
1分钟前
cat应助科研通管家采纳,获得10
1分钟前
cat应助科研通管家采纳,获得10
1分钟前
冷静的豪完成签到 ,获得积分10
1分钟前
Arctic完成签到 ,获得积分10
1分钟前
Magic完成签到 ,获得积分10
1分钟前
qianci2009完成签到,获得积分10
1分钟前
Lychee完成签到 ,获得积分10
1分钟前
研友_惊鸿发布了新的文献求助10
1分钟前
2分钟前
2分钟前
爱沉淀的太阳花完成签到,获得积分10
2分钟前
woxinyouyou完成签到,获得积分0
2分钟前
Lillianzhu1完成签到,获得积分10
2分钟前
shayeeeeee完成签到 ,获得积分10
2分钟前
qvb完成签到 ,获得积分10
2分钟前
温暖完成签到 ,获得积分10
2分钟前
直率的笑翠完成签到 ,获得积分10
2分钟前
研友_Z1eDgZ发布了新的文献求助10
2分钟前
Jessica英语好完成签到,获得积分10
2分钟前
flysky120完成签到,获得积分10
2分钟前
yanglinhai完成签到 ,获得积分10
2分钟前
菲菲完成签到 ,获得积分10
2分钟前
t铁核桃1985完成签到 ,获得积分0
2分钟前
笨笨完成签到 ,获得积分10
2分钟前
七月不远发布了新的文献求助10
3分钟前
cat应助科研通管家采纳,获得10
3分钟前
cat应助科研通管家采纳,获得10
3分钟前
cat应助科研通管家采纳,获得10
3分钟前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
China Pluperfect I: Epistemology of Past and Outside in Chinese Art 520
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
基于锂离子电池正极材料回收的绿色溶剂开发及工程化应用研究 500
Auslegungsgeschichte 500
Transdermal drug delivery systems market size report 500
Cosmos as Art Object: Studies in Plato's Timaeus and Other Dialogues 500
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7640456
求助须知:如何正确求助?哪些是违规求助? 9213463
关于积分的说明 19763511
捐赠科研通 7206352
什么是DOI,文献DOI怎么找? 3276078
关于科研通互助平台的介绍 2437709
邀请新用户注册赠送积分活动 2273531