Ensemble Experiments to Optimize Interventions Along the Customer Journey: A Reinforcement Learning Approach

心理干预 强化学习 计算机科学 随机对照试验 随机试验 互补性(分子生物学) 机器学习 贝叶斯概率 人工智能 心理学 数学 医学 统计 外科 精神科 生物 遗传学
作者
Yicheng Song,Tianshu Sun
出处
期刊:Management Science [Institute for Operations Research and the Management Sciences]
卷期号:70 (8): 5115-5130 被引量:12
标识
DOI:10.1287/mnsc.2023.4914
摘要

Firms adopt randomized experiments to evaluate various interventions (e.g., website design, creative content, and pricing). However, most randomized experiments are designed to identify the impact of one specific intervention. The literature on randomized experiments lacks a holistic approach to optimize a sequence of interventions along the customer journey. Specifically, locally optimal interventions unveiled by randomized experiments might be globally suboptimal when considering their interdependence as well as the long-term rewards. Fortunately, the accumulation of a large number of historical experiments creates exogenous interventions at different stages along the customer journey and provides a new opportunity. This study integrates multiple experiments within the reinforcement learning (RL) framework to tackle the questions that cannot be answered by stand-alone randomized experiments. How can we learn optimal policy with a sequence of interventions along the customer journey based on an ensemble of historical experiments? Additionally, how can we learn from multiple historical experiments to guide future intervention trials? We propose a Bayesian recurrent Q-network model that leverages the exogenous interventions from multiple experiments to learn their effectiveness at different stages of the customer journey and optimize them for long-term rewards. Beyond optimization within the existing interventions, the Bayesian model also estimates the distribution of rewards, which can guide subject allocation in the design of future experiments to optimally balance exploration and exploitation. In summary, the proposed model creates a two-way complementarity between RL and randomized experiments, and thus, it provides a holistic approach to learning and optimizing interventions along the customer journey. This paper was accepted by Anindya Ghose, information systems. Funding: This work was supported by Adobe Faculty Research Award and the Marketing Science Institute Research Grant. Supplemental Material: The data files and online appendix are available at https://doi.org/10.1287/mnsc.2023.4914 .
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
刚刚
1秒前
1秒前
1秒前
宋宋完成签到 ,获得积分10
1秒前
洛城l发布了新的文献求助10
2秒前
2秒前
2秒前
2秒前
张大忽悠发布了新的文献求助10
2秒前
刘欢完成签到,获得积分10
3秒前
Ava的应助被Musialucky采纳,获得10
3秒前
3秒前
共享精神的应助被北越城主采纳,获得10
3秒前
顾矜的应助被XYZONE采纳,获得10
3秒前
田様的应助被山复尔尔采纳,获得10
3秒前
FashionBoy的应助被Zhang采纳,获得10
3秒前
Hello的应助被幸福的背包采纳,获得10
3秒前
ming2026的应助被轻风叶爽采纳,获得20
4秒前
4秒前
小玉发布了新的文献求助10
4秒前
媛媛老公发布了新的文献求助10
4秒前
4秒前
xiaofeng5838发布了新的文献求助10
5秒前
田様的应助被叶95采纳,获得10
5秒前
feihua1完成签到 ,获得积分10
5秒前
5秒前
Crystal发布了新的文献求助10
5秒前
ememem发布了新的文献求助20
5秒前
5秒前
6秒前
CipherSage的应助被moya采纳,获得10
6秒前
xzcx发布了新的文献求助10
6秒前
大意的谷波完成签到,获得积分10
6秒前
等等发布了新的文献求助10
6秒前
叶落发布了新的文献求助10
6秒前
6秒前
流彩完成签到,获得积分10
6秒前
我能私信骂你吗的应助被DrM采纳,获得10
6秒前
陈陌与发布了新的文献求助30
6秒前
高分求助中
(应助此贴封号)通过应助OA文献获取积分 10000
Organizational Behavior 510
A Silent Apostrophe:The Fayum Portraits 350
Sing with Understanding: Introduction to Theology in Christian Congregational Song, 3rd ed 330
Fractal analysis evaluation of regenerated bone in grafted and graftless maxillary sinus elevation procedures 300
Protection enhancement strategies of potential outbreaks during Hajj 300
Management of a religious mass gathering in North India: Parkash Utsav 550 300
热门求助领域 (近24小时)
化学 材料科学 医学 生物 计算机科学 工程类 纳米技术 有机化学 化学工程 内科学 物理 生物化学 复合材料 催化作用 细胞生物学 人工智能 心理学 无机化学 基因 遗传学
热门帖子
关注 科研通微信公众号,转发送积分 7841929
求助须知:如何正确求助?哪些是违规求助? 9363352
关于积分的说明 20632333
捐赠科研通 7436918
什么是DOI,文献DOI怎么找? 3340107
关于科研通互助平台的介绍 2484568
邀请新用户注册赠送积分活动 2362152