已入深夜,您辛苦了!由于当前在线用户较少,发布求助请尽量完整地填写文献信息,科研通机器人24小时在线,伴您度过漫漫科研夜!祝你早点完成任务,早点休息,好梦!

Deep Reinforcement Learning for Equilibrium Computation in Multistage Auctions and Contests

计算机科学 序贯平衡 强化学习 均衡选择 斯塔克伯格竞赛 投标 数理经济学 数学优化 纳什均衡 博弈论 共同价值拍卖 离散化 贝叶斯博弈 动作(物理) 解决方案概念 李普希茨连续性 马尔可夫完全平衡 马尔可夫决策过程 计算 实施理论 收入 差速器(机械装置) 重复博弈 信号(编程语言) 最佳反应 随机博弈 放松(心理学) 广泛形式游戏 零和博弈 非线性系统 一般均衡理论
作者
Fabian R. Pieroth,Nils Kohring,Martin Bichler
出处
期刊:Management Science [Institute for Operations Research and the Management Sciences]
标识
DOI:10.1287/mnsc.2024.06771
摘要

We compute equilibrium strategies in multistage games with continuous signal and action spaces as they are widely used in the management sciences and economics. Examples include sequential sales via auctions, multistage elimination contests, and Stackelberg competitions. In sequential auctions, analysts performing equilibrium analysis are required to derive not just single bids but bid functions for all possible signals or values that a bidder might have in multiple stages. Because of the continuity of the signal and action spaces, these bid functions come from an infinite dimensional space. Although such models are fundamental to game theory and its applications, equilibrium strategies are rarely known. The resulting system of nonlinear differential equations is considered intractable for all but elementary models. This has been limiting progress in game theory and is a barrier to its adoption in the field. We show that deep reinforcement learning and self-play can learn equilibrium bidding strategies for various multistage games. Verifying an equilibrium in such games is challenging because of the continuous signal and action spaces. We introduce a verification algorithm and prove that the error of this verifier decreases when considering Lipschitz continuous strategies with increasing levels of discretization and sample sizes. Leveraging the novel verification algorithm, we find equilibrium in models that have not yet been explored analytically and new asymmetric equilibrium bid functions for established models of sequential auctions. This paper was accepted by David Simchi-Levi, revenue management and market analytics. Funding: This work was supported by the Deutsche Forschungsgemeinschaft [Grant BI 1057/9]. Additionally, this project has received funding from the European Research Council under the European Union’s Horizon Europe research and innovation programme [Grant Agreement 101198689]. Supplemental Material: The online appendices and data files are available at https://doi.org/10.1287/mnsc.2024.06771 .
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
海绵宝宝完成签到 ,获得积分10
刚刚
3秒前
zz完成签到 ,获得积分10
4秒前
4秒前
赘婿应助懒羊羊采纳,获得10
4秒前
5秒前
生动曲奇完成签到,获得积分10
5秒前
6秒前
小贤发布了新的文献求助10
6秒前
haru发布了新的文献求助10
9秒前
xzbnvsg完成签到,获得积分10
9秒前
ice完成签到 ,获得积分10
10秒前
thaogaa02发布了新的文献求助10
11秒前
刘阳发布了新的文献求助10
11秒前
highkick发布了新的文献求助10
12秒前
17秒前
共享精神应助坚强莞采纳,获得10
17秒前
17秒前
19秒前
21秒前
21秒前
yzy完成签到 ,获得积分10
21秒前
yimei发布了新的文献求助10
22秒前
23秒前
23秒前
叶小点发布了新的文献求助10
23秒前
zzz完成签到 ,获得积分10
23秒前
li5498693完成签到,获得积分20
24秒前
懒羊羊发布了新的文献求助10
26秒前
26秒前
wodeqiche2007发布了新的文献求助30
28秒前
28秒前
刘长绪完成签到,获得积分20
29秒前
满意的天思完成签到 ,获得积分10
29秒前
li5498693发布了新的文献求助10
29秒前
32秒前
32秒前
乐乐应助叶小点采纳,获得10
35秒前
坚强莞发布了新的文献求助10
35秒前
36秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Effects of Two Weeks of Red Light Therapy on Choroidal Thickness and Axial Length in Young Adults 700
内視鏡的に摘除しえた十二指腸乳頭部腫瘍の2例 660
The Foundation of Positive Psychology 600
Management and the Arts 510
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
The Neuroscience of Language 400
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7676551
求助须知:如何正确求助?哪些是违规求助? 9242609
关于积分的说明 19917748
捐赠科研通 7246841
什么是DOI,文献DOI怎么找? 3286487
关于科研通互助平台的介绍 2444508
邀请新用户注册赠送积分活动 2289468