计算机科学
随机博弈
排队论
钥匙(锁)
数学优化
排队
服务(商务)
资源配置
控制(管理)
分布式计算
运筹学
人工智能
计算机网络
计算机安全
数学
工程类
经济
数理经济学
经济
作者
Wei-Kang Hsu,Jiaming Xu,Xiaojun Lin,Mark R. Bell
出处
期刊:Operations Research
[Institute for Operations Research and the Management Sciences]
日期:2021-03-09
卷期号:70 (2): 1166-1181
被引量:16
标识
DOI:10.1287/opre.2021.2100
摘要
Many online service platforms have dedicated algorithms to match their available resources to incoming clients to maximize client satisfaction. One of the key challenges is to balance the generation of higher payoffs from existing clients and exploration of new clients’ unknown characteristics while at the same time satisfy the resource capacity constraints. In “Integrated Online Learning and Adaptive Control in Queueing Systems with Uncertain Payoffs,” Hsu, Xu, Lin, and Bell show that traditional approaches such as maximizing instantaneous payoffs with current knowledge or using queue-length based controls guided by “shadow prices,” would lead to suboptimal long-term payoffs. Instead, they propose a novel utility-guided assignment algorithm that seamlessly integrates online learning and adaptive control to provide high system payoffs with performance guarantees. The theoretical performance bound also lends system design insights into the impact of uncertain client dynamics, payoff learning, and backlogged clients. They further develop a decentralized version of the algorithm, which is applicable to large systems and performs well even when the service rates are random.
科研通智能强力驱动
Strongly Powered by AbleSci AI