计算机科学
约束(计算机辅助设计)
订单(交换)
预算约束
激励
约束规划
期限(时间)
公制(单位)
运筹学
人工智能
数学优化
经济
数学
微观经济学
运营管理
财务
物理
几何学
量子力学
随机规划
作者
Jiahui Sun,Haiming Jin,Zhaoxing Yang,Lü Su
标识
DOI:10.1109/tkde.2023.3348491
摘要
Ride-hailing platforms (e.g., Uber and Didi Chuxing) have become increasingly popular in recent years. Efficiency has always been an important metric for such platforms. However, only focusing on efficiency inevitably ignores the fairness of driver incomes, which could impair the sustainability of ride-hailing systems. To optimize such two essential objectives, order dispatching and driver repositioning play an important role, as they impact not only the immediate, but also the future order-serving outcomes of drivers. In practice, the platform offers monetary incentives to drivers for completing the repositioning and has a budget for the repositioning cost. Therefore, in this paper, we aim to exploit joint order dispatching and driver repositioning to optimize both long-term efficiency and fairness in ride-hailing under the budget constraint. To this end, we propose JDRCL, a novel multi-agent reinforcement learning framework, which integrates a group-based action representation that copes with the variable action space, and a primal-dual iterative training algorithm to learn a constraint-satisfying policy that maximizes both the worst and the overall incomes of drivers. Furthermore, we prove the asymptotic convergence rate of our training algorithm. Extensive experiments based on three real-world ride-hailing order datasets show that JDRCL outperforms state-of-the-art baselines on both efficiency and fairness.
科研通智能强力驱动
Strongly Powered by AbleSci AI