Complex Network Optimization for Fixed-Time Continuous Action Iteration Dilemma by Using Reinforcement Learning

强化学习 数学优化 趋同(经济学) 计算机科学 随机博弈 理论(学习稳定性) 纳什均衡 人工智能 数学 机器学习 数理经济学 经济增长 经济
作者
Zhanxiao Jia,Dengxiu Yu,Zhen Wang,C. L. Philip Chen,Xuelong Li
出处
期刊:IEEE Transactions on Network Science and Engineering [Institute of Electrical and Electronics Engineers]
卷期号:11 (4): 3771-3781 被引量:8
标识
DOI:10.1109/tnse.2024.3384509
摘要

In this paper, an optimization algorithm based on deep reinforcement learning is proposed to optimize complex networks in fixed-time convergence of continuous action iteration dilemmas. The field of continuous action iterative dilemmas has long been studied, with prior research primarily emphasizing the effectiveness of strategy selection and the stability of strategy evolution. However, the impact of topology on strategy evolution has remained under-explored. The present study fills this gap by examining how the structure of complex networks influences the time required for players to reach Nash Equilibrium and overall payoff. To identify the optimal complex network that ensures fixed-time convergence of continuous action iteration dilemma, achieves the shortest time, and attains the highest overall payoff in the Nash Equilibrium state, a deep reinforcement learning algorithm is designed to optimize the complex network. Firstly, the paper applies the Lyapunov stability theory to analyze the convergence of the fixed-time continuous action iteration dilemma and compute the upper bound of convergence time. Secondly, based on the fixed-time convergence of continuous action iteration dilemma, we establish evaluation criteria based on the time taken by players to reach the Nash Equilibrium and the overall payoff, subsequently designing evaluation functions for complex networks utilizing these criteria. Thirdly, this paper applies a deep reinforcement learning algorithm to resolve the optimization issue associated with the proposed evaluation function, while analyzing the convergence of complex network optimization methods. Lastly, the effectiveness of the proposed method is verified by simulating the dynamic model of snowdrift games and prisoner dilemmas.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
Fossil@1017完成签到,获得积分10
1秒前
黎云完成签到,获得积分10
1秒前
XiaoMaomi完成签到,获得积分10
2秒前
孙萌萌发布了新的文献求助10
2秒前
2秒前
果汁豆浆完成签到 ,获得积分10
2秒前
乐子真帅发布了新的文献求助10
3秒前
怕孤单的寒天完成签到,获得积分10
3秒前
3秒前
贾方硕完成签到,获得积分10
3秒前
核桃发布了新的文献求助10
3秒前
以七发布了新的文献求助10
3秒前
bkagyin应助扒鸭屁股是块宝采纳,获得10
3秒前
3秒前
keaisile12138完成签到,获得积分10
3秒前
共享精神应助han采纳,获得10
3秒前
poohpooh发布了新的文献求助10
4秒前
dq发布了新的文献求助10
4秒前
4秒前
Rowlet完成签到,获得积分10
5秒前
5秒前
勤快的树懒完成签到,获得积分10
5秒前
6秒前
麦麦脆汁鸡完成签到,获得积分10
6秒前
6秒前
Vincent发布了新的文献求助10
7秒前
7秒前
YUE完成签到,获得积分10
7秒前
aaa发布了新的文献求助10
8秒前
8秒前
shijiu发布了新的文献求助10
8秒前
123完成签到,获得积分10
8秒前
9秒前
9秒前
9秒前
10秒前
10秒前
小姜发布了新的文献求助10
10秒前
猫猫生威发布了新的文献求助10
10秒前
guofuyan完成签到 ,获得积分10
10秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Essentials of Carbohydrate Chemistry and Biochemistry, 4th Edition 800
Navigating Normative Orders. Interdisciplinary Perspectives 800
Organizational Behavior 510
Management and the Arts 510
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
CLSI VET01S-2024 Performance Standards for Antimicrobial Disk and Dilution Susceptibility Tests for Bacteria Isolated From Animals (7th Ed) 500
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 计算机科学 化学工程 工程类 有机化学 物理 复合材料 生物化学 内科学 细胞生物学 基因 遗传学 免疫学 冶金 光电子学 癌症研究
热门帖子
关注 科研通微信公众号,转发送积分 7763335
求助须知:如何正确求助?哪些是违规求助? 9307868
关于积分的说明 20302760
捐赠科研通 7348209
什么是DOI,文献DOI怎么找? 3313962
关于科研通互助平台的介绍 2463728
邀请新用户注册赠送积分活动 2328148