强化学习
计算机科学
水准点(测量)
测试套件
增强学习
人工智能
机制(生物学)
反向
算法
一套
元启发式
机器学习
优化算法
功能(生物学)
函数优化
数学优化
测试用例
数学
地理
认识论
考古
回归分析
哲学
几何学
历史
生物
进化生物学
大地测量学
遗传算法
作者
Fuqing Zhao,Qiaoyun Wang,Ling Wang
标识
DOI:10.1016/j.knosys.2023.110368
摘要
A reward function is learned from the expert examples by inverse reinforcement learning (IRL), which is more reliable than an artificial method. The moth–flame optimization algorithm (MFO), which is based on the navigation mechanism of a moth flying at night, has been extensively employed to address the complex optimization problem. An inverse reinforcement learning framework with the Q-learning mechanism (IRLMFO) is designed to strengthen the performance of the MFO algorithm in a large-scale real-parameter optimization problem. The right strategy is chosen by the Q-learning mechanism, using historical data provided by the relevant approach in the strategy pool, which stores strategies that include diverse functions. The competition mechanism is designed to strengthen the exploitation capability of the IRLMFO algorithm. The performance of the IRLMFO is verified on the benchmark test suite in CEC 2017. Experimental results illustrate that the IRLMFO outperforms state-of-the-art algorithms.
科研通智能强力驱动
Strongly Powered by AbleSci AI