强化学习
计算机科学
粒子群优化
过程(计算)
趋同(经济学)
多目标优化
帕累托原理
数学优化
人工智能
机器学习
数学
经济
经济增长
操作系统
作者
Ping Zhou,Xuan Wang,Tianyou Chai
标识
DOI:10.1109/tcyb.2022.3164476
摘要
This article proposes a multiobjective operation optimization method based on reinforcement self-learning and knowledge guidance for quality assurance and consumption reduction of wastewater treatment process (WWTP) with nonstationary time-varying dynamics. First, operation optimization models are developed by online sequential random vector functional-link (OS-RVFL) neural network, which can realize online sequential learning of model parameters. Then, a knowledge base is established to store typical optimization cases for knowledge guiding the subsequent optimizations. Based on it, a reinforcement self-learning-based multiobjective particle swarm optimization (RSL-MOPSO) algorithm is proposed to perform optimization calculation. In this algorithm, reinforcement self-learning is used for interaction learning between environment and action in optimization, and the particle motion trend of algorithm is adjusted according to the feedback information of the optimization process. The effects of wastewater state parameters on particles are recorded and reused to improve the solution quality and calculation efficiency of optimization. Moreover, to make good use of the information of the previous optimizations and balance the coordination between global search in the early stage and local search in the later stage, a selective information feedback mechanism is further proposed to ensure the diversity and convergence of the algorithm. Finally, prediction-based intelligent decision making is performed to select the final optimization solution as the final setpoints for the lower-level controllers from the Pareto frontier with considering specific technical requirements. Data experiments show that the proposed method can effectively reduce energy consumption and ensure effluent quality.
科研通智能强力驱动
Strongly Powered by AbleSci AI