强化学习
控制理论(社会学)
非线性系统
信息物理系统
跟踪(教育)
计算机科学
控制(管理)
控制工程
人工神经网络
人工智能
工程类
心理学
物理
教育学
量子力学
操作系统
作者
Penghao Chen,Hamid Reza Karimi,Xiaoli Luan,Fei Liu
摘要
ABSTRACT In this article, the identifier–critic–actor neural adaptive optimal control issue is addressed for a class of fully nonaffine pure‐feedback nonlinear cyber‐physical systems with input quantization and time‐reference‐dependent output constraints. By constructing a time‐varying asymmetric barrier Lyapunov function and integrating it with the dynamic surface control method and a reinforcement learning algorithm, a controller is designed based on a neural network approximation of the identifier–critic–actuator structure. In this framework, the identifier estimates unknown dynamics, the critic evaluates system performance, and the actor executes the control action. The control scheme involves designing the actual control inputs for all virtual and dynamic surface controls as the optimal solutions of their corresponding subsystems. The updated law is derived by taking the negative gradient of a simple positive function, which is constructed from the partial derivatives of the Hamilton–Jacobi–Bellman equation. In parallel, the proposed quantizer combines the benefits of both hysteretic and uniform quantization. A key aspect of this article is the simultaneous consideration of constraint boundaries related to both the reference signal and time, which adds complexity to the design of the control algorithm. Stability analysis confirms that all signals remain bounded and adhere to the time‐ and reference‐dependent constraints imposed on output.
科研通智能强力驱动
Strongly Powered by AbleSci AI