反推
万向节
控制理论(社会学)
PID控制器
控制工程
计算机科学
控制器(灌溉)
补偿(心理学)
非线性系统
过程(计算)
跟踪(教育)
自适应控制
控制系统
强化学习
工程类
人工智能
控制(管理)
跟踪误差
自适应系统
作者
Kang Wang,Peng Si,Jie Song,Ke Zhang,Zhongxin Li,Zhilin Wu
摘要
ABSTRACT To improve the tracking and aiming performance of the aerial two‐DOF gimbal system, this paper proposes a backstepping‐based control strategy compensated by the Deep Deterministic Policy Gradient (DDPG) algorithm to address the system's strong nonlinearity and susceptibility to external disturbances. The backstepping control (BSC) method constructs a hierarchical control law to ensure system asymptotic stability, whereas the DDPG algorithm compensates for modeling errors and external disturbances via a deep reinforcement learning process based on policy optimization. A mathematical model of the aerial two‐DOF gimbal system is first established. Two DDPG‐compensated backstepping controllers are designed: a centralized single‐agent compensation structure (BSC‐DDPG‐C) and a decentralized multi‐agent compensation structure (BSC‐DDPG‐D). Simulation experiments were conducted to analyze the training behaviors of the two BSC‐DDPG controllers and reinforcement‐learning‐based PID (RL‐PID) controller, followed by performance comparisons with three conventional controllers: a PID controller, a backstepping controller, and an adaptive backstepping controller. The results show that both BSC‐DDPG controllers significantly outperform the traditional ones, while the RL‐PID controller provides moderate improvement but remains inferior to the DDPG‐compensated methods. Among all, BSC‐DDPG‐D achieves the best yaw and pitch tracking accuracy and stability, validating the effectiveness and superiority of the proposed approach.
科研通智能强力驱动
Strongly Powered by AbleSci AI