Lifelong Safe Optimal Adaptive Tracking Control of Nonlinear Strict‐Feedback Discrete‐Time Systems

计算机科学 反推 汉密尔顿-雅各比-贝尔曼方程 控制理论(社会学) 人工神经网络 遗忘 数学优化 贝尔曼方程 非线性系统 集合(抽象数据类型) 最优控制 趋同(经济学) 自适应控制 控制(管理) 数学 人工智能 物理 哲学 量子力学 经济 经济增长 语言学 程序设计语言
作者
Behzad Farzanegan,Suresh Jagannathan
出处
期刊:International Journal of Adaptive Control and Signal Processing [Wiley]
卷期号:39 (3): 451-470 被引量:3
标识
DOI:10.1002/acs.3950
摘要

ABSTRACT This paper presents a comprehensive approach for achieving multi‐task safe optimal adaptive tracking (MSOAT) for a class of nonlinear discrete‐time systems, particularly those in strict‐feedback form, utilizing a multi‐layer neural network (MNN)‐based framework. To begin, a cost function with a novel Barrier function (BF) term is introduced for each subsystem to address the weak safely reachable problem, serving as a crucial tool for guiding the system's trajectory toward the safe set while avoiding unwanted sets. To deal with the tracking problem, the Hamilton‐Jacobi‐Bellman (HJB) framework is used through the actor‐critic MNN‐based backstepping technique to estimate the solution of the value functions and obtain both virtual and actual optimal control policies for each subsystem, effectively circumventing non‐causality issues. Further, to mitigate catastrophic forgetting in multi‐tasking scenarios, a regularizer term, which is derived from the online version of the Elastic Weight Consolidation (EWC) method, is included in the critic and actor MNN update laws without directly computing the Fisher information matrix. To enhance the convergence rate, the critic MNN is tuned with a hybrid learning technique involving weight adjustments both at specific sampling instants and iteratively within those intervals. A control barrier function (CBF) with a time‐varying BF is also integrated into the actor update law, collaborating with the BF to keep the trajectory in the safe set with a smaller trade‐off factor, simultaneously validating the safety condition in real‐time. Finally, the overall stability is established. An example of a 6‐DOF autonomous underwater vehicle (AUV) is used to assess the effectiveness of the proposed approach.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
1秒前
1秒前
ASH应助林子采纳,获得10
5秒前
5秒前
乐乐应助吱吱采纳,获得10
6秒前
木羽发布了新的文献求助10
6秒前
殿下发布了新的文献求助10
7秒前
洁净代容发布了新的文献求助10
7秒前
8秒前
9秒前
恩恩爸学科研完成签到,获得积分10
10秒前
10秒前
shirley发布了新的文献求助10
10秒前
aajhajkahna应助SUK采纳,获得10
10秒前
上官若男应助默流采纳,获得10
11秒前
可爱的函函应助吱吱采纳,获得10
12秒前
yin完成签到,获得积分20
13秒前
14秒前
14秒前
15秒前
刘骁萱发布了新的文献求助10
15秒前
白衣修身发布了新的文献求助10
17秒前
Asura完成签到,获得积分10
18秒前
TOP完成签到,获得积分10
19秒前
久处完成签到,获得积分10
19秒前
yin发布了新的文献求助10
19秒前
吱吱发布了新的文献求助10
19秒前
xsx1111完成签到,获得积分10
20秒前
L77完成签到,获得积分0
20秒前
21秒前
21秒前
叶武林完成签到,获得积分10
21秒前
21秒前
24秒前
bigstone完成签到,获得积分10
24秒前
24秒前
25秒前
liuzengzhang666发布了新的文献求助100
26秒前
orixero应助小鱼要变咸采纳,获得10
26秒前
27秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Health Psychology 800
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
Electric machines: theory, operating applications, and controls 500
The Analytical and Numerical Solution of Electric and Magnetic Fields 500
When Is Two-Stage Sample Robust Optimization Asymptotically Optimal? 500
Discerning Saints: Moralization of Intrinsic Motivation and Selective Prosociality at Work 500
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7593551
求助须知:如何正确求助?哪些是违规求助? 9170720
关于积分的说明 19629554
捐赠科研通 7171351
什么是DOI,文献DOI怎么找? 3267626
关于科研通互助平台的介绍 2432450
邀请新用户注册赠送积分活动 2260268