Statistical Learning Theory for Control: A Finite-Sample Perspective

工具箱 计算机科学 背景(考古学) 透视图(图形) 动力系统理论 国家(计算机科学) 人工智能 控制(管理) 推荐系统 样品(材料) 控制器(灌溉) 理论计算机科学 机器学习 算法 生物 农学 古生物学 物理 化学 色谱法 程序设计语言 量子力学
作者
Anastasios Tsiamis,Ingvar Ziemann,Nikolai Matni,George J. Pappas
出处
期刊:IEEE Control Systems Magazine [Institute of Electrical and Electronics Engineers]
卷期号:43 (6): 67-97 被引量:44
标识
DOI:10.1109/mcs.2023.3310345
摘要

Learning algorithms have become an integral component to modern engineering solutions. Examples range from self-driving cars and recommender systems to finance and even critical infrastructure, many of which are typically under the purview of control theory. While these algorithms have already shown tremendous promise in certain applications [1] , there are considerable challenges, in particular, with respect to guaranteeing safety and gauging fundamental limits of operation. Thus, as we integrate tools from machine learning into our systems, we also require an integrated theoretical understanding of how they operate in the presence of dynamic and system-theoretic phenomena. Over the past few years, intense efforts toward this goal—an integrated theoretical understanding of learning, dynamics, and control—have been made. While much work remains to be done, a relatively clear and complete picture has begun to emerge for (fully observed) linear dynamical systems. These systems already allow for reasoning about concrete failure modes, thus helping to indicate a path forward. Moreover, while simple at a glance, these systems can be challenging to analyze. Recently, a host of methods from learning theory and high-dimensional statistics, not typically in the control-theoretic toolbox, have been introduced to our community. This tutorial survey serves as an introduction to these results for learning in the context of unknown linear dynamical systems (see “Summary”). We review the current state of the art and emphasize which tools are needed to arrive at these results. Our focus is on characterizing the sample efficiency and fundamental limits of learning algorithms. Along the way, we also delineate a number of open problems. More concretely, this article is structured as follows. We begin by revisiting recent advances in the finite-sample analysis of system identification. Next, we discuss how these finite-sample bounds can be used downstream to give guaranteed performance for learning-based offline control. The final technical section discusses the more challenging online control setting. Finally, in light of the material discussed, we outline a number of future directions.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
1秒前
tulip77发布了新的文献求助10
1秒前
1秒前
章鱼完成签到,获得积分10
1秒前
1秒前
1秒前
852应助海盐采纳,获得10
1秒前
无极微光应助高贵振家采纳,获得20
1秒前
2秒前
Steven完成签到,获得积分10
2秒前
贪玩的谷兰完成签到,获得积分10
2秒前
2秒前
王大伟2023发布了新的文献求助10
2秒前
王大伟2023发布了新的文献求助10
2秒前
2秒前
王大伟2023发布了新的文献求助10
2秒前
lan发布了新的文献求助10
2秒前
体贴的冷之完成签到,获得积分20
3秒前
科研通AI6.2应助WATQ采纳,获得10
3秒前
王大伟2023发布了新的文献求助10
3秒前
王大伟2023发布了新的文献求助10
3秒前
木木完成签到,获得积分10
3秒前
4秒前
王大伟2023发布了新的文献求助10
4秒前
干脆面完成签到,获得积分10
4秒前
4秒前
三叶发布了新的文献求助10
4秒前
5秒前
可靠冥幽完成签到,获得积分10
5秒前
5秒前
王大伟2023发布了新的文献求助10
5秒前
张大侠发布了新的文献求助10
5秒前
5秒前
6秒前
NERV完成签到,获得积分10
6秒前
王大伟2023发布了新的文献求助10
6秒前
bobo完成签到,获得积分20
6秒前
王大伟2023发布了新的文献求助10
6秒前
王大伟2023发布了新的文献求助10
6秒前
linger发布了新的文献求助10
7秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
HYDROLYSE ACIDE DE QUELQUES DIOXASPIROCYCLANES 1000
Navigating Normative Orders. Interdisciplinary Perspectives 800
1 Peter and Christ's Descent to the Dead in Its Early Christian Reception 700
Essentials of Carbohydrate Chemistry and Biochemistry, 4th Edition 600
Organizational Behavior 510
Management and the Arts 510
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7741413
求助须知:如何正确求助?哪些是违规求助? 9290023
关于积分的说明 20198572
捐赠科研通 7319745
什么是DOI,文献DOI怎么找? 3306727
关于科研通互助平台的介绍 2458937
邀请新用户注册赠送积分活动 2317134