强化学习
计算机科学
控制系统
比例(比率)
加速度
控制(管理)
时差学习
系统设计
能源消耗
控制工程
最优控制
模拟
人工智能
工程类
数学优化
软件工程
经典力学
电气工程
物理
量子力学
数学
作者
Eddieb Sadat,Mostaan Lotfalian Saremi,Alparslan Emrah Bayrak
标识
DOI:10.1115/detc2023-116567
摘要
Abstract Design of smart (or active) systems that perform automated tasks intelligently based on the interaction with their environments requires a collective solution of the physical and control system design problems together. In this paper, we present a model-free on-policy reinforcement learning approach to solve control co-design problems for such smart systems. This approach uses a discrete two timescale reinforcement learning that addresses the control system design in an inner loop with a fast time scale and the physical system design in an outer loop with a slower time scale. Both design problems use the same temporal difference-based Q-learning formulation. We apply this two-time-scale reinforcement approach to the online video game EcoRacer where the physical system involves the design of a gear ratio for an electric vehicle and the control system involves acceleration and braking decisions over time to finish a track with minimum energy consumption within a limited time. The results show the ability of the proposed approach to find the system optimal solution for the EcoRacer case study within a reasonable computation time without requiring any knowledge of the physics governing the system. The proposed method is generalizable and has the potential to take advantage of the ongoing developments in the field of reinforcement learning.
科研通智能强力驱动
Strongly Powered by AbleSci AI