强化学习
马尔可夫决策过程
计算机科学
人工智能
时差学习
领域(数学)
钢筋
人工神经网络
机器学习
工程类
结构工程
作者
Sutton, Richard S.,Barto, Andrew
标识
DOI:10.1109/tnn.2004.842673
摘要
An account of key ideas and algorithms in reinforcement learning. The discussion ranges from the history of the field's intellectual foundations to recent developments and applications. Areas studied include reinforcement learning problems in terms of Markov decision problems and solution methods.
科研通智能强力驱动
Strongly Powered by AbleSci AI