强化学习
机器人
计算机科学
粒子群优化
移动机器人
增强学习
群机器人
机器人学习
人工智能
群体行为
过程(计算)
网格
机器学习
数学
几何学
操作系统
作者
Orawan Watchanupaporn,Peerapun Pudtuan
标识
DOI:10.1109/iccar.2016.7486700
摘要
In this paper, a group of mobile robots learns to solve a target reaching problem in a simulated grid environment filled with obstacles. Each robot knows its distance to the target and can communicate with each other. The proposed learning algorithm combines a reinforcement learning algorithm and a swarm optimization algorithm. Q-learning, which is a reinforcement learning algorithm, is modified to learn a policy by specifying rewards and punishment for certain robot actions. Particle swarm optimization (PSO), which is a swarm optimization algorithm, is modified for grid environment and used to accelerate the learning process for multiple robots. The proposed algorithm outperforms the original Q-learning in both training and testing. It learns 2.23 times faster and required 6.52 fewer steps to reach the destination. Moreover, it uses less memory than the original Q-learning. We also experiment on various numbers of robots. The result shows that more robots learn faster in most cases.
科研通智能强力驱动
Strongly Powered by AbleSci AI