Multi-Agent Connected Autonomous Driving using Deep Reinforcement Learning

强化学习计算机科学可扩展性自主代理人交叉口（航空）领域（数学分析）马尔可夫决策过程人工智能集合（抽象数据类型）部分可观测马尔可夫决策过程过程（计算）多智能体系统分布式计算马尔可夫链人机交互马尔可夫过程机器学习马尔可夫模型工程类统计程序设计语言数据库操作系统航空航天工程数学分析数学

作者

Praveen Palanisamy

链接

arxiv.org arxiv.org arxiv.orgdoi.org

标识

DOI：10.1109/ijcnn48605.2020.9207663

摘要

The capability to learn and adapt to changes in the driving environment is crucial for developing autonomous driving systems that are scalable beyond geo-fenced operational design domains. Deep Reinforcement Learning (RL) provides a promising and scalable framework for developing adaptive learning based solutions. Deep RL methods usually model the problem as a (Partially Observable) Markov Decision Process in which an agent acts in a stationary environment to learn an optimal behavior policy. However, driving involves complex interaction between multiple, intelligent (artificial or human) agents in a highly non-stationary environment. In this paper, we propose the use of Partially Observable Markov Games(POSG) for formulating the connected autonomous driving problems with realistic assumptions. We provide a taxonomy of multi-agent learning environments based on the nature of tasks, nature of agents and the nature of the environment to help in categorizing various autonomous driving problems that can be addressed under the proposed formulation. As our main contributions, we provide MACAD-Gym, a Multi-Agent Connected, Autonomous Driving agent learning platform for furthering research in this direction. Our MACAD-Gym platform provides an extensible set of Connected Autonomous Driving (CAD) simulation environments that enable the research and development of Deep RL- based integrated sensing, perception, planning and control algorithms for CAD systems with unlimited operational design domain under realistic, multi-agent settings. We also share the MACAD-Agents that were trained successfully using the MACAD-Gym platform to learn control policies for multiple vehicle agents in a partially observable, stop-sign controlled, 3-way urban intersection environment with raw (camera) sensor observations.

求助该文献

最长约 10秒，即可获得该文献文件

Multi-Agent Connected Autonomous Driving using Deep Reinforcement Learning

今日热心研友