强化学习
计算机科学
人工智能
增强学习
分类器(UML)
球(数学)
占有(语言学)
背景(考古学)
任务(项目管理)
机器学习
数学
哲学
古生物学
数学分析
经济
管理
生物
语言学
作者
Saikat Sarkar,Dipti Prasad Mukherjee,Amlan Chakrabarti
出处
期刊:IEEE Transactions on Cognitive and Developmental Systems
[Institute of Electrical and Electronics Engineers]
日期:2022-07-26
卷期号:15 (2): 914-924
被引量:2
标识
DOI:10.1109/tcds.2022.3194103
摘要
We propose a reinforcement learning (RL) based technique to detect passes from the video of a soccer match. The detection of passes determines ball possession statistics of a soccer match. A sequence of video frames is mapped to a sequence of states, such as ball with team A or team B or ball not possessed either by team A or B. The agent of RL learns the frame-to-state mapping and the optimal policy to decide the mapping task. We propose a novel reward function by utilizing contextual information of the soccer game in order to help the agent decide the optimal policy. In this context, the advantage of RL is in the integration of a reward system in choosing an action that maps a video frame of a soccer match to one of three possible states. Unlike competing methods, we design the RL model in a way so that explicit identification of team labels of players is not required. We introduce a deep recurrent $Q$ -network (DRQN) to learn the optimal policy. For efficient training of the DRQN, we have proposed decorrelated experience replay (DER), a strategy that selects important experiences based on the correlations of the experiences stored in the replay memory. Experimental results show that at least 5.75% and 2.1% better accuracy are achieved in calculating pass and possession statistics, respectively, compared to similar approaches.
科研通智能强力驱动
Strongly Powered by AbleSci AI