强化学习
情绪识别
脑电图
计算机科学
感知
人工智能
钢筋
心理学
认知心理学
神经科学
社会心理学
作者
Dongdong Li,Li Xie,Zhe Wang,Hai Yang
标识
DOI:10.1109/tnnls.2023.3265730
摘要
Inspired by the well-known Papez circuit theory and neuroscience knowledge of reinforcement learning, a double dueling deep Q network (DQN) is built incorporating the electroencephalogram (EEG) signals of the frontal lobe as prior information, which is named frontal lobe double dueling DQN (FLD3QN). The framework of FLD3QN is constructed in accord with the brain emotion mechanism which takes the frontal lobe and the thalamus as the core, in which the part of the Papez circuit is simulated by the bifrontal lobe residual convolution neural network (BiFRCNN). Moreover, a step penalty factor is designed to constrain the number of mistakes of the agent. The ablation studies results on the public EEG emotion dataset DEAP verified the important roles of the frontal lobe and the Papez circuit in modeling the procedure of learning rewards during the perception of emotions, with a great increase in the average accuracies by 25.24% and 23.31% in valence and arousal dimensions.
科研通智能强力驱动
Strongly Powered by AbleSci AI