清晨好,您是今天最早来到科研通的研友!由于当前在线用户较少,发布求助请尽量完整地填写文献信息,科研通机器人24小时在线,伴您科研之路漫漫前行!

A Semiparametric Inverse Reinforcement Learning Approach to Characterize Decision Making for Mental Disorders

灵敏度(控制系统) 概率逻辑 统计推断 参数统计 推论 估计员 心理学 强化学习 人工智能 机器学习 计算机科学 计量经济学 数学 统计 电子工程 工程类
作者
Xingche Guo,Donglin Zeng,Yuanjia Wang
出处
期刊: 卷期号:119 (545): 27-38 被引量:1
标识
DOI:10.1080/01621459.2023.2261184
摘要

Major depressive disorder (MDD) is one of the leading causes of disability-adjusted life years. Emerging evidence indicates the presence of reward processing abnormalities in MDD. An important scientific question is whether the abnormalities are due to reduced sensitivity to received rewards or reduced learning ability. Motivated by the probabilistic reward task (PRT) experiment in the EMBARC study, we propose a semiparametric inverse reinforcement learning (RL) approach to characterize the reward-based decision-making of MDD patients. The model assumes that a subject’s decision-making process is updated based on a reward prediction error weighted by the subject-specific learning rate. To account for the fact that one favors a decision leading to a potentially high reward, but this decision process is not necessarily linear, we model reward sensitivity with a nondecreasing and nonlinear function. For inference, we estimate the latter via approximation by I-splines and then maximize the joint conditional log-likelihood. We show that the resulting estimators are consistent and asymptotically normal. Through extensive simulation studies, we demonstrate that under different reward-generating distributions, the semiparametric inverse RL outperforms the parametric inverse RL. We apply the proposed method to EMBARC and find that MDD and control groups have similar learning rates but different reward sensitivity functions. There is strong statistical evidence that reward sensitivity functions have nonlinear forms. Using additional brain imaging data in the same study, we find that both reward sensitivity and learning rate are associated with brain activities in the negative affect circuitry under an emotional conflict task. Supplementary materials for this article are available online.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
滕皓轩完成签到 ,获得积分20
12秒前
酷波er应助科研通管家采纳,获得10
18秒前
25秒前
Thunnus001完成签到 ,获得积分10
49秒前
vitamin完成签到 ,获得积分0
51秒前
lq完成签到 ,获得积分10
56秒前
mufcyang完成签到,获得积分10
1分钟前
曾经不言完成签到 ,获得积分10
1分钟前
青竹完成签到,获得积分10
1分钟前
1分钟前
naczx完成签到,获得积分0
1分钟前
耕牛热完成签到,获得积分10
1分钟前
顺心的寻双完成签到 ,获得积分10
1分钟前
Artin完成签到,获得积分10
2分钟前
2分钟前
寒冷的月亮完成签到 ,获得积分10
2分钟前
我是你爹完成签到,获得积分10
2分钟前
2分钟前
小熊饼干完成签到,获得积分10
2分钟前
kyokyoro完成签到,获得积分10
2分钟前
2012csc完成签到 ,获得积分0
3分钟前
酷波er应助海绵baby采纳,获得10
3分钟前
心随以动完成签到 ,获得积分10
3分钟前
qin完成签到 ,获得积分10
3分钟前
鸡鸡大魔王完成签到,获得积分10
3分钟前
woxinyouyou完成签到,获得积分0
3分钟前
razz1618完成签到 ,获得积分10
3分钟前
3分钟前
huluwa完成签到,获得积分10
3分钟前
dnpl完成签到,获得积分10
3分钟前
Gary完成签到 ,获得积分10
3分钟前
清平道人完成签到,获得积分0
3分钟前
超男完成签到 ,获得积分10
3分钟前
SUNNYONE完成签到 ,获得积分10
3分钟前
ze完成签到 ,获得积分10
4分钟前
香蕉觅云应助阿泽采纳,获得10
4分钟前
tlh完成签到 ,获得积分10
4分钟前
4分钟前
alee完成签到,获得积分10
4分钟前
无心的亦玉完成签到,获得积分10
4分钟前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
内視鏡的に摘除しえた十二指腸乳頭部腫瘍の2例 660
Cognitive Psychology in a Changing World 600
On nonlinear stability of contact discontinuities. In: Hyperbolic problems: theory, numerics, applications (Stony Brook, NY, 1994) 510
Management and the Arts 510
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
微电子器件实验教程 400
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7681476
求助须知:如何正确求助?哪些是违规求助? 9245554
关于积分的说明 19935299
捐赠科研通 7251971
什么是DOI,文献DOI怎么找? 3287851
关于科研通互助平台的介绍 2445583
邀请新用户注册赠送积分活动 2291449