心理学
结果(博弈论)
背景(考古学)
认知心理学
P3b页
事件相关电位
价(化学)
脑电图
大脑活动与冥想
神经活动
任务(项目管理)
动力学(音乐)
预测(人工智能)
发展心理学
钢筋
社会心理学
神经系统
信息处理
认知
强化学习
意识的神经相关物
唤醒
作者
Matthew D. Bachman,Kaya Scheman,René San Martín,Scott A. Huettel,Marty G. Woldorff
摘要
Reward expectations are fundamental to theories of decision making and reinforcement learning. While prior research has focused on how expectations can influence outcome processing, far fewer studies have investigated how these expectations are actually formed. To address this gap, we measured EEG activity from participants as they completed a two-stage binary-choice task designed to separate the formation of expectations about outcome probabilities from the processing of actual reward outcomes. To more fully examine the neural mechanisms underlying each stage, we measured the Reward Positivity (RewP) and P3b time-domain event-related potential (ERP) components, as well as the delta- and theta-band activity underlying each ERP. During the Outcome Probability stage, participants learned the likelihood of their choice winning on that trial. Each measure of RewP-latency activity (ERP, delta, theta) was larger for outcomes that were certain to occur, but each measure diverged in its relationship to outcome valence. Conversely, all P3b-latency measures were increased for losses that were certain to occur. Notably, changes in RewP-Theta, not in ERP components, provided the earliest marker of sensitivity to certain losses. At the Actual Outcome stage, the RewP and P3b-ERPs were larger for unexpected wins, consistent with theories of reward prediction errors and context updating. Delta activity generally followed the patterns observed in its temporally matched ERP but displayed an inconsistent relationship with outcome valence, suggesting that it may reflect contextualized feedback processing rather than a specific win-related signal. Theta was insensitive to outcome valence and only sensitive to expectations at longer latencies, indicating a shift from valence-sensitive processing during expectation formation to a broader role in monitoring expectancy violations. Together, the results underscore the importance of temporally and functionally distinguishing between the expectation formation and outcome phases, while demonstrating the value of multimethodological analytical approaches to fully capture the dynamic nature of reward-based decision-making.
科研通智能强力驱动
Strongly Powered by AbleSci AI