脑磁图
计算机科学
稳健性(进化)
公制(单位)
人工智能
二元分类
解码方法
机器学习
班级(哲学)
随机森林
性能指标
接收机工作特性
脑电图
大脑活动与冥想
二进制数
灵敏度(控制系统)
模式识别(心理学)
支持向量机
数学
心理学
算法
精神科
工程类
基因
算术
经济
生物化学
化学
管理
电子工程
运营管理
作者
Philipp Thölke,Yorguin-José Mantilla-Ramos,Hamza Abdelhedi,Charlotte Maschke,Arthur Dehgan,Yann Harel,Anirudha Kemtur,Loubna Mekki Berrada,Myriam Sahraoui,Tammy Young,Antoine Bellemare,Clara El Khantour,Mathieu Landry,Annalisa Pascarella,Vanessa Hadid,Etienne Combrisson,Jordan O’Byrne,Karim Jerbi
出处
期刊:NeuroImage
[Elsevier BV]
日期:2023-06-27
卷期号:277: 120253-120253
被引量:200
标识
DOI:10.1016/j.neuroimage.2023.120253
摘要
Machine learning (ML) is increasingly used in cognitive, computational and clinical neuroscience. The reliable and efficient application of ML requires a sound understanding of its subtleties and limitations. Training ML models on datasets with imbalanced classes is a particularly common problem, and it can have severe consequences if not adequately addressed. With the neuroscience ML user in mind, this paper provides a didactic assessment of the class imbalance problem and illustrates its impact through systematic manipulation of data imbalance ratios in (i) simulated data and (ii) brain data recorded with electroencephalography (EEG), magnetoencephalography (MEG) and functional magnetic resonance imaging (fMRI). Our results illustrate how the widely-used Accuracy (Acc) metric, which measures the overall proportion of successful predictions, yields misleadingly high performances, as class imbalance increases. Because Acc weights the per-class ratios of correct predictions proportionally to class size, it largely disregards the performance on the minority class. A binary classification model that learns to systematically vote for the majority class will yield an artificially high decoding accuracy that directly reflects the imbalance between the two classes, rather than any genuine generalizable ability to discriminate between them. We show that other evaluation metrics such as the Area Under the Curve (AUC) of the Receiver Operating Characteristic (ROC), and the less common Balanced Accuracy (BAcc) metric - defined as the arithmetic mean between sensitivity and specificity, provide more reliable performance evaluations for imbalanced data. Our findings also highlight the robustness of Random Forest (RF), and the benefits of using stratified cross-validation and hyperprameter optimization to tackle data imbalance. Critically, for neuroscience ML applications that seek to minimize overall classification error, we recommend the routine use of BAcc, which in the specific case of balanced data is equivalent to using standard Acc, and readily extends to multi-class settings. Importantly, we present a list of recommendations for dealing with imbalanced data, as well as open-source code to allow the neuroscience community to replicate and extend our observations and explore alternative approaches to coping with imbalanced data.
科研通智能强力驱动
Strongly Powered by AbleSci AI