缺少数据
人工神经网络
阿达布思
计算机科学
集成学习
氢脆
人工智能
模式识别(心理学)
数据挖掘
机器学习
材料科学
腐蚀
分类器(UML)
冶金
作者
Xujie Gong,Ruichao Lei,Ruize Sun,Xue Jiang,Yanjing Su,Yu Yan
出处
期刊:
[Wiley]
日期:2024-05-01
卷期号:2 (2)
被引量:2
摘要
Abstract Accurately and quickly predicting hydrogen embrittlement performance is critical for the service of metal materials. However, due to multi‐source heterogeneity, existing hydrogen embrittlement data are missing, making it impractical to train reliable machine learning models. In this study, we proposed an ensemble learning training strategy for missing data based on the Adaboost algorithm. This method introduced a mask matrix with missing data and enabled each round of training to generate sub‐datasets, considering missing value information. The strategy first trained a subset of features based on the existing dataset and a selected method and continuously focused on the combination of features with the highest error for iterative training, where the mask matrix of the missing data was used as the input to fit the weights of each base learner using a neural network. Compared with directly modeling on highly sparse data, the predictive ability of this strategy was significantly improved by approximately 20%. In addition, in the testing of new samples, the predicted mean absolute error of the new model was successfully reduced from 0.2 to 0.09. This strategy offers good adaptability to the hydrogen embrittlement sensitivity of different sizes and can avoid interference from feature importance caused by filling data.
科研通智能强力驱动
Strongly Powered by AbleSci AI