机器学习
计算机科学
可靠性
人工智能
转化式学习
可靠性(半导体)
集合(抽象数据类型)
质量(理念)
特征(语言学)
工具箱
生成语法
钥匙(锁)
风险分析(工程)
启发式
工作(物理)
稀缺
可转让性
分类
数据集
数据挖掘
启发式
数据质量
数据建模
作者
Yina Fan,Jun Li,Yi Ren,Chao Liu,Weiming Zhang
标识
DOI:10.1021/acs.est.6c00882
摘要
Adsorption behavior of emerging contaminants (ECs) plays a central role in their environmental fate and the efficiency of artificial removal systems. Machine learning (ML) has been extensively leveraged for adsorption prediction. However, most existing research focuses primarily on model construction and application, with limited attention to systematic challenges and corresponding solutions across the modeling workflow, thereby constraining model reliability and transferability. This review systematically summarizes key challenges in model construction, including data scarcity and bias, a low sample-to-feature ratio (SFR), inadequate feature representativeness, and inadequate model applicability and credibility. To address these challenges, this work consolidates a set of targeted improvement strategies, encompassing rigorous data quality assessment, automated feature extraction, cross-system modeling approaches, and the incorporation of external validation. Furthermore, we explore future development directions such as generative data augmentation, physics-informed machine learning (PIML), and automated literature data extraction via large language models. This work provides guidance for prospective ML model developers to build more robust and reliable models, laying the foundation for transformative advances in EC adsorption prediction research and applications.
科研通智能强力驱动
Strongly Powered by AbleSci AI