特征选择
帕累托原理
计算机科学
冗余(工程)
特征(语言学)
模式识别(心理学)
人工智能
多目标优化
选择(遗传算法)
最小冗余特征选择
数据挖掘
算法
机器学习
数学优化
数学
哲学
操作系统
语言学
作者
Amin Hashemi,Mohammad Bagher Dowlatshahi,Hossein Nezamabadi–pour
标识
DOI:10.1016/j.ins.2021.09.052
摘要
Multi-label learning algorithms have significant challenges due to high-dimensional feature space and noises in multi-label datasets. Feature selection methods are effective techniques to deal with these problems. ParetoCluster is an effective multi-label feature selection algorithm based on Pareto dominance and cluster analysis concepts which considers each label an objective function. This algorithm loses its effectiveness to differentiate features when dealing with high labeled datasets and makes most features incomparable (e.g., when most features fall into the first layer). Thus, a cluster analysis criterion in ParetoCluster will play a decisive role in determining the most relevant features. Bearing this in mind, in this paper, we have modeled the multi-label feature selection problem into a bi-objective optimization problem regarding the relevancy and redundancy degree of the features. We then handle it using Pareto dominance in this bi-objective domain. To illustrate the optimality and efficiency of the proposed method, we have compared our approach against some similar techniques.
科研通智能强力驱动
Strongly Powered by AbleSci AI