聚类分析
计算机科学
高维数据聚类
数据挖掘
共识聚类
分类
多样性(控制论)
水准点(测量)
领域(数学)
模糊聚类
概念聚类
分拆(数论)
相关聚类
钥匙(锁)
CURE数据聚类算法
数据科学
机器学习
人工智能
数学
大地测量学
组合数学
纯数学
地理
计算机安全
作者
Guoxian Yu,Liang-Rui Ren,Jun Wang,Carlotta Domeniconi,Xiangliang Zhang
标识
DOI:10.1016/j.cosrev.2024.100621
摘要
Clustering is a fundamental data exploration technique to discover hidden grouping structure of data. With the proliferation of big data, and the increase of volume and variety, the complexity of data multiplicity is increasing as well. Traditional clustering methods can provide only a single clustering result, which restricts data exploration to one single possible partition. In contrast, multiple clustering can simultaneously or sequentially uncover multiple non-redundant and distinct clustering solutions, which can reveal multiple interesting hidden structures of the data from different perspectives. For these reasons, multiple clustering has become a popular and promising field of study. In this survey, we have conducted a systematic review of the existing multiple clustering methods. Specifically, we categorize existing approaches according to four different perspectives (i.e., multiple clustering in the original space, in subspaces and on multi-view data, and multiple co-clustering). We summarize the key ideas underlying the techniques and their objective functions, and discuss the advantages and disadvantages of each. In addition, we built a repository of multiple clustering resources (i.e., benchmark datasets and codes). Finally, we discuss the key open issues for future investigation.
科研通智能强力驱动
Strongly Powered by AbleSci AI