引用本文:钱卫宁,周傲英.从多角度分析现有聚类算法.软件学报,2002,13(8):1382-1394
【打印本页】   【下载PDF全文】   查看/发表评论  【EndNote】   【RefMan】   【BibTex】
←前一篇|后一篇→ 过刊浏览    高级检索
本文已被:浏览 4911次   下载 8720 本文二维码信息
码上扫一扫!
分享到: 微信 更多
从多角度分析现有聚类算法
钱卫宁1,2, 周傲英1,2
1.复旦大学,智能信息处理开放实验室,上海,200433;2.复旦大学,计算机科学系,上海,200433
摘要:
聚类是数据挖掘中研究的重要问题之一.聚类分析就是把数据集分成簇,以使得簇内数据尽量相似,簇间数据尽量不同.不同的聚类方法采用不同的相似测度和技术.从以下3个角度分析现有流行聚类算法: (1)聚类尺度; (2)算法框架; (3)簇的表示.在此基础上,分析了一些综合或概括了一些其他方法的算法.由于分析从3个角度进行,所提出的方法能够涵盖,并区分绝大多数现有聚类算法.所做的工作是自调节聚类方法以及聚类基准测试研究的基础.
关键词:  数据挖掘  聚类分析  算法
DOI:
分类号:
基金项目:Supported by the National Grand Fundamental Research 973 Program of China under Grant No.G1998030414 (国家重点基础研究发展规划973项目); the National Research Foundation for the Doctoral Program of Higher Education of China under Grant No.99038 (国家教育部博士点基金)
Analyzing Popular Clustering Algorithms from Different Viewpoints
QIAN Wei-ning,ZHOU Ao-ying
Abstract:
Clustering is widely studied in data mining community. It is used to partition data set into clusters so that intra-cluster data are similar and inter-cluster data are dissimilar. Different clustering methods use different similarity definition and techniques. Several popular clustering algorithms are analyzed from three different viewpoints: (1) clustering criteria, (2) cluster representation, and (3) algorithm framework. Furthermore, some new built algorithms, which mix or generalize some other algorithms, are introduced. Since the analysis is from several viewpoints, it can cover and distinguish most of the existing algorithms. It is the basis of the research of self-tuning algorithm and clustering benchmark.
Key words:  data mining  clustering  algorithm

引用本文:
【打印本页】   【下载PDF全文】   查看/发表评论  【EndNote】   【RefMan】   【BibTex】
←前一篇|后一篇→ 过刊浏览    高级检索
本文已被:浏览次   下载  
分享到: 微信 更多
摘要:
关键词:  
DOI:
分类号:
基金项目:
Abstract:
Key words: