节点文献

基于最小生成树的多层次k-Means聚类算法及其在数据挖掘中的应用

Multi-level k-Means Clustering Algorithm Based on Minimum Spanning Tree and Its Application in Data Mining

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 金晓民张丽萍

【Author】 JIN Xiaomin;ZHANG Liping;Institute of Transportation,Inner Mongolia University;Inner Mongolia Engineering Research Center of Testing and Strengthening for Bridges;College of Computer Science and Technology,Inner Mongolia Normal University;

【机构】 内蒙古大学交通学院内蒙古自治区桥梁检测与维修加固工程技术研究中心内蒙古师范大学计算机科学技术学院

【摘要】 针对传统聚类算法存在挖掘效率慢、准确率低等问题,提出一种基于最小生成树的多层次k-means聚类算法,并应用于数据挖掘中.先分析聚类样本的数据类型,根据分析结果设计聚类准则函数;再通过最小生成树对样本数据进行划分,并选取初始聚类中心,将样本的数据空间划分为矩形单元,在矩形单元中对样本对象数据进行计算、降序和选取,得到有效的初始聚类中心,减少数据挖掘时间.实验结果表明,与传统算法相比,该算法可快速、准确地挖掘数据,且挖掘效率提升约50%.

【Abstract】 Aiming at the problem of slow mining efficiency and low accuracy in traditional clustering algorithm,we proposed a multi-level k-means clustering algorithm based on minimum spanning tree,and applied to data mining.Firstly,we analyzed the data types of the clustering samples and designed the clustering criterion function according to the analysis results.Secondly,we divided the sample data by the minimum spanning tree,and selected the initial clustering center.The data space of the sample was divided into rectangular unit,the sample object data was calculated,descended and selected in the rectangular unit,the effective initial clustering center was obtained to reduce the time spent in data mining.The experimental results show that,compared with the traditional algorithm,the proposed method can quickly and accurately excavate the data,and the efficiency of mining is increased by about 50%.

【基金】 国家自然科学基金(批准号:61462071)
  • 【文献出处】 吉林大学学报(理学版) ,Journal of Jilin University(Science Edition) , 编辑部邮箱 ,2018年05期
  • 【分类号】TP311.13
  • 【被引频次】23
  • 【下载频次】648
节点文献中: 

本文链接的文献网络图示:

本文的引文网络