节点文献

一种基因数据的聚类并行算法研究

A Study on Parallel Algorithm of the Gene Expression Data Clustering Analysis

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 毛韶阳; 李肯立;

【Author】 MAO Shao-yang1,2, LI Ken-li2 (1 Department of Mathematics, Hunan Institute of Humanities, Science and Technology, Loudi 417000, China; 2 School of Computer and Communication, Hunan University, Changsha 410082, China)

【机构】 湖南人文科技学院数学系; 湖南大学计算机与通信学院 湖南娄底417000 湖南大学计算机与通信学院; 湖南长沙410082;

【摘要】 提出了一种基于密度的聚类并行算法,在APRAM模型的分布式存储系统中,通过欧几里德距离矩阵和密度函数两次时间复杂度为O(n2)的计算,可使聚类过程的时间复杂度变为O(n),以增加一次计算的代价来降低聚类过程的时间复杂度。基于8结点的机群计算实验表明本算法能够达到较同类算法更高的并行加速比,能提高高维生物数据的聚类速度。

【Abstract】 Put forward a clustering parallel algorithms based on the density. Use MPI under the APRAM model, passing twice computing with time complexity is O(n2) that of the Euclidean distance matrix and the density function, can make the time complexity of clustering procedure be O(n), reduce the time complexity of clustering through adding once computing. The experiment based on eight nodes indicates that this algorithm can attain higher parallel accelerate ratio than the same kind algorithm, raise the clustering rate of the high dimension living data.

【基金】 国家自然科学基金项目(60603053);教育部重点项目(105128)
  • 【文献出处】 微电子学与计算机 ,Microelectronics & Computer , 编辑部邮箱 ,2007年09期
  • 【分类号】TP301.6
  • 【被引频次】2
  • 【下载频次】142
节点文献中: 

本文链接的文献网络图示:

本文的引文网络