节点文献

基于聚类模式的多数据源记录匹配算法

Matching Data Records Among Multi Data Sources Based on Clustering Techniques

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 唐懿芳钟达夫严小卫

【Author】 TANG Yi-fang~1, ZHONG Da-fu~1, YAN Xiao-wei~ 1,2 ~1 (Department of Computer Science, Guangxi Normal University, Guilin 541004, China) ~2 (Department of Computer Science, Sydney University, Austrilia)

【机构】 广西师范大学计算机科学系广西师范大学计算机科学系 广西桂林541004广西桂林541004广西桂林541004澳大利亚悉尼大学计算机科学系悉尼澳大利亚

【摘要】 提出了一种基于聚类技术的多数据源记录匹配算法,该算法运用的罩盖(Canopy)聚类技术是一种专门对付大型数据的聚类方法,此算法不仅是一个与应用领域无关的算法,跟其它模型相比,在保证原有准确程度的前提下,大大地减少了必需的计算量,提高了记录匹配的效率.

【Abstract】 This paper put forward an algorithm, by using the canopy clustering technique which focuses on large data set, to match data records among multi data sources. The algorithm is a kind of domain-independent method, and compare to other model, when it promises the algorithm’s accuracy, this method increases the effectiveness.

【基金】 广西师范大学青年基金资助.
  • 【文献出处】 小型微型计算机系统 ,Mini-micro Systems , 编辑部邮箱 ,2005年09期
  • 【分类号】TP301
  • 【被引频次】19
  • 【下载频次】195
节点文献中: 

本文链接的文献网络图示:

本文的引文网络