节点文献
基于WDB特征和用户查询请求的Web数据库选择
Data Sources Selection Based on WDB’s Characters and User’s Query
【Author】 Lin Peiguang,Zhao Lin,Zhang Yan,and Nie Peiyao (School of Computer & Information Engineering,Shandong University of Finance,Jinan 250014)
【机构】 山东财政学院计算机信息工程学院;
【摘要】 要实现Deep Web领域中的数据集成,提供一个高效的数据检索策略是集成系统要解决的首要问题.面对众多的Web数据库,选择最恰当的数据库进行查询,实现以更小的代价返回更多的数据是研究的核心问题.针对此问题,提出了基于Web数据库独立样本的Web数据库特征表示和抽取方法,并基于该特征,提出了一种综合考虑查询相关度、返回数据量和数据冗余度3个要素的数据源选择方法.实验证明,该方法能够达到预期的研究目标,能较好地满足集成系统的需求.
【Abstract】 In order to implement the data integration in the domain of Deep Web,providing an effective data retrieval strategy bears the brunt.In the face of a large number of Web databases,it is the core issue of the study that we should select the most appropriate composition of databases to query and obtain the more data at a smaller cost.For this problem,we offer a Web database feature representation and extraction method based on its independent samples,and we also put forward a method of selecting the data source with comprehensively considering three elements,including the relevant degree of query,the amount of data and the data redundancy,based on the WDB’s characters.Experiments show that this method can achieve the desired objectives and can meet the demand of integrating system very well.
【Key words】 characters of Web database; Deep Web; relevant degree; data sources selection;
- 【会议录名称】 NDBC2010第27届中国数据库学术会议论文集(B辑)
- 【会议名称】NDBC2010第27届中国数据库学术会议
- 【会议时间】2010-10-13
- 【会议地点】中国北京
- 【分类号】TP311.13
- 【主办单位】中国计算机学会数据库专业委员会(CCF DBTC)