节点文献

基于本体的文档引文元数据信息抽取

Ontology-based document citation metadata extraction

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 郭志鑫

【Author】 Guo,Zhixin

【机构】 华中科技大学计算机科学与技术学院集群与网格计算实验室

【摘要】 结合本体技术,提出了一种新的从文档中抽取引文元数据信息的方法。该方法采用模式匹配方式,可以从文档中提取作者、标题、日期等信息,并使用OWL本体描述语言进行形式化,为进一步的语义搜索和语义存储奠定基础。实验数据证明了该方法的有效性。

【Abstract】 A new method using ontology to extract citation metadata from technical documents is proposed in this paper. By the wayof pattern matching, it can get metadata such as authors, titles and publishing date, and it uses OWL ontology describing language toformulate the extracted metadata, which assists the semantic searching and storage. The experiment proved its efficiency.

【关键词】 信息抽取语义网本体模式匹配
【Key words】 information extractionsemantic webontologypattern matching
【基金】 973计划基金(2003CB317003)项目支持
  • 【文献出处】 微计算机信息 , 编辑部邮箱 ,2006年18期
  • 【分类号】TP391.1
  • 【被引频次】28
  • 【下载频次】384
节点文献中: