节点文献

利用VB读取方正排版文件提取元数据

Research on metadata extraction by using VB to read from founder typesetting files

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 杨海亮徐用吉

【Author】 YANG Hailiang;XU Yongji;Journal of Northeastern University;

【机构】 东北大学学报编辑部

【摘要】 【目的】为科技期刊自动提取更加全面的元数据提供方法和借鉴。【方法】以方正排版文件为对象,建立了提取元数据的数学模型,同时提出尾部分割算法。然后利用基于对象的VB编程软件编写了自动提取元数据程序。【结果】在分析了方正排版语言特点之后,对方正排版文件进行了字符串替换处理,并建立了分割关键词列表文件,最后将提取的元数据保存到Excel文件中。【结论】实际应用表明,仅几秒钟就可以完成一期数据的提取工作,大大提高了工作效率。

【Abstract】 [Purpose]The objective of this paper is to automatically extract more comprehensive metadata from the journals of science and technology. [Methodology] A mathematical model which takes founder typesetting files as the object is established to extract the metadata,and also the tail segmentation algorithm is advanced. Then,the automatic metadata extraction software is programmed based on VB programming software. [Findings] The strings of founder typesetting files are replaced,with analyzing the founder typesetting language features,and then a segmentation keywords list file is established. Finally,the extracted metadata is saved to the Excel file. [Conclusions]The actual application shows that completing the extraction work of 1 issue consumes only a fewseconds,which greatly improves the work efficiency.

  • 【文献出处】 中国科技期刊研究 ,Chinese Journal of Scientific and Technical Periodicals , 编辑部邮箱 ,2015年06期
  • 【分类号】G230.7
  • 【被引频次】8
  • 【下载频次】116
节点文献中: 

本文链接的文献网络图示:

本文的引文网络