节点文献

基于视觉词汇形状描述的图像表示方法

Image representation based on visual vocabulary shape description

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 王红霞杨克俭张敏艾浩军陈先桥

【Author】 WANG Hongxia 1,YANG Kejian 1,ZHANG Min 2,AI Haojun 2,CHEN Xianqiao 11.School of Computer Science & Technology,Wuhan University of Technology,Wuhan 430063,China 2.School of Computer,Wuhan University,Wuhan 430072,China

【机构】 武汉理工大学计算机科学与技术学院武汉大学计算机学院

【摘要】 针对目前图像表示中引入空间位置信息的空间金字塔匹配方法缺乏对图像中视觉物体平移、缩放和旋转的考虑,提出一种基于视觉词汇形状描述模型的图像表示方法。该方法相对于每个视觉单词的几何中心建立空间几何模型,保证平移不变性;给出对数极坐标空间金字塔匹配,对对数极半径做归一化,保证缩放不变性;在空间金字塔划分过程中确定极角的主方向,从而保证旋转不变性。分别在Caltech-101数据集和自建图像数据集上对该方法进行了验证和比较。实验结果表明,该方法提高了分类识别准确率,特别是对于包含明显平移、缩放和旋转变化的图像数据集;该方法的方差较小,说明其鲁棒性更强。

【Abstract】 The Spatial Pyramid Matching(SPM)approach,which is based on approximate global geometric correspondence,disregards invariance to translation,scale and rotation of visual objects in images.This paper proposes an image representation method based on visual vocabulary shape description model.According to this method,spatial geometric model relative to the geometric center of each visual word is constructed to guarantee translation invariance;this paper presents log polar spatial pyramid matching,log polar radius is normalized and a consistent orientation to visual word is assigned in order to achieve scaling and rotation invariance.Experiments have been conducted for comparing and evaluating the proposed method utilizing the Caltech-101 dataset and this paper’s dataset.Experimental results show that the proposed method improves the classification accuracy,especially for the dataset containing images with obvious translation,scaling and rotation changes,and is more robust because of its smaller variance.

【基金】 国家自然科学基金(No.51179146);武汉市科学技术局科技攻关计划项目(No.201010621208)
  • 【文献出处】 计算机工程与应用 ,Computer Engineering and Applications , 编辑部邮箱 ,2012年21期
  • 【分类号】TP391.41
  • 【被引频次】3
  • 【下载频次】240
节点文献中: 

本文链接的文献网络图示:

本文的引文网络