节点文献

基于时空信息和非负成分表示的动作识别

Action recognition based on spatio-temporal information and nonnegative component representation

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 王健弘张旭章品正姜龙玉罗立民

【Author】 Wang Jianhong;Zhang Xu;Zhang Pinzheng;Jiang Longyu;Luo Limin;Laboratory of Image Science and Technology,Southeast University;

【机构】 东南大学影像科学与技术实验室

【摘要】 为充分利用时空分布信息及视觉单词间的关联信息,提出了一种新的时空非负成分表示方法(ST-NCR)用于动作识别.首先,基于视觉词袋(Bo VW)表示,利用混合高斯模型对每个视觉单词所包含的局部特征的时空位置分布进行建模,计算时空Fisher向量(STFV)来描述特征位置的时空分布;然后,利用非负矩阵分解从Bo VW表示中学习动作基元并对动作视频进行编码.为有效融合时空信息,采用基于图正则化的非负矩阵分解,并且将STFV作为图正则化项的一部分.在3个公共数据库上对该方法进行了测试,结果表明,相比于Bo VW表示和不带时空信息的非负成分表示方法,该方法能够提高动作识别率.

【Abstract】 To make full use of spatial-temporal information and the relationship among different visual words,a novel spatial-temporal nonnegative component representation method( ST-NCR) is proposed for action recognition. First,based on Bo VW( bag of visual words) representation,the locations of local features belonging to each visual word are modeled with the Gaussian mixture model,and a spatio-temporal Fisher vector( STFV) is calculated to describe the location distribution of local features. Then,nonnegative matrix factorization( NMF) is employed to learn the action components and encode the action video samples. To incorporate the spatial-temporal cues for final representation,the graph regularized NMF( GNMF) is adopted,and STFV is used as part of graph regularization. The proposed method is extensively evaluated on three public datasets. Experimental results demonstrate that compared with Bo VW representation and nonnegative component representation without spatio-temporal information,the method can obtain better action recognition accuracy.

【基金】 国家自然科学基金青年科学基金资助项目(61401085);教育部留学归国人员科研启动基金资助项目(2015)
  • 【文献出处】 东南大学学报(自然科学版) ,Journal of Southeast University(Natural Science Edition) , 编辑部邮箱 ,2016年04期
  • 【分类号】TP391.41
  • 【下载频次】74
节点文献中: 

本文链接的文献网络图示:

本文的引文网络