节点文献
基于短文本的独立语义特征抽取算法
Independent semantic feature extraction algorithm based on short text
【摘要】 提出了一种基于短文本的独立语义特征抽取算法,旨在降低文本向量的稀疏性并提其高语义表示能力。该算法首先采用潜在语义分析降低文本的维数并去除噪声,然后运用独立成份分析方法在潜在语义特征中提取出最能表达语义且相互统计独立的特征。实验表明此算法优于潜在语义索引算法。
【Abstract】 An independent semantic feature extraction algorithm was proposed,aiming at reducing the sparseness of short text and enhancing its capability of semantic expression.The algorithm first makes use of latent semantic indexing to re-duce the dimension and wipe off noise,and then it introduces independent component analysis to extract statistic inde-pendent and semantic features.Experimental results prove the feasibility of the algorithm and demonstrate it is superior to latent semantic indexing.
【基金】 国家自然科学基金资助项目(60475007,60675001)~~
- 【文献出处】 通信学报 ,Journal on Communications , 编辑部邮箱 ,2007年12期
- 【分类号】TP391.1
- 【被引频次】24
- 【下载频次】548