节点文献
基于SVM的日文网页分类
Classify Japanese Document by Support Vector Machine
【摘要】 网页分类是使用机器学习算法实现网页类别的自动标注。提出了一种基于SVM的日文网页分类方法,针对日文的特点,设计日文词素词典与规则库,并以此为基础进行日文分词和特征表示,然后使用互信息度进行特征选择,最后应用SVM来构造分类超平面,对日文网页进行分类。最后通过实验进行了验证。
【Abstract】 Web classification uses Machine Learning algorithm to tag Web automatically.This paper propose a method to classify Japanese Web pages based on support vector machine.The morpheme diction-ary and rule library are designed according to the feature of Japanese,which are used to segment and present features.Then use Mutual Information to select feature,build the hyperplane to classify the Japan-ese Web page.The positive results demonstrate the performance on a challenging problem.
【基金】 国家“863”计划基金资助项目(2004AA117010-05)
- 【文献出处】 广西师范大学学报(自然科学版) ,Journal of Guangxi Normal University(Natural Science Edition) , 编辑部邮箱 ,2007年02期
- 【分类号】TP391.1
- 【被引频次】1
- 【下载频次】74