节点文献
基于最大熵的汉语短语结构识别方法
Recognition Method of Chinese Phrase Structure Based on Maximum Entropy
【摘要】 为提高计算机对汉语信息的处理能力,更好地进行浅层句法分析,提出一种基于最大熵的汉语短语结构识别方法。利用词语之间的互信息知识对句子的短语结构边界进行预测,应用最大熵模型建立原子模板与复合模板,选择有效的特征构成特征集,实现对句子短语结构的识别。实例证明,基于互信息的最大熵模型能取得较好的精确率和召回率。
【Abstract】 To improve the computer’s processing capacity on Chinese information,and do better shallow parsing,this paper presents a recognition method of Chinese phrase structure based on Maximum Entropy(ME).The Mutual Information(MI) among the phrases is proposed to achieve boundary prediction of the sentences structure,and the ME model is used to set up atomic and composite templates,selects more effective features for constituting the final feature set.The identification of phrase structure is completed by using the ME method,and good precision and recall are proved in the ME model based on MI by the practical experiment.
【Key words】 shallow parsing; Mutual Information(MI); boundary prediction; Maximum Entropy(ME) model; feature selection;
- 【文献出处】 计算机工程 ,Computer Engineering , 编辑部邮箱 ,2011年16期
- 【分类号】TP391.1
- 【被引频次】9
- 【下载频次】198