节点文献

基于隐马可夫模型的邻近方言差异系数研究

Research on Coefficient of Neighboring Dialect Differences Based on Hidden Markov Model

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 王雪飞刘珺

【Author】 WANG Xuefei;LIU Jun;Modern Education Technology Center,Huangshan University;Department of Computer Science and Technology,College of Information Engineering,Huangshan University;

【机构】 黄山学院现代教育技术中心黄山学院信息工程学院计算机科学技术系

【摘要】 量化邻近地域的方言差异性研究,运用方言朗读独立字词文本A形成声音文件M,使用HTK工具包将M文件构造为声学特征参数集S_M,计算方言差异系数。在邻近连续i个地域基础上得到相应的Si_Mi,同时使声音Mi结合对比样本区域(i=0)音-字(词)映射表,形成i村落并对应文本Ai。差异系数ξ定义为Ai与A0(样本区域或村落)之间的文本内容差异之比。分析连续古村落ξ值特征结果表明,方言在邻近3个村落(地理位置)的ξ值介于0.88~1时,差异较小,而当邻近9个村落的ξ值(综合)小于0.6及词组ξ值小于0.2时,差异快速变大,建立方言距离并提出方言半径概念,确认所测试方言的半径为8(8个村落)。

【Abstract】 In the research of quantifying neighboring area s dialect difference,this paper makes people read the independent word text A in dialect to form the sound file M,uses the Hidden Markov Toolkit(HTK) to structure the acoustic feature parameter set SM for the file M,calculates and forms the diversity factor.Using this method in continuous i neighboring regions forms the homologous parameter set SiMi,while using the sound file Mi to compare to the sound-character(word) mapping table of the sample area(i =0),obtaining the text Ai of the village i.The ratio of text content differences between Ai and A0(sample area or village) is defined as diversity factor ξ.By analyzing ξ feature of continuous villages,the paper finds that the dialect has less difference when the ξ value is between 0.88 and 1 in neighboring 3 villages(geographic location),while in 9 villages distance,the ξ value(synthesize) less than 0.6,and the ξvalue of phrases less than 0.2,this difference changes quickly,so this paper establishes the dialect distance and proposes the concept of dialect radius,confirming that the dialect radius is eight(eight villages).

【基金】 国家文物局文化-遗产保护领域科学和技术研究基金资助项目(2013-YB-HT-015)
  • 【文献出处】 计算机工程 ,Computer Engineering , 编辑部邮箱 ,2016年04期
  • 【分类号】TN912.3
  • 【下载频次】131
节点文献中: