首页 | 本学科首页   官方微博 | 高级检索  
     

声学模型区分性训练中的动态加权数据选取方法
引用本文:陈斌,牛铜,张连海,李弼程,屈丹.声学模型区分性训练中的动态加权数据选取方法[J].自动化学报,2014,40(12):2899-2907.
作者姓名:陈斌  牛铜  张连海  李弼程  屈丹
作者单位:1.解放军信息工程大学信息系统工程学院 郑州 450002
基金项目:国家自然科学基金(61175017)资助
摘    要:提出了一种基于动态加权的数据选取方法, 并应用到连续语音识别的声学模型区分性训练中. 该方法联合后验概率和音素准确率选取数据, 首先, 采用后验概率的Beam算法裁剪词图, 在此基础上依据候选词所在候选路径的错误率, 基于后验概率动态的赋予候选词不同的权值; 其次, 通过统计音素对之间的混淆程度, 给易混淆音素对动态地加以不同的惩罚权重, 计算音素准确率; 最后, 在估计得到弧段期望准确率分布的基础上, 采用高斯函数形式对所有竞争弧段的期望音素准确率软加权.实验结果表明, 与最小音素错误准则相比, 该动态加权方法识别准确率提高了0.61%, 可有效减少训练时间.

关 键 词:区分性训练    语音识别    训练数据选取    动态加权
收稿时间:2013-12-30

A Variable Weighting Based Training Data Selection Method for Discriminative Training of Acoustic Models
CHEN Bin,NIU Tong,ZHANG Lian-Hai,LI Bi-Cheng,QU Dan.A Variable Weighting Based Training Data Selection Method for Discriminative Training of Acoustic Models[J].Acta Automatica Sinica,2014,40(12):2899-2907.
Authors:CHEN Bin  NIU Tong  ZHANG Lian-Hai  LI Bi-Cheng  QU Dan
Affiliation:1.Institute of Information System Engineering, PLA Information Engineering University, Zhengzhou 450002
Abstract:By combining the phone posterior and phone accuracy, a data selection method based on variable weighting is proposed to improve the discriminative training performance of the acoustic model for continuous speech recognition. Firstly, the word lattice is reduced by using a posterior-based Beam pruning method, and for each hypothesis word a weight is derived from the word error rates of the path containing that word with the posterior. Then, each pair of confusing phones is variably weighted according to a phone confusion matrix, and the modified phone accuracy is calculated by applying those weights. Finally, the distribution of the expected phone accuracies is estimated and all competing arcs are soft weighted using Gaussian functions. Experimental results show that compared with the minimum phone error criterion, the variable weighting method not only improves the recognition rate by 0.61%, but also reduces the required training time.
Keywords:Discriminative training  speech recognition  training data selection  variable weighting
点击此处可从《自动化学报》浏览原始摘要信息
点击此处可从《自动化学报》下载全文
设为首页 | 免责声明 | 关于勤云 | 加入收藏

Copyright©北京勤云科技发展有限公司  京ICP备09084417号