声学模型区分性训练中的动态加权数据选取方法 A Variable Weighting Based Training Data Selection Method for Discriminative Training of Acoustic Models期刊界 All Journals 搜尽天下杂志传播学术成果专业期刊搜索期刊信息化学术搜索

声学模型区分性训练中的动态加权数据选取方法

引用本文：	陈斌,牛铜,张连海,李弼程,屈丹.声学模型区分性训练中的动态加权数据选取方法[J].自动化学报,2014,40(12):2899-2907.

作者姓名：	陈斌牛铜张连海李弼程屈丹

作者单位：	1.解放军信息工程大学信息系统工程学院郑州 450002

基金项目：	国家自然科学基金(61175017)资助

摘要：	提出了一种基于动态加权的数据选取方法, 并应用到连续语音识别的声学模型区分性训练中. 该方法联合后验概率和音素准确率选取数据, 首先, 采用后验概率的Beam算法裁剪词图, 在此基础上依据候选词所在候选路径的错误率, 基于后验概率动态的赋予候选词不同的权值; 其次, 通过统计音素对之间的混淆程度, 给易混淆音素对动态地加以不同的惩罚权重, 计算音素准确率; 最后, 在估计得到弧段期望准确率分布的基础上, 采用高斯函数形式对所有竞争弧段的期望音素准确率软加权.实验结果表明, 与最小音素错误准则相比, 该动态加权方法识别准确率提高了0.61%, 可有效减少训练时间.
关键词：	区分性训练语音识别训练数据选取动态加权
收稿时间：	2013-12-30
A Variable Weighting Based Training Data Selection Method for Discriminative Training of Acoustic Models

CHEN Bin,NIU Tong,ZHANG Lian-Hai,LI Bi-Cheng,QU Dan.A Variable Weighting Based Training Data Selection Method for Discriminative Training of Acoustic Models[J].Acta Automatica Sinica,2014,40(12):2899-2907.

Authors:	CHEN Bin NIU Tong ZHANG Lian-Hai LI Bi-Cheng QU Dan

Affiliation:	1.Institute of Information System Engineering, PLA Information Engineering University, Zhengzhou 450002

Abstract:	By combining the phone posterior and phone accuracy, a data selection method based on variable weighting is proposed to improve the discriminative training performance of the acoustic model for continuous speech recognition. Firstly, the word lattice is reduced by using a posterior-based Beam pruning method, and for each hypothesis word a weight is derived from the word error rates of the path containing that word with the posterior. Then, each pair of confusing phones is variably weighted according to a phone confusion matrix, and the modified phone accuracy is calculated by applying those weights. Finally, the distribution of the expected phone accuracies is estimated and all competing arcs are soft weighted using Gaussian functions. Experimental results show that compared with the minimum phone error criterion, the variable weighting method not only improves the recognition rate by 0.61%, but also reduces the required training time.

Keywords:	Discriminative training speech recognition training data selection variable weighting

	点击此处可从《自动化学报》浏览原始摘要信息
	点击此处可从《自动化学报》下载全文

设为首页 | 免责声明 | 关于勤云 | 加入收藏