首页 | 本学科首页   官方微博 | 高级检索  
     


A codebook compensative voice morphing algorithm based on maximum likelihood estimation
Authors:Ning Xu  Zhen Yang  Linhua Zhang
Affiliation:1. Institute of Signal Processing and Transmission,Nanjing University of Post & Telecommunications,Nanjing 213002,China
2. College of Telecommunication & Information Engineering,Nanjing University of Post & Telecommunications,Nanjing 210003,China
Abstract:This paper presents an improved voice morphing algorithm based on Gaussian Mixture Model (GMM) which overcomes the traditional one in the terms of overly smoothed problems of the converted spectral and discontinuities between frames.Firstly,a maximum likelihood estimation for the model is introduced for the alleviation of the inversion of high dimension matrixes caused by traditional conversion function.Then,in order to resolve the two problems associated with the baseline,a codebook compensation technique and a time domain medial filter are applied.The results of listening evaluations show that the quality of the speech converted by the proposed method is significantly better than that by the traditional GMM method,and the Mean Opinion Score (MOS) of the converted speech is improved from 2.5 to 3.1 and ABX score from 38% to 75%.
Keywords:Maximum-Likelihood(ML)estimation  Codebook compensation  Medial filter  Voice morphing
本文献已被 CNKI 维普 万方数据 SpringerLink 等数据库收录!
点击此处可从《电子科学学刊(英文版)》浏览原始摘要信息
点击此处可从《电子科学学刊(英文版)》下载全文
设为首页 | 免责声明 | 关于勤云 | 加入收藏

Copyright©北京勤云科技发展有限公司  京ICP备09084417号