Voice Conversion Using Partial Least Squares Regression

13 years 10 months ago

Download www.cs.tut.fi

Abstract--Voice conversion can be formulated as finding a mapping function which transforms the features of the source speaker to those of the target speaker. Gaussian mixture model (GMM)based conversion is commonly used, but it is subject to overfitting. In this paper, we propose to use partial least squares (PLS)-based transforms in voice conversion. To prevent overfitting, the degrees of freedom in the mapping can be controlled by choosing a suitable number of components. We propose a technique to combine PLS with GMMs, enabling the use of multiple local linear mappings. To further improve the perceptual quality of the mapping where rapid transitions between GMM components produce audible artefacts, we propose to low-pass filter the component posterior probabilities. The conducted experiments show that the proposed technique results in better subjective and objective quality than the baseline joint density GMM approach. In speech quality conversion preference tests, the proposed met...

Elina Helander, Tuomas Virtanen, Jani Nurminen, Mo

Real-time Traffic

Density Gmm Method | Joint Density Gmm | Software Engineering | TASLP 2010 | Voice Conversion |

claim paper

Post Info
More Details (n/a)

Added	21 May 2011
Updated	21 May 2011
Type	Journal
Year	2010
Where	TASLP
Authors	Elina Helander, Tuomas Virtanen, Jani Nurminen, Moncef Gabbouj

Comments (0)

Sciweavers

Voice Conversion Using Partial Least Squares Regression

Density Gmm Method | Joint Density Gmm | Software Engineering | TASLP 2010 | Voice Conversion |

Explore & Download

Productivity Tools

Document Tools

Image Tools

Sciweavers