Author
Orhan, Mustafa Cem, Demiroğlu, Cenk
Publication Date
2011
Publication Place
-
IEEE
Subject
Hidden Markov models, Speaker recognition, Speech synthesis
Type
Document
Language
Turkish
Digital
Yes
Manuscript
No
Library
Özyeğin University
Library Asset ID
978-1-4577-0462-8
Record ID
7c63e4c5-3622-4ce7-9037-c908f89a327f
Library Location
Electrical & Electronics Engineering
Date
2011
Notes
Due to copyright restrictions, the access to the full text of this article is only available via subscription.
Sample Text
Hidden Markov Model (HMM) based text-to-speech (TTS) systems offer many advantages compared to the concatenative approach. One of those advantages is the ability to interpolate between different speakers to generate new voices. In this paper, speaker interpolation for HMM-based TTS (HTS) is described and listening test results for the interpolation of English and Turkish voices are presented. Similar to English, we obtained Turkish speech that strongly reflect the interpolation ratio in perceptual similarity. Some insight into the interpolation process is also provided by analysing the spectra of the reference and final voices.
DOI
10.1109/SIU.2011.5929767