SMM-based text-to-speech synthesis system with speaker interpolation

Title SMM-based text-to-speech synthesis system with speaker interpolation
Author Orhan, Mustafa Cem, Demiroğlu, Cenk
Publication Date: 2011
Publication Place - IEEE
Subject Hidden Markov models, Speaker recognition, Speech synthesis
Type Document
Language Turkish
Digital Yes
Manuscript No
Library: Özyeğin University
Library Asset ID 978-1-4577-0462-8
Record ID 7c63e4c5-3622-4ce7-9037-c908f89a327f
Library Location Electrical & Electronics Engineering
Date 2011
Notes Due to copyright restrictions, the access to the full text of this article is only available via subscription.
Sample Text Hidden Markov Model (HMM) based text-to-speech (TTS) systems offer many advantages compared to the concatenative approach. One of those advantages is the ability to interpolate between different speakers to generate new voices. In this paper, speaker interpolation for HMM-based TTS (HTS) is described and listening test results for the interpolation of English and Turkish voices are presented. Similar to English, we obtained Turkish speech that strongly reflect the interpolation ratio in perceptual similarity. Some insight into the interpolation process is also provided by analysing the spectra of the reference and final voices.
DOI 10.1109/SIU.2011.5929767
View in source Özyeğin University Özyeğin University - Ottoman library catalog search
Özyeğin University - Ottoman library catalog search Özyeğin University

SMM-based text-to-speech synthesis system with speaker interpolation

Author Orhan, Mustafa Cem, Demiroğlu, Cenk
Publication Date 2011
Publication Place - IEEE
Subject Hidden Markov models, Speaker recognition, Speech synthesis
Type Document
Language Turkish
Digital Yes
Manuscript No
Library Özyeğin University
Library Asset ID 978-1-4577-0462-8
Record ID 7c63e4c5-3622-4ce7-9037-c908f89a327f
Library Location Electrical & Electronics Engineering
Date 2011
Notes Due to copyright restrictions, the access to the full text of this article is only available via subscription.
Sample Text Hidden Markov Model (HMM) based text-to-speech (TTS) systems offer many advantages compared to the concatenative approach. One of those advantages is the ability to interpolate between different speakers to generate new voices. In this paper, speaker interpolation for HMM-based TTS (HTS) is described and listening test results for the interpolation of English and Turkish voices are presented. Similar to English, we obtained Turkish speech that strongly reflect the interpolation ratio in perceptual similarity. Some insight into the interpolation process is also provided by analysing the spectra of the reference and final voices.
DOI 10.1109/SIU.2011.5929767
Özyeğin University - Ottoman library catalog search
Özyeğin University You are being redirected...

Please wait