Universiti Teknologi Malaysia Institutional Repository

Corpus design for Malay corpus-based speech synthesis system

Tan, Tian-Swee and Sh-Hussain, Sh-Hussain (2009) Corpus design for Malay corpus-based speech synthesis system. American Journal of Applied Sciences, 6 (4). pp. 696-702. ISSN 15469239

Full text not available from this repository.

Official URL: http://dx.doi.org/10.3844/ajas.2009.696.702

Abstract

Problem statement: Speech corpus is one of the major components in corpus-based synthesis. The quality and coverage in speech corpus will affect the quality of synthesis speech sound. Approach: This study proposes a corpus design for Malay corpus-based speech synthesis system. This includes the study of design criteria in corpus-based speech synthesis, Malay corpus based database design and the concatenation engine in Malay corpus-based synthesis system. A set of 10 millions digital text corpuses for Malay language has been collected from Malay internet news. This text corpus had been analyzed using word frequency count to find out all high frequency words to be used for designing the sentences for speech corpus. Results: Altogether 381 sentences for speech corpus had been designed using 70% of high frequency words from 10 million text corpus. It consists of 16826 phoneme units and the total storage size is 37.6Mb. All the phone units are phonetically transcribed to preserve the phonetic context of its origin that will be used for phonetic context unit. This speech corpus had been labeled at phoneme level and used for variable length continuous phoneme based concatenation. Speech corpus is one of the major components in corpus-based synthesis. The quality and coverage in speech corpus will affect the quality of synthesized speech sound. Conclusion/Recommendation: This study has proposed a platform for designing speech corpus especially for Malay Text to Speech which can be further enhanced to support more coverage and higher naturalness of synthetic speech.

Item Type:Article
Uncontrolled Keywords:concatenation, corpus-based speech synthesis, speech synthesis, text to speech, unit selection, variable length unit selection
Subjects:R Medicine > R Medicine (General)
Divisions:?? FBSK ??
ID Code:13281
Deposited By: Ms Zalinda Shuratman
Deposited On:29 Jul 2011 09:40
Last Modified:29 Jul 2011 09:40

Repository Staff Only: item control page