论文信息 - Document Structure Analysis and Text Normalization for Chinese Putonghua and Cantonese Text-to-Speech Synthesis

Document Structure Analysis and Text Normalization for Chinese Putonghua and Cantonese Text-to-Speech Synthesis

This paper describes our recent effort on document structure analysis (DSA) and text normalization (NORM) for Chinese Putonghua and Cantonese text-to-speech synthesis. A unified framework has been proposed, where DSA and NORM procedures are language-independent for the two-dialects of Chinese. For document structure analysis, regular expressions have been utilized to detect and identify the non-standard-words (NSWs) and punctuations related to document structure; a new document segmentation approach is then proposed by considering the information provided by NSWs and punctuations. For text normalization, a method which considers the contextual information is put forward to handle the ambiguity of the NSWs, symbols and punctuations.

Yuzhuo Zhong | Xinxin Zhou | Chun Yuan | Zhiyong Wu

[1] Dave Burke. Speech Synthesis Markup Language (SSML) , 2007 .

[2] David Yarowsky,et al. Homograph Disambiguation in Text-to-Speech Synthesis , 1997 .

[3] Varol Akman,et al. An Analysis of English Punctuation , 1998 .

[4] Pak-Chung Ching,et al. CU VOCAL: corpus-based syllable concatenation for Chinese speech synthesis across domains and dialects , 2002, INTERSPEECH.

[5] Chiu-yu Tseng,et al. Fluent speech prosody: Framework and modeling , 2005, Speech Commun..

[6] G. Clark,et al. Reference , 2008 .

[7] Shankar Kumar,et al. Normalization of non-standard words , 2001, Comput. Speech Lang..