A new duration modeling approach for Mandarin speech

doi:10.1109/TSA.2003.814377

Full metadata record

DC Field	Value	Language
dc.contributor.author	Chen, SH	en_US
dc.contributor.author	Lai, WH	en_US
dc.contributor.author	Wang, YR	en_US
dc.date.accessioned	2014-12-08T15:40:41Z	-
dc.date.available	2014-12-08T15:40:41Z	-
dc.date.issued	2003-07-01	en_US
dc.identifier.issn	1063-6676	en_US
dc.identifier.uri	http://dx.doi.org/10.1109/TSA.2003.814377	en_US
dc.identifier.uri	http://hdl.handle.net/11536/27760	-
dc.description.abstract	In this paper, a new duration modeling approach for Mandarin speech is proposed. It explicitly takes several major affecting factors as multiplicative companding factors (CFs) and estimates all model parameters by an EM algorithm. Besides, the three basic Tone 3 patterns (i.e., full tone, half tone and sandhi tone) are also properly considered via using three different Us to separate their affections on syllable duration. Experimental results showed that the variance of the syllable duration was greatly reduced from 180.17 to 2.52 frame(2) (1 frame =5 ms) by the syllable duration modeling to eliminate effects from those affecting factors. Moreover, the estimated Us of those affecting factors agreed well to our prior linguistic knowledge. Two extensions of the duration modeling method are also performed. One is the use of the same technique to model initial and final durations. The other is to replace the multiplicative model with an additive one. Lastly, a preliminary study of applying the proposed model to predict syllable duration for TTS is also performed. Experimental results showed that it outperformed the conventional regressive prediction method.	en_US
dc.language.iso	en_US	en_US
dc.subject	duration modeling	en_US
dc.subject	Mandarin	en_US
dc.subject	text-to-speech	en_US
dc.title	A new duration modeling approach for Mandarin speech	en_US
dc.type	Article	en_US
dc.identifier.doi	10.1109/TSA.2003.814377	en_US
dc.identifier.journal	IEEE TRANSACTIONS ON SPEECH AND AUDIO PROCESSING	en_US
dc.citation.volume	11	en_US
dc.citation.issue	4	en_US
dc.citation.spage	308	en_US
dc.citation.epage	320	en_US
dc.contributor.department	電信工程研究所	zh_TW
dc.contributor.department	Institute of Communications Engineering	en_US
dc.identifier.wosnumber	WOS:000184375100002	-
dc.citation.woscount	13	-
Appears in Collections:	Articles

Files in This Item:

000184375100002.pdf

If it is a zip file, please download the file and unzip it, then open index.html in a browser to view the full text content.