
doi: 10.1400/82445
handle: 20.500.14243/58643 , 11390/856464
A new sinusoidal model based engine for FESTIVAL TTS system which performs the DSP (Digital Signal Pro- cessing) operations (i.e. converting a phonetic input into audio signal) of a diphone-based TTS concatenative sys- tem, taking as input the NLP (Natural Language Process- ing) data (a sequence of phonemes with length and into- nation values elaborated from the text script) computed by FESTIVAL is described. The engine aims to be an alternative to MBROLA and makes use of SMS ("Spectral Modeling Synthesis") repre- sentation, implemented with the CLAM (C++ Library for Audio and Music) framework. This program will be released with open source license (GPL), and will compile everywhere gcc and CLAM do (i.e.: Windows, Linux and Mac OS X operating systems).
Text-to-Speech synthesis, SMS, Sintesi Sinusoidale, TTS
Text-to-Speech synthesis, SMS, Sintesi Sinusoidale, TTS
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 0 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Average | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Average | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Average |
