Speech Synthesis with Neural Networks

dc.creatorKaraali, Orhan
dc.creatorCorrigan, Gerald
dc.creatorGerson, Ira
dc.date1998-11-24
dc.date.accessioned2026-07-07T03:23:50Z
dc.date.available2026-07-07T03:23:50Z
dc.descriptionText-to-speech conversion has traditionally been performed either by concatenating short samples of speech or by using rule-based systems to convert a phonetic representation of speech into an acoustic representation, which is then converted into speech. This paper describes a system that uses a time-delay neural network (TDNN) to perform this phonetic-to-acoustic mapping, with another neural network to control the timing of the generated speech. The neural network system requires less memory than a concatenation system, and performed well in tests comparing it to commercial systems using other technologies.
dc.description6 pages, PostScript
dc.identifierhttps://arxiv.org/abs/cs/9811031
dc.identifierhttp://arxiv.org/abs/cs/9811031
dc.identifierWorld Congress on Neural Networks (1996) 45-50. San Diego
dc.identifier.urihttp://salesiana.dossiersoluciones.com/handle/123456789/33101
dc.subjectNeural and Evolutionary Computing
dc.subjectHuman-Computer Interaction
dc.subjectI.2.6; K.3.2
dc.titleSpeech Synthesis with Neural Networks
dc.typetext

Files

Collections