Speech Synthesis with Neural Networks
| dc.creator | Karaali, Orhan | |
| dc.creator | Corrigan, Gerald | |
| dc.creator | Gerson, Ira | |
| dc.date | 1998-11-24 | |
| dc.date.accessioned | 2026-07-07T03:23:50Z | |
| dc.date.available | 2026-07-07T03:23:50Z | |
| dc.description | Text-to-speech conversion has traditionally been performed either by concatenating short samples of speech or by using rule-based systems to convert a phonetic representation of speech into an acoustic representation, which is then converted into speech. This paper describes a system that uses a time-delay neural network (TDNN) to perform this phonetic-to-acoustic mapping, with another neural network to control the timing of the generated speech. The neural network system requires less memory than a concatenation system, and performed well in tests comparing it to commercial systems using other technologies. | |
| dc.description | 6 pages, PostScript | |
| dc.identifier | https://arxiv.org/abs/cs/9811031 | |
| dc.identifier | http://arxiv.org/abs/cs/9811031 | |
| dc.identifier | World Congress on Neural Networks (1996) 45-50. San Diego | |
| dc.identifier.uri | http://salesiana.dossiersoluciones.com/handle/123456789/33101 | |
| dc.subject | Neural and Evolutionary Computing | |
| dc.subject | Human-Computer Interaction | |
| dc.subject | I.2.6; K.3.2 | |
| dc.title | Speech Synthesis with Neural Networks | |
| dc.type | text |