Blind Normalization of Speech From Different Channels and Speakers
| dc.creator | Levin, David N. | |
| dc.date | 2002-04-02 | |
| dc.date.accessioned | 2026-07-07T03:18:14Z | |
| dc.date.available | 2026-07-07T03:18:14Z | |
| dc.description | This paper describes representations of time-dependent signals that are invariant under any invertible time-independent transformation of the signal time series. Such a representation is created by rescaling the signal in a non-linear dynamic manner that is determined by recently encountered signal levels. This technique may make it possible to normalize signals that are related by channel-dependent and speaker-dependent transformations, without having to characterize the form of the signal transformations, which remain unknown. The technique is illustrated by applying it to the time-dependent spectra of speech that has been filtered to simulate the effects of different channels. The experimental results show that the rescaled speech representations are largely normalized (i.e., channel-independent), despite the channel-dependence of the raw (unrescaled) speech. | |
| dc.description | 4 pages, 2 figures | |
| dc.identifier | https://arxiv.org/abs/cs/0204003 | |
| dc.identifier | http://arxiv.org/abs/cs/0204003 | |
| dc.identifier.uri | http://salesiana.dossiersoluciones.com/handle/123456789/31035 | |
| dc.subject | Computation and Language | |
| dc.subject | I.2.7 | |
| dc.title | Blind Normalization of Speech From Different Channels and Speakers | |
| dc.type | text |