Subsegmental segmental and suprasegmental processing of linear prediction residual for speaker information
Loading...
Date
item.page.authors
Journal Title
Journal ISSN
Volume Title
Publisher
Abstract
The speaker specific information in speech is mostly attributed to the shape size and dynamics of the vocal tract and excitation source The excitation information can be viewed at subsegmental 3 5 msec segmental 10 30 msec and suprasegmental 100 300 msec levels These include glottal cycle activities periodicity and strength of vocal folds vibration and speaker learning habits This work proposes methods to model these information from the linear prediction LP residual and uses them