Delivery included to the United States

Speech Recognition Using the Mellin Transform

Speech Recognition Using the Mellin Transform

Paperback (26 Oct 2014)

  • $15.60
Add to basket

Includes delivery to the United States

10+ copies available online - Usually dispatched within 7 days

Publisher's Synopsis

The purpose of this research was to improve performance in speech recognition. Specifically, a new approach was investigating by applying an integral transform known as the Mellin transform (MT) on the output of an auditory model to improve the recognition rate of phonemes through the scale-invariance property of the Mellin transform. Scale-invariance means that as a time-domain signal is subjected to dilations, the distribution of the signal in the MT domain remains unaffected. An auditory model was used to transform speech waveforms into images representing how the brain "sees" a sound. The MT was applied and features were extracted. The features were used in a speech recognizer based on Hidden Markov Models. The results from speech recognition experiments showed an increase in recognition rates for some phonemes compared to traditional methods.

Book information

ISBN: 9781502959430
Publisher: Createspace Independent Publishing Platform
Imprint: Createspace Independent Publishing Platform
Pub date:
Language: English
Number of pages: 54
Weight: 149g
Height: 280mm
Width: 216mm
Spine width: 3mm