Extraction of tones of speech : an application to the Thai language

By: Call Number: AIT Thesis no.TC-95-5 Contributor(s): Material type: TextSeries: Asian Institute of Technology. Thesis ; no. TC-95-5Publication details: Bangkok : Asian Institute of Technology, 1995Description: 85 leaves : illSubject(s): Online resources: Dissertation note: Thesis (M.Eng.) - Asian Institute of Technology, 1995 Summary: The pitch vanat10n that convey lexical infonnation about the meaning of a word is commonly refeITed to as a tone. Automatic speech recognition of tonal languages requires the recognition of tones associated with the syllables in addition to the phonemes for proper identification of the syllable. Here, in this thesis work, a general model for the automatic recognition of tones of syllables of tonal languages is developed. The pitch contour of the syllable is initially estimated using Subharmonic Summation as the Pitch Determination Algorithm and given as input to vector quantizer for correct recognition of tones associated with the syllable. The input vector is time aligned and pitch n01malised to remove inter and intraspeaker variations before being applied to the vector quantizer. A codebook is implemented containing reference vectors corresponding to each of the tones of the tonal language. A distortion measure is computed between the test vector and each of the reference vectors. The reference vector c01Tesponding to the least distortion is identified as the tone. The perfo1mance evaluation is done in MATLAB environment. Thai language is chosen as the tonal language for the pe1formance evaluation. Four isolated syllables, uttered by four speakers for all the five tones, are used for simulation. For noise-free speech the system gave 100% correct recognition of tones for the four speakers. Recognition rates of 98, 97, 97, 98, 93 and 93 % were obtained for signal-to-noise-ratios of 40 dB, 30 dB, 20 dB, 10 dB, 5 dB and 2 dB respectively for the worst case speaker.
Tags from this library: No tags from this library for this title. Log in to add tags.
Star ratings
    Average rating: 0.0 (0 votes)
Holdings
Cover image Item type Current library Home library Collection Shelving location Call number Materials specified Vol info URL Copy number Status Notes Date due Barcode Item holds Item hold queue priority Course reserves
20-AIT Publication Asian Institute of Technology Library AIT Publications AIT Thesis no.TC-95-5 (Browse shelf(Opens below)) 1 Available 30050120813877
20-AIT Publication Asian Institute of Technology Library AIT Publications AIT Thesis no.TC-95-5 (Browse shelf(Opens below)) 2 Available 30050120813885
40-Archives Asian Institute of Technology Library Archives AIT Thesis no.TC-95-5 (Browse shelf(Opens below)) Available 30050160046081
20-AIT Publication Asian Institute of Technology Library Archives AIT Thesis no.TC-95-5 (Browse shelf(Opens below)) 3 Available 30050211018741

Thesis (M.Eng.) - Asian Institute of Technology, 1995

A thesis submitted in partial fulfilment of the requirement for the degree of Master of Engineering.

The pitch vanat10n that convey lexical infonnation about the meaning of a word is commonly refeITed to as a tone. Automatic speech recognition of tonal languages requires the recognition of tones associated with the syllables in addition to the phonemes for proper identification of the syllable. Here, in this thesis work, a general model for the automatic recognition of tones of syllables of tonal languages is developed. The pitch contour of the syllable is initially estimated using Subharmonic Summation as the Pitch Determination Algorithm and given as input to vector quantizer for correct recognition of tones associated with the syllable. The input vector is time aligned and pitch n01malised to remove inter and intraspeaker variations before being applied to the vector quantizer. A codebook is implemented containing reference vectors corresponding to each of the tones of the tonal language. A distortion measure is computed between the test vector and each of the reference vectors. The reference vector c01Tesponding to the least distortion is identified as the tone. The perfo1mance evaluation is done in MATLAB environment. Thai language is chosen as the tonal language for the pe1formance evaluation. Four isolated syllables, uttered by four speakers for all the five tones, are used for simulation. For noise-free speech the system gave 100% correct recognition of tones for the four speakers. Recognition rates of 98, 97, 97, 98, 93 and 93 % were obtained for signal-to-noise-ratios of 40 dB, 30 dB, 20 dB, 10 dB, 5 dB and 2 dB respectively for the worst case speaker.

There are no comments on this title.

to post a comment.
คัดลอกแล้ว!