<?xml version="1.0" encoding="UTF-8"?>
<mods xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns="http://www.loc.gov/mods/v3" version="3.1" xsi:schemaLocation="http://www.loc.gov/mods/v3 http://www.loc.gov/standards/mods/v3/mods-3-1.xsd">
  <titleInfo>
    <title>Recognition of syllables in tone languages</title>
  </titleInfo>
  <name type="personal">
    <namePart>Tanee Demeechai</namePart>
    <role>
      <roleTerm authority="marcrelator" type="text">creator</roleTerm>
    </role>
  </name>
  <name type="personal">
    <namePart>Makelainen, Kimmo</namePart>
    <role>
      <roleTerm type="text">Chairperson</roleTerm>
    </role>
  </name>
  <name type="personal">
    <namePart>Sadanada, Ramakoti</namePart>
    <role>
      <roleTerm type="text">Examination committee</roleTerm>
    </role>
  </name>
  <name type="personal">
    <namePart>Ahmed, Kazi Mohiuddin</namePart>
    <role>
      <roleTerm type="text">Examination Committee</roleTerm>
    </role>
  </name>
  <name type="personal">
    <namePart>Rajatheva, R.M.A.P</namePart>
    <role>
      <roleTerm type="text">Examination Committee</roleTerm>
    </role>
  </name>
  <name type="personal">
    <namePart>Lee, Chin-Hui</namePart>
    <role>
      <roleTerm type="text">Examination committee</roleTerm>
    </role>
  </name>
  <name type="corporate">
    <namePart>Royal Thai Govenment (RTG)</namePart>
    <role>
      <roleTerm type="text">Scholarship Donor</roleTerm>
    </role>
  </name>
  <typeOfResource>text</typeOfResource>
  <genre authority="marc">series</genre>
  <genre authority="marc">technical report</genre>
  <originInfo>
    <place>
      <placeTerm type="code" authority="marccountry">th</placeTerm>
    </place>
    <place>
      <placeTerm type="text">Bangkok</placeTerm>
    </place>
    <publisher>Asian Institute of Technology</publisher>
    <dateIssued>2000</dateIssued>
    <issuance>continuing</issuance>
  </originInfo>
  <language>
    <languageTerm authority="iso639-2b" type="code">eng</languageTerm>
  </language>
  <physicalDescription>
    <extent>91 p.</extent>
  </physicalDescription>
  <abstract>Spcech recognition of tone languages requires detection of the tone in addition to detection of the consonants and vowels of a syllable. Two approaches for recognition of tonal syllables have been proposed in the literature: joint detection and sequential detection. In joint detection, recognition is done by employing a hidden Markov model (HMM) of connected tonal syllasbles, in which the pitch and its time derivative are included into the feaure vector in addition to the phonetic features. In sequential detection, base syllables (syllables ignoring their tones) are recognized by using a HMM of connected base syllables only; the estimated syllable boundaries are then used for subsequent tone recognition in a separate HMM of tones. Joint detection performs better than sequential detection, but its computational complexity is higher. In this thesis, a new approach caled linked detection is proposed to achieve performance close to that of joint detection with computational complexity close to that of sequential detection. In linked detection, the recognition in the HMM of connected base sykkabkes is modified to periodically take into account also tonal likelihood computed form a HMM of tones. Likeed detection can provide performance that is comparable to the performance of joint detection and superior to that of sequential detection. For a large vocabulary task, the computational complexity of linked detection is much lower than that of joint detection while it is only slightly higher than that of sequential detection. In our experiment on recognition of 173 Thai-language syllables, the worst recognition rate obtained from sequential detection is 72% This result is comparable to results of the IBM Mandarin Call Home system (LIU et al., 1996), where a syllable recognition rate of 50% is reported for an experiment, in which the speech data are conversational telephone speech data and a word-pair language model is applied so that the perplexity is 15.</abstract>
  <note>A dissertaion submitted in partial fulfillment of the requirements for the degree of Doctor of Engineering, School of Engineering and Technology</note>
  <note>Thesis (Ph.D.) - Asian Institute of Technology</note>
  <subject authority="lcsh">
    <topic>Tone (Phonetics)</topic>
  </subject>
  <subject authority="lcsh">
    <topic>Speech perception</topic>
  </subject>
  <subject authority="lcsh">
    <topic>Speech processing systems</topic>
  </subject>
  <relatedItem type="series">
    <titleInfo>
      <title>Dissertation ; no. TC-00-01</title>
    </titleInfo>
    <name type="corporate">
      <namePart>Asian Institute of Technology.</namePart>
      <namePart/>
    </name>
  </relatedItem>
  <identifier type="uri">http://203.159.5.9/ait-thesis/detail.php?q=B00038</identifier>
  <location>
    <url displayLabel="Full-Text">http://203.159.5.9/ait-thesis/detail.php?q=B00038</url>
  </location>
  <recordInfo>
    <recordCreationDate encoding="marc">290501</recordCreationDate>
    <recordChangeDate encoding="iso8601">20260818175617.0</recordChangeDate>
  </recordInfo>
</mods>
