Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 5 of 5 for “"speech modeling"”.

  1. Speech Representation Models for Speech Synthesis and Multimodal Speech Recognition

    The field of speech recognition has seen steady advances over the last two decades, leading to the accurate, real-time recognition systems available on mobile phones today. In this thesis, I apply speech modeling techniques developed for recognition to two other speech problems: speech synthesis …

    mit Repository record for Speech Representation Models for Speech Synthesis and Multimodal Speech Recognition (opens in a new tab)

  2. Multimodal Fusion With Applications to Audio -Visual Speech Recognition

    … of the intermodal couplings in audio-visual speech recognition and in multichannel biometrics defy a universal fusion method for both applications. For audio-visual speech modeling, we propose a novel sensory fusion method based on the coupled hidden Markov models (CHMMs). The CHMM framework …

    uiuc Repository record for Multimodal Fusion With Applications to Audio -Visual Speech Recognition (opens in a new tab)

  3. Articulatory features for robust visual speech recognition

    This thesis explores a novel approach to visual speech modeling. Visual speech, or a sequence of images of the speaker's face, is traditionally viewed as a single stream of contiguous units, each corresponding to a phonetic segment. These units are defined heuristically by mapping several visually …

    mit Repository record for Articulatory features for robust visual speech recognition (opens in a new tab)

  4. Probabilistic generative modeling of speech

    Speech processing refers to a set of tasks that involve speech analysis and synthesis. Most speech processing algorithms model a subset of speech parameters of interest and blur the rest using signal processing techniques and feature extraction. However, evidence shows that many speech parameters …

    uiuc Repository record for Probabilistic generative modeling of speech (opens in a new tab)