Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 10 of 10 for “"speaker variability"”.

  1. Hierachical methods for large population speaker identification using telephone speech

    This study focuses on speaker identificat ion. Several problems such as acoustic noise, channel noise, speaker variability, large population of known group of speakers wi thin the system and many others limit good SiD performance. The SiD system extracts speaker specific features from digitised …

    cape-town Repository record for Hierachical methods for large population speaker identification using telephone speech (opens in a new tab)

  2. Bayesian distance metric learning on i-vector for speaker verification

    … metric learning (Bayes_dml) for the task of speaker verification using the i-vector feature representation. We propose a framework that explores the distance constraints between i-vector pairs from the same speaker and different speakers. With an approximation of the distance metric as a …

    mit Repository record for Bayesian distance metric learning on i-vector for speaker verification (opens in a new tab)

  3. Automatic Detection of Landmark Acoustic Cues in Human Speech

    … Mixture Models (GMMs). To remove the effects of speaker variability and different recording environments, methods for normalizing speech-related measurements are proposed and evaluated. For a new speech signal, the normalized speech-related measurements are extracted at each time frame and …

    mit Repository record for Automatic Detection of Landmark Acoustic Cues in Human Speech (opens in a new tab)

  4. Temporal and Aerodynamic Aspects of Velopharyngeal Coarticulation: Effects of Age, Gender and Vowel Height

    … secondary objective was to determine the within speaker variability of the segments.</p> <p>Speakers consisted of 20 children between the ages of 5 and 7 years, 20 children between 9 and 11 years and 20 adult speakers 18 years or older. Nasal and oral air flows were collected from the …

    tenn-hsc Repository record for Temporal and Aerodynamic Aspects of Velopharyngeal Coarticulation: Effects of Age, Gender and Vowel Height (opens in a new tab)

  5. Discriminative and adaptive training for robust speech recognition and understanding

    … conditions: (1) To suppress the effect of inter-speaker variability on speaker-independent DNN acoustic model, speaker-invariant training is proposed to learn a deep representation in the DNN that is both senone-discriminative and speaker-invariant through adversarial multi-task training (2) To …

    gatech Repository record for Discriminative and adaptive training for robust speech recognition and understanding (opens in a new tab)

  6. Brain plasticity in speech training in native English speakers learning mandarin tones

    … The levels of difficulty included progression in speaker variability from one to four speakers and progression through four levels of acoustic exaggeration of duration, pitch range, and pitch contour. Behavioral results for the natural speech stimuli revealed significant training-induced …

    umn Repository record for Brain plasticity in speech training in native English speakers learning mandarin tones (opens in a new tab)

  7. Speaker model adaptation in automatic speech recognition.

    … of automatic speech recognition is to achieve speaker independence. It is generally believed that the main difficulty is the inter-speaker variability in which the acoustic characteristics of different speakers are not the same. There are mainly three approaches to overcome this problem; …

    rgu Repository record for Speaker model adaptation in automatic speech recognition. (opens in a new tab)

  8. Accent Conversion via Formant-based Spectral Mapping and Pitch Contour Modification

    … conversion intends to change the accent of a speaker to a desired accent and preserve the speaker’s voice identity. This technology can offer a number of useful applications. For example, integrating accent conversion to a text-to-speech system (TTS) can produce a voice with a desired accent …

    wlv Repository record for Accent Conversion via Formant-based Spectral Mapping and Pitch Contour Modification (opens in a new tab)

  9. Normalization in the acoustic feature space for improved speech recognition

    … warping during signal analysis to reduce inter-speaker variability. The baseline procedure for training and test data normalization is introduced and optimized so that consistently large improvements in recognition performance are achieved under a variety of acoustic conditions. A technique for …

    aachen Repository record for Normalization in the acoustic feature space for improved speech recognition (opens in a new tab)

  10. Inter- and Intra-Subject Variability: A Palatometric Study

    … articulation files is feasible by examining the variability which exists within and between speakers. Twenty standard American English dialect speakers were fitted with palatometer pseudopalates. Test stimuli were VCV nonsense words using a schwa in the initial position, the 15 palatal …

    byu Repository record for Inter- and Intra-Subject Variability: A Palatometric Study (opens in a new tab)