Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 20 of 53 for “"Human speech"”.

  1. An Investigation of Semantic Invariance in Human Speech

    … the hypothesis of phonetic symbolism in human speech. As a psycholinguistic and epistemological concept, the phonetic symbol is understood to be a structural a priori in sound which correlates with the semantic properties of physiognomic, muscular tension. An extension of this hypothesis …

    nodak Repository record for An Investigation of Semantic Invariance in Human Speech (opens in a new tab)

  2. Automatic Detection of Landmark Acoustic Cues in Human Speech

    … detection of the eight landmark acoustic cues in human speech. Landmarks are key articulatory events, produced as a result of minimal vocal tract constriction (e.g., vowels and glides) or closures and releases in the oral region (e.g., nasal, fricative, and stop consonants). A complete landmark …

    mit Repository record for Automatic Detection of Landmark Acoustic Cues in Human Speech (opens in a new tab)

  3. Optimization under ecological realism reproduces signatures of human speech perception

    … levels of performance similar to those of humans. In particular, optimizing models for ecologically realistic training datasets has helped to yield more human-like model results. In the field of speech recognition, models trained under realistic conditions with simulated cochlear input …

    mit Repository record for Optimization under ecological realism reproduces signatures of human speech perception (opens in a new tab)

  4. A stochastic measure of similarity between dolphin signature whistles

    … dolphin whistle production to a model of human speech production is discussed, providing a basis for the use of human speech recognition techniques for creating whistle models. Discrete hidden Markov models based on vector quantization of linear prediction coefficients are used to create …

    vt Repository record for A stochastic measure of similarity between dolphin signature whistles (opens in a new tab)

  5. Improving performance of a GSM-based speech recognizer

    [page 73 missing] Communication between human beings is important and the most effective way of passing information. Humans are also able to communicate with machines for instance, computers, where keyboards and typing are the means of communication. Most people can speak but not everyone can read …

    cape-town Repository record for Improving performance of a GSM-based speech recognizer (opens in a new tab)

  6. A virtual vocabulary speech recognizer

    A system for the automatic recognition of human speech is described. A commercially available speech recognizer sees its recognition vocabulary increased through the use of virtual memory management techniques. central to the design are issues concerning the nature of speech, its effectiveness as …

    mit Repository record for A virtual vocabulary speech recognizer (opens in a new tab)

  7. Prosody Dependent Speech Recognition on American Radio News Speech

    "Prosody (the melody and rhythm of natural speech), although important for human speech recognition, has not been fully utilized in large vocabulary continuous speech recognition. In this dissertation, we propose a novel ""prosody-dependent speech recognition"" framework, in which word and prosody …

    uiuc Repository record for Prosody Dependent Speech Recognition on American Radio News Speech (opens in a new tab)

  8. A computational memory and processing model for prosody

    … links processing in working memory to prosody in speech, and links different working memory capacities to different prosodic styles. It provides a causal account of prosodic differences and an architecture for reproducing them in synthesized speech. The implemented system mediates text-based …

    mit Repository record for A computational memory and processing model for prosody (opens in a new tab)

  9. Using phonetics to teach children between the ages 10 to 12 of the "Centro Infantil y Guarderia La Florida" to read in french at a basic level (level A1, according to the common european framework of reference for languages), year 2016

    … that studies the articulate language in human speech (Bertil Malmberg “La Fonética”, 1977). In the opinion of linguists such as James Gibson, Phillip Gregory, and many others, this discipline is strongly connected to the ability of children to learn to read, if it is applied with the …

    u-elsalvador Repository record for Using phonetics to teach children between the ages 10 to 12 of the "Centro Infantil y Guarderia La Florida" to read in french at a basic level (level A1, according to the common european framework of reference for languages), year 2016 (opens in a new tab)

  10. Studying dialects to understand human language

    … of dialect variations as a way to understand how humans might process speech. It evaluates some of the important research in dialect identification and draws conclusions about how their results can give insights into human speech processing. A study clustering dialects using k-means clustering is …

    mit Repository record for Studying dialects to understand human language (opens in a new tab)

  11. A "Living Political Dialect": The Science of Language and the Victorian Epic Impulse

    … the study of the origins and nature of human speech, as a powerful model for engaging the diminishing status of hereditary rule and the rise of popular sovereignty. Philologists and natural scientists presented a new understanding of language as a self-enclosed, evolutionary system. This …

    wustl Repository record for A "Living Political Dialect": The Science of Language and the Victorian Epic Impulse (opens in a new tab)

  12. Removing redundancy in speech by modeling forward masking

    … approaches to this, the first being automatic speech recognition (ASR). Even though many state-of-the-art analysis methods, such as LPC, STFT, and MFCC, have been used in speech recognition, the performance of ASR has reached a plateau and the speech decoding problem remains unresolved. …

    uiuc Repository record for Removing redundancy in speech by modeling forward masking (opens in a new tab)

  13. Evaluation of the usability and usefulness of automatic speech recognition among users in South Africa

    An automatic speech recognition (ASR) system is a software application which recognizes human speech, processes it as input, and displays a text version of the speech as output or uses the input as commands for another application's usage. ASR can either be speaker-dependent or speaker-independent. …

    cape-town Repository record for Evaluation of the usability and usefulness of automatic speech recognition among users in South Africa (opens in a new tab)

  14. Robust speech recognition based on spectro-temporal processing

    … for enhancing the robustness of automatic speech recognition systems (ASR) in adverse acoustical conditions. Recent physiological and psychoacoustical findings indicate that spectro-temporal processing plays an important role in human speech perception. Therefore, sigma-pi cells and Gabor …

    oldenburg Repository record for Robust speech recognition based on spectro-temporal processing (opens in a new tab)

  15. Time domain segmentation of speech signals

    … intense research is the computer recognition of human speech. A basic component of almost all speech recognition schemes is the capability to distinguish silence and noise from speech segments and voiced from unvoiced segments. This thesis examines the segmentation of isolated speech into voiced, …

    iastate Repository record for Time domain segmentation of speech signals (opens in a new tab)

  16. Decoding the Depths: Developing a Click Separator for Predictive Speaker Recognition in Sperm Whale Conversations Using Machine Learning

    … sperm whale communication similar to how human speech is recorded and transformed into written text. Sperm whales communicate using a sophisticated system of clicks and codas. Listening through hundreds of hours of sperm whale audio recordings, individuals have been producing annotations …

    mit Repository record for Decoding the Depths: Developing a Click Separator for Predictive Speaker Recognition in Sperm Whale Conversations Using Machine Learning (opens in a new tab)

  17. Classification of vocal fold vibration as regular or irregular in normal, voiced speech

    … serves an important communicative function in human speech and occurs allophonically in American English. This thesis uses cues from both the temporal and frequency domains - such as fundamental frequency, normalized RMS amplitude, smoothed-energy-difference amplitude (a measure of abruptness …

    mit Repository record for Classification of vocal fold vibration as regular or irregular in normal, voiced speech (opens in a new tab)

  18. A sociophonetic analysis of female-sounding virtual assistants

    … Alexa) are increasingly anthropomorphized by humans and viewed as active interlocutors, it raises questions about the social information indexed by machine voices. This thesis provides a preliminary exploration of the relationship between human sociophonetics, social expectations, and …

    emich Repository record for A sociophonetic analysis of female-sounding virtual assistants (opens in a new tab)

  19. A verification experiment of the second formant transition feature as a perceptual cue in natural speech

    … features that are used as perceptual cues in human speech perception. A brief history of this research is given, with emphasis on one important feature, the second formant (F2) transition. A review of historical arguments made for and against its role in the perception of speech, as well as …

    uiuc Repository record for A verification experiment of the second formant transition feature as a perceptual cue in natural speech (opens in a new tab)

  20. Closed-loop auditory-based representation for robust speech recognition

    A closed-loop auditory based speech feature extraction algorithm is presented to address the problem of unseen noise for robust speech recognition. This closed-loop model is inspired by the possible role of the medial olivocochlear (MOC) efferent system of the human auditory periphery, which has …

    mit Repository record for Closed-loop auditory-based representation for robust speech recognition (opens in a new tab)

Page 1 of 3