Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 20 of 82 for “"speech data"”.

  1. Using question-specific vocabularies to support speech data collection with SALAAM

    … settings for information dissemination and data collection. This provides an opportunity to reduce the information gap in low-resource settings in which low-literacy is a huge hindrance to the adoption of Information Communication Technologies (ICTs). Since the languages spoken in these …

    cape-town Repository record for Using question-specific vocabularies to support speech data collection with SALAAM (opens in a new tab)

  2. Single-Case Pilot Study For Longitudinal Analysis Of Referential Failures And Sentiment In Schizophrenic Speech From Client-Centered Psychotherapy Recordings

    … language features in schizophrenic disordered speech, the relative stability of these language features in longitudinal samples is still unknown. This longitudinal pilot study analyzed schizophrenic disordered speech data from the archival therapy audio recordings of one patient spanning 23 …

    national-louis Repository record for Single-Case Pilot Study For Longitudinal Analysis Of Referential Failures And Sentiment In Schizophrenic Speech From Client-Centered Psychotherapy Recordings (opens in a new tab)

  3. Exploring the dimensionality of speech using manifold learning and dimensionality reduction methods

    Many previous investigations have indicated that speech data has inherent low-dimensional structure and that it may be possible to efficiently represent speech using only a small number of parameters. This view is motivated by the fact that articulatory movement is limited by physiological …

    dcu Repository record for Exploring the dimensionality of speech using manifold learning and dimensionality reduction methods (opens in a new tab)

  4. Unsupervised speech processing with applications to query-by-example spoken term detection

    … searching and extracting useful information from speech data in a completely unsupervised setting. In many real world speech processing problems, obtaining annotated data is not cost and time effective. We therefore ask how much can we learn from speech data without any transcription. To address …

    mit Repository record for Unsupervised speech processing with applications to query-by-example spoken term detection (opens in a new tab)

  5. Self-Supervised Learning for Speech Processing

    … learning algorithms on large amounts of labeled speech data have achieved remarkable performance on various spoken language processing applications, often being the state of the arts on the corresponding leaderboards. However, the fact that training these systems relies on large amounts of …

    mit Repository record for Self-Supervised Learning for Speech Processing (opens in a new tab)

  6. Fractal based speech recognition and synthesis

    … message is most often the primary purpose of speech com­munication and the recognition of this message by machine that would be most useful. This research consists of two major parts. The first part presents a novel and promis­ing approach for estimating the degree of recognition of speech

    de-montfort Repository record for Fractal based speech recognition and synthesis (opens in a new tab)

  7. Semi-supervised cycle-consistency training for end-to-end ASR using unpaired speech

    … a new method to train end-to-end automatic speech recognition (ASR) models using unpaired speech. In general, large amounts of paired data (speech and text) are needed to train an end-to-end automatic speech recognition system. To alleviate the problem of limited paired data, the idea of …

    uiuc Repository record for Semi-supervised cycle-consistency training for end-to-end ASR using unpaired speech (opens in a new tab)

  8. Automated Testbench Generation for Communication Systems

    … methods, a VHDL model was constructed for the speech-coding channel of the Global System for Mobile Communication (GSM). GSM is the Pan-European digital mobile telephony standard specified by the European Telecommunication Standards Institute (ETSI). This thesis emphasizes the error detection …

    vt Repository record for Automated Testbench Generation for Communication Systems (opens in a new tab)

  9. SPEECH DEVELOPMENT IN CHILDREN WITH CLEFT LIP AND PALATE

    … cleft palate and cleft lip and palate from pre-speech to the age of 4;6. The aim of the project was to investigate the extent to which the cleft palate condition affects the nature and chronology of phonetic and phonological development. The investigation comprised two studies. Eight children …

    de-montfort Repository record for SPEECH DEVELOPMENT IN CHILDREN WITH CLEFT LIP AND PALATE (opens in a new tab)

  10. Algorithms and low power hardware for keyword spotting

    … consumption while doing real-time processing of speech data. The algorithm based on convolutional neural network (CNN) delivers high accuracy with small model size that can be stored in on-chip memory. However, the state-of-the-art NN accelerators either target at complex tasks using large CNN …

    mit Repository record for Algorithms and low power hardware for keyword spotting (opens in a new tab)

  11. Autoregressive hidden Markov models and the speech signal

    … (HMM) and demonstrates its application to the speech signal. This new variant of the HMM is built upon the mathematical structure of the HMM and linear prediction analysis of speech signals. By incorporating these two methods into one inference algorithm, linguistic structures are inferred from …

    uiuc Repository record for Autoregressive hidden Markov models and the speech signal (opens in a new tab)

  12. A Comparison of Beijing and Taiwan Mandarin Tone Register: An Acoustic Analysis of Three Native Speech Styles

    … by means of an acoustic analysis of three speech styles. Speech styles included spontaneous interview, spontaneous descriptive, and controlled read sentential speech. Data analysis included long segments of recorded speech in order to discern any statistically significant pitch register …

    byu Repository record for A Comparison of Beijing and Taiwan Mandarin Tone Register: An Acoustic Analysis of Three Native Speech Styles (opens in a new tab)

  13. Unsupervised modeling of latent topics and lexical units in speech audio

    Zero-resource speech processing involves the automatic analysis of a collection of speech data in a completely unsupervised fashion without the benefit of any transcriptions or annotations of the data. In this thesis, we describe a zero-resource framework that automatically discovers important …

    mit Repository record for Unsupervised modeling of latent topics and lexical units in speech audio (opens in a new tab)

  14. Data-Centric Machine Learning for Speech and Audio

    … is growing recognition of the importance of data-centric methods for building machine learning systems. Data-centric methods assume a fixed model and iterate over the data to improve system performance. This is in contrast to traditional model-centric approaches, which assume a fixed dataset …

    cuny-grad Repository record for Data-Centric Machine Learning for Speech and Audio (opens in a new tab)

  15. Relative-fuzzy: a novel approach for handling complex ambiguity for software engineering of data mining models

    … ambiguity type of uncertainty that may exist in data, for software engineering of predictive Data Mining (DM) classification models. The proposed approach is based on Relative-Fuzzy Logic (RFL), a novel type of fuzzy logic. RFL defines a new formulation of the problem of ambiguity type of …

    de-montfort Repository record for Relative-fuzzy: a novel approach for handling complex ambiguity for software engineering of data mining models (opens in a new tab)

  16. Pronunciation learning for automatic speech recognition

    … remains the Achilles heel of modern automatic speech recognizers (ASRs). Unlike stochastic acoustic and language models that learn the values of their parameters from training data, the baseform pronunciations of words in an ASR vocabulary are typically specified manually, and do not change, …

    mit Repository record for Pronunciation learning for automatic speech recognition (opens in a new tab)

  17. A study of adaptive enhancement methods for improved distant speech recognition

    Automatic speech recognition systems trained on speech data recorded by microphones placed close to the speaker tend to perform poorly on speech recorded by microphones placed farther away from the speaker due to reverberation effects and background noise. I designed and implemented a variety of …

    mit Repository record for A study of adaptive enhancement methods for improved distant speech recognition (opens in a new tab)

  18. Echolocation: Using Word-Burst Analysis to Rescore Keyword Search Candidates in Low-Resource Languages

    <p>State of the art technologies for speech recognition are very accurate for heavily studied languages like English. They perform poorly, though, for languages wherein the recorded archives of speech data available to researchers are relatively scant. In the context of these low-resource …

    cuny-grad Repository record for Echolocation: Using Word-Burst Analysis to Rescore Keyword Search Candidates in Low-Resource Languages (opens in a new tab)

  19. Computer assisted audiometric evaluation system

    … personal computer to perform pure tone and speech tests and · comprises a plug-in card and custom software. The card contains pure tone and masking noise generators, together with amplifiers for a. set of headphones .and bone conduction transducer, patient and audiologist microphone …

    cape-town Repository record for Computer assisted audiometric evaluation system (opens in a new tab)

Page 1 of 5