Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 20 of 59 for “"Speech synthesis"”.

  1. Complex Waveform Phonetic Speech Synthesis

    Made available in DSpace on 2014-12-10T19:07:23Z (GMT). No. of bitstreams: 1 7511692.pdf: 5802945 bytes, checksum: 48625785c12e8319a6404f2f0e422568 (MD5) Previous issue date: 1974

    uiuc Repository record for Complex Waveform Phonetic Speech Synthesis (opens in a new tab)

  2. An investigation of speech synthesis parameters

    The model of speech production generally used in speech synthesis is that of a source modified by a digital filter. The major difference between a number of models is the form of the digital filter. The purpose of this research is to compare the properties of these filters when used for speech

    soton Repository record for An investigation of speech synthesis parameters (opens in a new tab)

  3. Speech Enhancement Using Speech Synthesis Techniques

    <p>Traditional speech enhancement systems reduce noise by modifying the noisy signal to make it more like a clean signal, which suffers from two problems: under-suppression of noise and over-suppression of speech. These problems create distortions in enhanced speech and hurt the quality of the …

    cuny-grad Repository record for Speech Enhancement Using Speech Synthesis Techniques (opens in a new tab)

  4. Time-domain concatenative text-to-speech synthesis.

    … framework for time-domain concatenative speech synthesis (TDCSS) is presented and evaluated. In this framework, speech segments are extracted from CV, VC, CVC and CC waveforms, and abutted. Speech rhythm is controlled via a single duration parameter, which specifies the initial portion of …

    bournemouth Repository record for Time-domain concatenative text-to-speech synthesis. (opens in a new tab)

  5. Articulatory Speech Synthesis and Speech Production Modelling

    … work, we carefully argue that the articulatory speech production model has the potential to flexibly synthesize natural-quality speech sounds and to provide a compact computational model for speech production that can be beneficial to a wide range of areas in speech signal processing.

    uiuc Repository record for Articulatory Speech Synthesis and Speech Production Modelling (opens in a new tab)

  6. Speech synthesis by rule: an acoustic domain approach.

    Massachusetts Institute of Technology. Dept. of Electrical Engineering. Thesis. 1967. Ph.D.

    mit Repository record for Speech synthesis by rule: an acoustic domain approach. (opens in a new tab)

  7. Natural-sounding speech synthesis using variable-length units

    Thesis (M.Eng.)--Massachusetts Institute of Technology, Dept. of Electrical Engineering and Computer Science, 1998.

    mit Repository record for Natural-sounding speech synthesis using variable-length units (opens in a new tab)

  8. Corpus-based unit selection for natural-sounding speech synthesis

    Speech synthesis is an automatic encoding process carried out by machine through which symbols conveying linguistic information are converted into an acoustic waveform. In the past decade or so, a recent trend toward a non-parametric, corpus-based approach has focused on using real human speech as …

    mit Repository record for Corpus-based unit selection for natural-sounding speech synthesis (opens in a new tab)

  9. Language generation and speech synthesis in dialogues for language learning

    Since 1989, the Spoken Language Systems group has developed an array of applications that allow users to interact with computers using natural spoken language. A recent project of interest is to develop an interactive conversational system to assist students in mastering a foreign language. The …

    mit Repository record for Language generation and speech synthesis in dialogues for language learning (opens in a new tab)

  10. Speech Representation Models for Speech Synthesis and Multimodal Speech Recognition

    The field of speech recognition has seen steady advances over the last two decades, leading to the accurate, real-time recognition systems available on mobile phones today. In this thesis, I apply speech modeling techniques developed for recognition to two other speech problems: speech synthesis

    mit Repository record for Speech Representation Models for Speech Synthesis and Multimodal Speech Recognition (opens in a new tab)

  11. Automatic generation of fundamental frequency for text-to-speech synthesis

    Thesis (M. Eng.)--Massachusetts Institute of Technology, Dept. of Electrical Engineering and Computer Science, 1997.

    mit Repository record for Automatic generation of fundamental frequency for text-to-speech synthesis (opens in a new tab)

  12. A computational model of prosody for Yorùbá text-to-speech synthesis

    … (SY) language in the context of computer text-to-speech synthesis applications. The thesis of this research is that it is possible to develop a practical prosody model by using appropriate computational tools and techniques which combines acoustic data with an encoding of the phonological and …

    aston Repository record for A computational model of prosody for Yorùbá text-to-speech synthesis (opens in a new tab)

  13. A general platform and markup language for text to speech synthesis

    Thesis (M. Eng.)--Massachusetts Institute of Technology, Dept. of Electrical Engineering and Computer Science, 1996.

    mit Repository record for A general platform and markup language for text to speech synthesis (opens in a new tab)

  14. Finding Sparse Subnetworks in Self-Supervised Speech Recognition and Speech Synthesis

    The modern paradigm in speech processing has demonstrated the importance of scale and compute for end-to-end speech recognition and synthesis. For instance, state-of-the-art self-supervised speech representation learning models typically consists of more than 300M model parameters and being trained …

    mit Repository record for Finding Sparse Subnetworks in Self-Supervised Speech Recognition and Speech Synthesis (opens in a new tab)

  15. Spectral discontinuity in concatenative speech synthesis – perception, join costs and feature transformations

    … of spectral discontinuity in concatenative speech synthesis. Such measures are used as join costs to quantify the compatibility of speech units for concatenation in unit selection synthesis. No previous study has reported a spectral measure that satisfactorily correlates with human …

    dcu Repository record for Spectral discontinuity in concatenative speech synthesis – perception, join costs and feature transformations (opens in a new tab)

  16. Diphone Speech Synthesis Based on a Pitch-Adaptive Short-Time Fourier Transform

    … of this work is to investigate a new method of speech synthesis from phonetic specifications. The investigation includes the design, computer simulation, and subjective evaluation of a speech analysis-synthesis system. The method is new in the sense that it utilizes two novel analytical …

    uiuc Repository record for Diphone Speech Synthesis Based on a Pitch-Adaptive Short-Time Fourier Transform (opens in a new tab)

  17. The design and construction of a special purpose computer for speech synthesis-by-rule.

    Thesis. 1976. M.S.--Massachusetts Institute of Technology. Dept. of Electrical Engineering and Computer Science.

    mit Repository record for The design and construction of a special purpose computer for speech synthesis-by-rule. (opens in a new tab)

  18. Concatenative speech synthesis: a Framework for Reducing Perceived Distortion when using the TD-PSOLA Algorithm

    … and evaluation of an approach to concatenative speech synthesis using the Titne-Domain Pitch-Synchronous OverLap-Add (I'D-PSOLA) signal processing algorithm. Concatenative synthesis systems make use of pre-recorded speech segments stored in a speech corpus. At synthesis time, the `best' segments …

    bournemouth Repository record for Concatenative speech synthesis: a Framework for Reducing Perceived Distortion when using the TD-PSOLA Algorithm (opens in a new tab)

  19. The effects of part–of–speech tagging on text–to–speech synthesis for resource–scarce languages

    … of the project is more natural text-to-speech (TTS) voices. Naturalness is primarily determined by prosody and it is shown that many aspects of prosodic modelling is, in turn, dependent on part-of-speech (POS) information. Solving the POS problem is, therefore, a prudent first step …

    nwu-za Repository record for The effects of part–of–speech tagging on text–to–speech synthesis for resource–scarce languages (opens in a new tab)

Page 1 of 3