Global ETD Search
Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.
Results
Showing 1 to 13 of 13 for “"Synthesized speech"”.
-
Generating expression in synthesized speech
Thesis (M.S.V.S.)--Massachusetts Institute of Technology, Dept. of Architecture, 1989.
-
The effects of speech rate, message repetition, and information placement on synthesized speech intelligibility
Recent improvements in speech technology have made synthetic speech a viable I/O alternative. However, little research has focused on optimizing the various speech parameters which influence system performance. This study examined the effects of speech rate, message repetition, and the placement of …
-
Synthesized speech intelligibility and preschool age children: comparing accuracy for single word repetition with repeated exposure.
… environments on the intelligibility of two synthesized speech voices and human recorded speech in preschool age children. Methods: Eighteen preschool aged participants listened to and repeated single words presented in human recorded speech, DECtalk™ Paul and AT & T Voice ™ Michael during …
-
Assessing human performance trade-offs of a telephone-based information system
… levels each. The four factors manipulated were: synthesized speech rate, time available for user input, subject age, and background music level. Subjects searched a fictitious department store database for 16 specific store items and transcribed 16 information messages which were spoken by a …
-
Text-Free Audio Captions of Short Videos from Latent Space Representation
… we re-implement previous work exploring image to speech captioning. We expand upon the work to implement video to speech captioning. Specifically, we implement a text-free image to speech captioning pipeline that integrates four distinct machine learning models. We alter the models to process …
-
A computational memory and processing model for prosody
… links processing in working memory to prosody in speech, and links different working memory capacities to different prosodic styles. It provides a causal account of prosodic differences and an architecture for reproducing them in synthesized speech. The implemented system mediates text-based …
-
Investigating Pilot Performance Using Mixed-Modality Simulated Data Link
… to compare the intelligibility of two text-to-speech (TTS) engines (DECtalk and AT&T's Natural Voices) as presented in 85 dB(A) aircraft cockpit engine noise. Results indicated significant differences in intelligibility (p £ 0.05) between the two speech synthesizers across the tested …
-
Towards a multimodal Ouija Board for aircraft carrier deck operations
… our model of the deck by issuing commands using speech and gestures. The system responds to users with its own synthesized speech and graphics. The result is a conversation of sorts between users and the Ouija Board as they work together to accomplish tasks.
-
Investigating Speaker Features From Very Short Speech Records
… containing single words and shorter segments of speech. By taking advantage of the fast convergence properties of adaptive filtering, the approach is capable of modeling the nonstationarities due to both the vocal tract and vocal cord dynamics. Specifically, the procedure extracts the vocal tract …
-
Quality of service (QoS) analysis frameworkn for text to speech (TTS) services
… is significant and necessary for text to speech web service applications. Text to speech media conversion quality measurements has general and specific mechanisms for its functional and nonfunctional requirements. The main objective of this thesis is to introduce QoS framework which is …
-
The use of the auditory lexical decision task as a method for assessing the relative quality of synthetic speech
… method for determining the quality of synthetic speech systems. The method involves the use of an auditory lexical decision task to assess the quality of synthetic speech generators relative to each other and to natural speech by using reaction time differences and error rates. Seven voices were …
-
Intelligibility of synthesized voice messages in commercial truck cab noise for normal-hearing and hearing-impaired listeners
… was conducted to assess the intelligibility of synthesized speech under a variety of noise conditions for both hearing-impaired and normal-hearing subjects. Modified Rhyme Test stimuli were used to determine intelligibility in four speech-to-noise (S/N) ratios (0, 5, 10, and 15 dB), and three …
-
The effects of five discrete variables on human performance in a telephone information system
… information system. The five variables were: speech rate (120 or 240 words per minute), length of input time-out (two or ten seconds), feedback (available or not available), wallet guide - a graphical representation of the information (available or not available), and the database structure …