Abstract
dc:description.abstractWe optimize the synthesis procedure of a videorealistic speech animation system [7] to achieve real-time speech animation synthesis. A synthesis rate must be high enough for real-time video streaming for speech animation systems to be viable in industry and deployed as applications for user machine interactions and real-time dialogue systems. In this thesis we apply various approaches to develop a parallel system that is capable of synthesizing real-time videos and is adaptable to various distributed computing architectures. After optimizing the synthesis algorithm, we integrate a videorealistic speech animation system, called Mary101, with a speech synthesizer, a speech recognizer, and a RTMP server to implement a web-based text to speech animation system.
Degree
thesis:*- Department dc:contributor.department
- Massachusetts Institute of Technology. Dept. of Electrical Engineering and Computer Science.
- Grantor dc:publisher
- Massachusetts Institute of Technology
- Year dc:date.issued
- 2011
Author and committee
dc:creator, dc:contributor.*- Author dc:creator
-
- Fu, Jieyun
- Advisor dc:contributor.advisor
-
- James Glass and D. Scott Cyphers.
Subjects
dc:subject × 1Rights
dc:rights- Statement dc:rights
-
- M.I.T. theses are protected by copyright. They may be viewed from this source for any purpose, but reproduction or distribution in any format is prohibited without written permission. See provided URL for inquiries about permission.
- Licence dc:rights.uri
- Language dc:language.iso
- eng
Identifiers
dc:identifier.*- Handle dc:identifier.uri
- http://hdl.handle.net/1721.1/66416
- OAI identifier oai:identifier
- oai:dspace.mit.edu:1721.1/66416