Global ETD Search
Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.
Results
Showing 1 to 10 of 10 for “"Speech translation"”.
-
Direct Speech Translation Toward High-Quality, Inclusive, and Augmented Systems
When this PhD started, the translation of speech into text in a different language was mainly tackled with a cascade of automatic speech recognition (ASR) and machine translation (MT) models, as the emerging direct speech translation (ST) models were not yet competitive. To close this gap, part of …
-
Korean language generation in an interlingua-based speech translation system
Thesis (M. Eng.)--Massachusetts Institute of Technology, Dept. of Electrical Engineering and Computer Science, 1995.
-
The Use of Language Models in End-to-End Spoken Language Processing
… spoken language processing including automatic speech recognition (ASR) and speech translation (ST). E2E spoken language processing systems have many advantages compared to cascaded systems, including global optimisation, a simplified pipeline, and a compact structure. However, E2E training …
-
Language Modeling from Visually Grounded Speech
… processing have significantly reduced automatic speech recognition (ASR) error rates, driven by large-scale supervised training on paired speech–text data and, more recently, self-supervised pre-training on unpaired speech and audio. These methods have facilitated robust transfer learning across …
-
Speech Foundation Models for Audio Processing
… shaped the development of natural language and speech technologies. Foundation speech models such as Whisper have shown strong performance across a variety of audio processing tasks, including automatic speech recognition (ASR) and speech translation. Unlike traditional systems that require …
-
Speech processing with less supervision : learning from weak labels and multiple modalities
… learning has achieved great success in speech processing with powerful neural network models and vast quantities of in-domain labeled data. However, collecting a labeled dataset covering all domains can be either expensive due to the diversity of speech or almost impossible for some …
-
Automatic subtitling: A new paradigm
Audiovisual Translation (AVT) is a field where Machine Translation (MT) has long found limited success mainly due to the multimodal nature of the source and the formal requirements of the target text. Subtitling is the predominant AVT type, quickly and easily providing access to the vast amounts of …
-
On the Security of Speech-based Machine Translation Systems: Vulnerabilities and Attacks
… reliance onmultilingual communication, speech-based Machine Translation (MT) systems have emerged as essential technologies for facilitating seamless cross-lingual interaction. These systems enable individuals and organizations to overcome linguistic boundaries by automatically …
-
Transfer Learning For Spoken Language Processing
… domain adaptation in the context of Automatic Speech Recognition (ASR) and Cross-Lingual Learning in Automatic Speech Translation (AST). The first part of the thesis develops an algorithm for unsupervised domain adaptation of End-to-End ASR models. In recent years, ASR performance has improved …
-
Speech Adaptation Modeling for Statistical Machine Translation
Spoken language translation (SLT) exists within one of the most challenging intersections of speech and natural language processing. While machine translation (MT) has demonstrated its effectiveness on the translation of textual data, the translation of spoken language remains a challenge, largely …