Global ETD Search
Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.
Results
Showing 1 to 4 of 4 for “"Audio-visual automatic speech recognition"”.
-
Designing a Visual Front End in Audio-Visual Automatic Speech Recognition System
<p>Audio-visual automatic speech recognition (AVASR) is a speech recognition technique integrating audio and video signals as input. Traditional audio-only speech recognition system only uses acoustic information from an audio source. However the recognition performance degrades significantly in …
-
Lip Detection and Adaptive Tracking
<p>Performance of automatic speech recognition (ASR) systems utilizing only acoustic information degrades significantly in noisy environments such as a car cabins. Incorporating audio and visual information together can improve performance in these situations. This work proposes a lip detection and …
-
IR-Depth Face Detection and Lip Localization Using Kinect V2
<p>Face recognition and lip localization are two main building blocks in the development of audio visual automatic speech recognition systems (AV-ASR). In many earlier works, face recognition and lip localization were conducted in uniform lighting conditions with simple backgrounds. However, such …
-
From bits to information : learning meets compressive sensing
… such as computer vision classification tasks, Audio-Visual Automatic Speech Recognition, lossy image compression and retrieval via locality sensitive hashing, locally linear estimation in large scale learning and Fourier sampling for phase retrieval - of particular interest in X-ray …