Global ETD Search
Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.
Results
Showing 1 to 3 of 3 for “"speech foundation models"”.
-
Speech Foundation Models for Audio Processing
Foundation models are trained on massive amounts of data and can be adapted to support a wide range of downstream tasks. Their emergence has significantly shaped the development of natural language and speech technologies. Foundation speech models such as Whisper have shown strong performance …
-
Efficient Knowledge Transfer and Adaptation for Speech and Beyond
… transfer and adaptation in the realm of speech processing. It is structured to address the limitations of transfer learning in dynamically evolving audio and speech processing contexts, particularly through novel approaches for class-incremental learning, parameter-efficient adaptation, …
-
Neural Time Alignment for End-to-End Automatic Speech Recognition Systems
Automatic Speech Recognition (ASR) is an important component for machines to interact with humans. With the development of deep learning, ASR systems can obtain very low error rates when there is a lot of training data available. Apart from the recognised text, it is also important to determine …